> iv. Your use of the SAM Materials will not involve or encourage others to reverse engineer, decompile or discover the underlying components of the SAM Materials.
> v. You are not the target of Trade Controls and your use of SAM Materials must comply with Trade Controls. You agree not to use, or permit others to use, SAM Materials for any activities subject to the International Traffic in Arms Regulations (ITAR) or end uses prohibited by Trade Controls, including those related to military or warfare purposes, nuclear industries or applications, espionage, or the development or use of guns or illegal weapons.
> b. If you institute litigation or other proceedings against Meta or any entity (including a cross-claim or counterclaim in a lawsuit) alleging that the SAM Materials, outputs or results, or any portion of any of the foregoing, constitutes infringement of intellectual property or other rights owned or licensable by you, then any licenses granted to you under this Agreement shall terminate as of the date such litigation or claim is filed or instituted. You will indemnify and hold harmless Meta from and against any claim by any third party arising out of or related to your use or distribution of the SAM Materials.
It's always nice when a model's weights are released, but Meta's models are not open source because their weights always come with weird restrictions.
momojo 1 hours ago [-]
I'm no fan of Facebook or its social effects but I can't deny the wonderful downstream effect of their open source.
Popular microscopy models like Cellpose[0] have leaned heavily on the cornucopia of open and SOTA power. I have no doubt thousands of biologists have benefitted from the capabilities these models bring. I think it was unthinkable just 5 years ago that a single biologist with just a laptop could do mass-segmentation at this kind of fidelity.
Then there's Napari and it's plugin ecosystem[1] that wouldn't exist without the Chan Zuckerberg Initiative. Again, I'm not trying to glaze them but as someone in the biotech/microscopy space I can't understate how often I use and benefit from their open source.
> but I can't deny the wonderful downstream effect of their open source.
ReactJS has been a very mixed-bag...
37 minutes ago [-]
righthand 30 minutes ago [-]
React is more of a plague than a wonderful downstream effect or mixed bag. It’s essentially cemented Js as the way to build a website even if you don’t need the complexity. It is the leader in brain dead Js evangelism.
DaiPlusPlus 21 minutes ago [-]
Credit where credit is due, though: ReactJS became #1 on its merits.
...yes, Redux (not React) deserves to be #1, but Redux isn't a complete, all-under-one-roof framework the way React is; but regardless of that: The Redux/React approach is just fundamentally a better design than what we had before: stateful-controls/widgets and two-way data-binding.
If you'd like to relive how UI devs suffered throughout the 1990s, 2000s, and most of the 2010s I invite you to try making a native Windows 11 desktop UI using WinUI3 using the MVVM (anti-) pattern: nothing but mutable objects of indeterminable state getting caught in infinite-loops or unbound recursion due to INotifyPropertyChanged - and Microsoft is still pretending that's the "right" way to build a UI.
Sorry am ranting on about something I have very little control over; it's just frustating.
32 minutes ago [-]
nine_k 2 hours ago [-]
The scoop: X-ray imaging of various strictures for scientific purposes produces colossal reams of data, previously hard to analyze. Meta provides machine analysis, both segmentation and classification, using unsupervised learning models.
> a fully reconstructed, semantically labeled 3D volume delivered back to the scientist physically standing at the beamline [x-ray] instrument, ready for interpretation while the experiment is still running. Total turnaround: approximately 15 minutes.
sailingparrot 58 minutes ago [-]
> Meta's open-source approach makes this possible
I'm sorry, was this article drafted in 2024 and never updated?
kingstnap 18 minutes ago [-]
This sort of random out of date comment is a common LLM writing trope. Though it is true that the segment anything model is open source.
> A100 GPUs — the high-performance computing chips that power today's most advanced AI systems.
They might not have written it with AI, but the article has a lot of em dashs and colons and not this but that statements.
hgoel 29 minutes ago [-]
The models in question have source and weights available. AFAIK not fully FOSS because the weights require registration to access.
Zaheer 2 hours ago [-]
This fits with my impression of the 'personality' of various models:
Meta: Perceptive (strong vision)
Gemini: Fastest
Claude: Smartest
OpenAI: Prettiest
hgoel 3 minutes ago [-]
They aren't talking about vision LLMs though. SAM and DINO are vision models, no LLM involved.
noodlescb 49 minutes ago [-]
I love how consistently none of us even vaguely consider Grok an actual player
pram 33 minutes ago [-]
To Grok's credit I think it's fairly good as a creative writing tool because it can be very "spontaneous" and it naturally seems to use an informal style. It also lacks a lot of the words and phrasing Claude and OpenAI get hyper-fixated on.
IDK if this is emergent from being trained on an endless trough of Twitter shitposts but compared to how stiff the rest are, I consider it a feature. I wouldn't use it for anything important though, heh.
airstrafer 44 minutes ago [-]
The models might be good but the product design, user story, and marketing is so terrible that it’s difficult to see it as more than an also-ran
noodlescb 38 minutes ago [-]
Even that is kind tbh. The company is so poorly run and the leader so controversial that it makes it irresponsible to build anything serious that relies on their products. Other than the rocket part of the business, everything else under the SpaceX umbrella is a nonstarter.
giancarlostoro 18 minutes ago [-]
Starlink arguably is not the rocket part and its the most profitable piece, it held together the rocket side of SpaceX and he expanded research and development.
kubrickslair 1 hours ago [-]
Why do you think OpenAI is the prettiest?
Claude often makes better looking interfaces and designs. And I think OpenAI has solved more open math/ stats/ CS problems.
giancarlostoro 16 minutes ago [-]
> And I think OpenAI has solved more open math/ stats/ CS problems.
Still surprises me that OpenAI seems to lead in this one weird niche, I wonder what causes GPT to be able to routinely pull this off, there was one instance where some random 18 year old broke some mathematical question without knowing more than high school math if I remember correctly, all because of GPT.
embedding-shape 59 minutes ago [-]
> Claude often makes better looking interfaces and designs
How do you even qualify this? Either by "Well, when you're not specifying anything about it in the prompt" and then it almost doesn't matter at all, or by what actually goes into the prompt, then again it doesn't matter at all what model you use, more about the person driving it.
economistbob 2 hours ago [-]
Qwen: Zestiest
Nemo: Straightest
DeepSeek: Craftiest
munk-a 2 hours ago [-]
And, of course, Grok: Sir-not-appearing-in-this-listest
radiorental 1 hours ago [-]
Drunkleist?
ben_w 1 hours ago [-]
Muskiest
2 hours ago [-]
jyr0s 2 hours ago [-]
this page hijacks your tab's back button history :\
abirch 2 hours ago [-]
Well Meta has hijacked my privacy. Now IRL Meta can track me from the people wearing their sunglasses.
brcmthrowaway 1 hours ago [-]
ELI5.
embedding-shape 60 minutes ago [-]
SAM 3 (Segment Anything Model 3) and DINOv3, projects released by Facebook, were used to do science and research.
Petersipoi 57 minutes ago [-]
AI model inspects hundreds of thousands of scientific images. A job that previously took an expert roughly a month can now be completed in around 15 minutes.
nonameiguess 2 hours ago [-]
Gonna need to have a talk with LLNL. I'm sure they didn't choose the name, but seems a tad leaning in to use the name of tech from Star Trek meant to produce untold abundance that instead became an unintentional doomsday device.
iLoveOncall 2 hours ago [-]
So, not LLM models, right?
Also on this:
> The numbers are staggering: The DOE's light and neutron source facilities now produce tens of petabytes of data annually
Come on, petabytes are not staggering for entreprise software.
munk-a 2 hours ago [-]
I used to work with images representing scans of brain tissue - for a full brain visualization at one horizontal slice terabytes was a common measure and the resolution of those images wasn't even particularly detailed - all the full resolution stuff was taken of tiny sub-sections of interest. This was also two decades ago - so I'm sure they've upped their game.
momojo 1 hours ago [-]
My company does whole-brain scans of mice on a Zeiss Z.1. Lower resolutions are typically in the low hundreds of GB. Higher-resolutions and multichannel staining can get you in the TB range. When a typical client is doing 10's of brains it definitely adds up. But even for us (and I consider us a smaller operation) we aren't output PB's. So I'd consider the above claim still pretty impressive.
munk-a 40 minutes ago [-]
Yeah - our multi TB scan from two decades ago was a dolphin brain image which is a fair bit larger than mouse - and it was a stained sample that ended up being used as our demo image frequently because the contrast dyes set well and resulted in a very pretty visual overall.
I didn't meant to say that PBs of image data is common place - we had no image that approached that size - but people outside the domain of microscope scan results might be unfamiliar with just how chonky these image files would get traditionally.
hgoel 22 minutes ago [-]
I think a big thing that the enterprise comparison misses is that this is petabytes worth of dense data that has to be put through fairly heavy processing (some of which is currently custom for the specific experiment) and studied by a human.
It isn't just a giant database of small files and metadata blindly feeding a recommender system.
Several petabytes is definitely a staggering amount of data in that context of being analyzed by human eyes to extract some scientific value.
ben_w 58 minutes ago [-]
I'm not sure exactly which enterprises you have in mind, but sure: quantities which can be expressed as "a year's worth fits on my desk" should not be described as "staggering", and 10 PB of hard drives will (just about) fit on my desk.
The LHC, on the other hand, that generates a petabyte a second and has to throw most of it away for obvious reasons:
As a former proposal specialist (B2B, B2G, non DoD) I looked into the Genesis Mission procurement site and process.
Unless someone can correct me, the total amount of grant monies is $280,000,000 or so.
It became obvious that it’s not worth my time to engage in the “mission” as they call it, even if I could benefit some worthwhile causes.
That’s a pittance and pretty insulting to the purported benefit of funding scientific endeavors. I’m not even attempting to be political here. $280 Million versus $XX Billion for warfighting is a seriously gross misallocation of public monies, IMHO.
Total lackluster reporting on the scale and scope of the actual numbers, but not surprising.
orochimaaru 24 minutes ago [-]
This is a really weird comparison. The Chan Zuckerberg foundation is nowhere involved with any war. All they’re doing is making money available for research. In which universe is $280m not enough?
philipwhiuk 1 hours ago [-]
It's a weird world where $280M is not considered a lot of money.
mawadev 2 hours ago [-]
That is — incredible
rozap 1 hours ago [-]
Really fascinating writeup — I particularly appreciated how they didn't hesitate to delve into the load bearing design choices — the implications are staggering.
Rendered at 19:35:31 GMT+0000 (Coordinated Universal Time) with Vercel.
SAM License (https://github.com/facebookresearch/sam3/blob/main/LICENSE):
> iv. Your use of the SAM Materials will not involve or encourage others to reverse engineer, decompile or discover the underlying components of the SAM Materials.
> v. You are not the target of Trade Controls and your use of SAM Materials must comply with Trade Controls. You agree not to use, or permit others to use, SAM Materials for any activities subject to the International Traffic in Arms Regulations (ITAR) or end uses prohibited by Trade Controls, including those related to military or warfare purposes, nuclear industries or applications, espionage, or the development or use of guns or illegal weapons.
> b. If you institute litigation or other proceedings against Meta or any entity (including a cross-claim or counterclaim in a lawsuit) alleging that the SAM Materials, outputs or results, or any portion of any of the foregoing, constitutes infringement of intellectual property or other rights owned or licensable by you, then any licenses granted to you under this Agreement shall terminate as of the date such litigation or claim is filed or instituted. You will indemnify and hold harmless Meta from and against any claim by any third party arising out of or related to your use or distribution of the SAM Materials.
The DINOv3 License (https://github.com/facebookresearch/dinov3/blob/main/LICENSE...) is similar but with the model names swapped.
It's always nice when a model's weights are released, but Meta's models are not open source because their weights always come with weird restrictions.
Popular microscopy models like Cellpose[0] have leaned heavily on the cornucopia of open and SOTA power. I have no doubt thousands of biologists have benefitted from the capabilities these models bring. I think it was unthinkable just 5 years ago that a single biologist with just a laptop could do mass-segmentation at this kind of fidelity.
Then there's Napari and it's plugin ecosystem[1] that wouldn't exist without the Chan Zuckerberg Initiative. Again, I'm not trying to glaze them but as someone in the biotech/microscopy space I can't understate how often I use and benefit from their open source.
[0] https://cellpose.readthedocs.io/en/latest/models.html [1] https://chanzuckerberg.com/rfa/napari-plugin-grants/
ReactJS has been a very mixed-bag...
...yes, Redux (not React) deserves to be #1, but Redux isn't a complete, all-under-one-roof framework the way React is; but regardless of that: The Redux/React approach is just fundamentally a better design than what we had before: stateful-controls/widgets and two-way data-binding.
If you'd like to relive how UI devs suffered throughout the 1990s, 2000s, and most of the 2010s I invite you to try making a native Windows 11 desktop UI using WinUI3 using the MVVM (anti-) pattern: nothing but mutable objects of indeterminable state getting caught in infinite-loops or unbound recursion due to INotifyPropertyChanged - and Microsoft is still pretending that's the "right" way to build a UI.
Sorry am ranting on about something I have very little control over; it's just frustating.
> a fully reconstructed, semantically labeled 3D volume delivered back to the scientist physically standing at the beamline [x-ray] instrument, ready for interpretation while the experiment is still running. Total turnaround: approximately 15 minutes.
I'm sorry, was this article drafted in 2024 and never updated?
https://github.com/facebookresearch/sam3
They have another with calling A100s modern.
> A100 GPUs — the high-performance computing chips that power today's most advanced AI systems.
They might not have written it with AI, but the article has a lot of em dashs and colons and not this but that statements.
Meta: Perceptive (strong vision)
Gemini: Fastest
Claude: Smartest
OpenAI: Prettiest
IDK if this is emergent from being trained on an endless trough of Twitter shitposts but compared to how stiff the rest are, I consider it a feature. I wouldn't use it for anything important though, heh.
Claude often makes better looking interfaces and designs. And I think OpenAI has solved more open math/ stats/ CS problems.
Still surprises me that OpenAI seems to lead in this one weird niche, I wonder what causes GPT to be able to routinely pull this off, there was one instance where some random 18 year old broke some mathematical question without knowing more than high school math if I remember correctly, all because of GPT.
How do you even qualify this? Either by "Well, when you're not specifying anything about it in the prompt" and then it almost doesn't matter at all, or by what actually goes into the prompt, then again it doesn't matter at all what model you use, more about the person driving it.
Nemo: Straightest
DeepSeek: Craftiest
Also on this:
> The numbers are staggering: The DOE's light and neutron source facilities now produce tens of petabytes of data annually
Come on, petabytes are not staggering for entreprise software.
I didn't meant to say that PBs of image data is common place - we had no image that approached that size - but people outside the domain of microscope scan results might be unfamiliar with just how chonky these image files would get traditionally.
It isn't just a giant database of small files and metadata blindly feeding a recommender system.
Several petabytes is definitely a staggering amount of data in that context of being analyzed by human eyes to extract some scientific value.
The LHC, on the other hand, that generates a petabyte a second and has to throw most of it away for obvious reasons:
https://www.itnews.com.au/news/computing-for-the-large-hadro...
Unless someone can correct me, the total amount of grant monies is $280,000,000 or so.
It became obvious that it’s not worth my time to engage in the “mission” as they call it, even if I could benefit some worthwhile causes.
That’s a pittance and pretty insulting to the purported benefit of funding scientific endeavors. I’m not even attempting to be political here. $280 Million versus $XX Billion for warfighting is a seriously gross misallocation of public monies, IMHO.
Total lackluster reporting on the scale and scope of the actual numbers, but not surprising.