Rendered at 00:13:50 GMT+0000 (Coordinated Universal Time) with Cloudflare Workers.
flowerlad 46 minutes ago [-]
It seems every DeepSeek paper/patent has a huge number of authors, and this one is no exception. They couldn't even fit everyone on the page, there are 31 others not shown. This could be an asset protection strategy (i.e., human assets). Imagine if there were only 3 authors. Those authors may get hired away by competitors. If you list every employee on every paper then competitors don't know who to lure away.
ahurmazda 8 minutes ago [-]
Just look for the “corresponding author”
ycui7 3 minutes ago [-]
click the link, or read the PDF.
all authors are listed. there is no conspiracy to hide authorship.
arxiv simply want to keep the page short not too long.
yipinwong 1 hours ago [-]
The topic isn't as interesting as how 131 authors communicated to get this out.
14 minutes ago [-]
stefan_ 15 minutes ago [-]
I don't understand why half the comments here are about the author list, this is very common practice in e.g. large-scale physics experiments and biology, and every new GPT release from OpenAI equally had papers with tons of authors
hodgehog11 13 minutes ago [-]
Agreed, I'm very confused as well. It's like no-one here has been paying attention to research papers.
Or they are just rushing to say anything, and it's much easier to comment on that than the content of the paper.
embedding-shape 1 hours ago [-]
Is it possible they're doing a research lab "socialism style" and everyone gets equal credit for just being a part of the lab, regardless of actual input into the specific papers? If they're innovating in computer, maybe they're not so afraid of innovating in social/academic structures as well?
tonmoy 1 hours ago [-]
This is common practice in Biology labs in the west I think
vblanco 4 hours ago [-]
380.000 concurrent sandboxes on 160 Epyc based server nodes. Crazy stuff
dangoodmanUT 2 hours ago [-]
that's only 12 sandboxes per core
aabhay 2 hours ago [-]
Um, how is that not impressive
r_lee 42 minutes ago [-]
if it's agentic stuff, they likely aren't hammering a core constantly and they will maybe sit idle quite often between model requests, so it makes sense. I do wonder how much memory they allocate to each one though.
it's just very efficient use of shared cores that is required to make these kinds of workloads cost efficient
doc_ick 3 hours ago [-]
Will admit that I haven’t read it yet, just saw the crazy number of authors and think this may compete for one of the papers with the most authors.
justinnk 3 hours ago [-]
This one [0] about the Higgs boson is hard to beat. Pages 25-32 are just full of around 5150 „authors“, pages 33-37 are their affilliations.
Even without that particularly special case, high-energy physics collaborations have long since broken the idea of authors. There are now many collaborations with hundreds of authors publishing regularly and quite a few that are into the thousands.
mentalgear 1 hours ago [-]
I like it, represents science's general 'standing on the shoulders of our precedents' far more realistic then the 'genius solo' PR mythos.
elashri 49 minutes ago [-]
I think the number is ~2930 authors according to the CDS [1] which is comparable to the corresponding CMS paper which has about ~2900 authors [2]
how do you even keep up with the amount of research coming out these days
throwaway7783 48 minutes ago [-]
Is this like agent substrate?
redat00 2 hours ago [-]
So.. serverless ?
redat00 2 hours ago [-]
Still very impressive! Love how it's done!
tipiirai 2 hours ago [-]
Can you give a brief for what this is and why it is impressive?
redat00 1 hours ago [-]
It just describes the platform they built for scheduling workloads, and running those workloads. After a second thought it is not that impressive and probably doesn't deserve any kind of hype. It's the same kind of setup AWS is running for Lambda, as well as anyone else basically running SLURM clusters out there.
Still giving them credit because creating such as scheduler/platform from scratch is quite complex, and I know that they probably struggled a lot to get it right.
r_lee 37 minutes ago [-]
I'd say it's cool that they're openly writing about it and how they run their workloads. shows a nice window into how these things are actually deployed at scale.
it also shows how much density you can get easily from a single core if you wanted to replicate this
Vaslo 2 hours ago [-]
That number of authors though
swingboy 4 hours ago [-]
Is there a lab more innovative than DeepSeek? Imagine if they had the same compute resources that Anthropic and OpenAI have.
ProphetOfParado 4 hours ago [-]
Food for thought: Constraints are the source of creativity.
conception 4 hours ago [-]
Yeah if they had the resources of an OpenAI or anthropic they’d be OpenAI or Anthropic. Scrappy underdogs have to be nimble and innovative.
mirekrusin 3 hours ago [-]
They are as “underdog” as Linux is to Windows.
ianm218 1 hours ago [-]
Not really, they are underdogs in the true sense of the word.
rozim 3 hours ago [-]
Possibly relevant: 突破技术壁垒, "break the technical barricade" -> overcome an obstacle through innovation (in this case, sub SOTA GPUs at least).
Native Chinese speakers to confirm....
michaellee8 3 hours ago [-]
your translation is correct, I would say Chinese labs may be able to figure out the current capability of latest frontier models in 3-6 months, but then Anthropic and OpenAI may have already been ASI in that time already. China's main problem is still lack of (good) chips, and that is a hardware issue that is unlikely to be solved for a while. and more effort for efficiency means less effort for actual capability improvements. we have already seen what anthropic can do if they focus on efficiency with opus 5.5
impulser_ 4 hours ago [-]
Short term they might have less compute, but long term they will most definitely have more compute. They don't have to worry about energy, they don't have to worry about people blocking them building data centers the only thing stopping them is there no Chinese manufacturer that can produce a chip as good as Nvidia but I would bet that solved in a year or so.
tucnak 49 minutes ago [-]
They don't have to worry about people blocking DC construction in the US either. All new AI datacenters are designated "dual-use" so the federal government is already letting local councils know to fuck off.
SmartestUnknown 3 hours ago [-]
Just because other companies don't write papers about what they do doesn't mean they aren't innovative...
broodbucket 2 hours ago [-]
I'm actually the most innovative, I've written thousands of papers advancing the state of human knowledge. They're just in my basement and I don't show anyone.
doc_ick 3 hours ago [-]
Sure, but we’ll never know what they do or if it is innovate because we won’t know what they do.
ijidak 3 hours ago [-]
"Necessity is the mother of invention."
Not sure they'd be the same without the constraints.
all authors are listed. there is no conspiracy to hide authorship.
arxiv simply want to keep the page short not too long.
Or they are just rushing to say anything, and it's much easier to comment on that than the content of the paper.
it's just very efficient use of shared cores that is required to make these kinds of workloads cost efficient
[0] https://arxiv.org/abs/1207.7214
[1] https://cds.cern.ch/record/1471031
[2] https://impact.ornl.gov/en/publications/observation-of-a-new...
Still giving them credit because creating such as scheduler/platform from scratch is quite complex, and I know that they probably struggled a lot to get it right.
it also shows how much density you can get easily from a single core if you wanted to replicate this
Native Chinese speakers to confirm....
Not sure they'd be the same without the constraints.