Posts / ai
Open Source AI Won by Being Open (And Why That Should Worry Nvidia More Than Us)
I fell down a r/LocalLLaMA hole last night instead of doing anything sensible, and came out the other side with a thought I can’t quite shake: the open source AI crowd isn’t winning because they’re smarter than OpenAI or Anthropic. They’re winning because they’re not allowed to keep secrets from each other.
There’s a thread doing the rounds with a fairly simple thesis: closed labs have to constantly reinvent the wheel to protect their lead, while the open labs, mostly the big Chinese ones plus a scattering of others, are effectively working in public. Someone drops a new attention mechanism, someone else’s next release quietly has it baked in a fortnight later. One commenter put it well: it’s “almost like working together makes improving technology easier.” Which, yes. Obviously. But it’s a genuinely useful thing to see stated plainly, because for a couple of years the story we were told was that scale and secrecy were the moat. Enough compute, enough proprietary technique, and nobody catches up.
That doesn’t seem to be how it’s playing out. Reading through the comments, there’s a decent case that closed labs are stuck rotating the same small pool of star researchers, the way a football club buys the best striker every transfer window and hopes it’s enough. Someone made that exact comparison, actually, comparing it to buying Messi: technically you’ve got him, his style doesn’t change much, and meanwhile every other club is figuring out how to beat that style collectively. Open labs don’t have a Messi. They have a few hundred people who all read each other’s papers on a Tuesday and ship the good ideas by Friday. It’s not romantic, it’s just structurally faster.
I don’t fully buy the tidier version of this story though, the one where open source is simply, morally, the better path and closed labs are doomed. Someone in the thread pushed back hard on the idea that any open model actually rivals the frontier closed ones, and that’s a fair challenge. Benchmarks are noisy, vibes-based comparisons are worse, and “beats GPT in coding” claims age like milk most months. I hold both of these at once: open models have gotten shockingly good, fast, and also I don’t fully trust my own read on how good, because I’m reading Reddit threads written by people who are, like me, a bit invested in a particular outcome.
What struck me more, honestly, was a separate but related thread worrying about Nvidia potentially buying up Hugging Face and poaching the llama.cpp team. That one hit differently for me, because it’s less about who’s smarter and more about who owns the plumbing. I’ve spent enough years in IT watching infrastructure quietly consolidate into two or three companies’ hands to know how that story usually ends. Doesn’t have to be malicious. Just slow, boring enterprise gravity: acquire the thing, “streamline” it, and five years later the free version does 60% of what it used to and the interesting bits are paywalled. One commenter reckoned it’s a defensive buy, Nvidia hoovering up Hugging Face so nobody else gets to control the on-ramp to their hardware. Plausible. Also not exactly comforting.
None of this is abstract for me. I run a couple of models locally on an ageing GPU mostly out of curiosity and a mild allergy to sending my daughter’s homework questions to a server in Virginia. The gap between what I can run in my study in Berwick and what OpenAI charges a subscription for has closed a lot faster than I expected two years ago. That’s genuinely exciting, and I don’t say that about much in tech anymore. It’s also the exact reason I keep half an eye on who owns the tools that make it possible. The community building this stuff has no shareholders to answer to, which is its whole advantage, and also the reason it’s vulnerable to one large company deciding to buy the parts it likes.
I don’t have a neat conclusion here, and I don’t think one exists yet. Open models catching up to closed ones is a genuinely good thing for anyone who’d rather not have three companies deciding what a model is allowed to say about, say, Taiwan or wages or vaccines. But “open” only stays open as long as enough people keep hoarding weights, seeding torrents, and forking the code the moment ownership gets weird. Somebody in that Nvidia thread said the community should treat this as a hedge, not a guarantee. That feels like the right amount of worried to be. Not panicked. Just paying attention, which is more than I can say for most of what passes for tech commentary these days.