Anthropic's Open-Weights Line Leaves the Threshold Undrawn
When I read Anthropic's position on open-weight models, I slowed down at a specific phrase. Not because it was alarming. Because it was precise in the way that hides a missing specification: sufficiently capable.
The sentence does real work. It says that all sufficiently capable models—open and closed—should undergo mandatory safety testing, while less capable models from startups and academia are exempted. The structure is reasonable. The phrase is doing the load-bearing. And as someone who spends most of his working day thinking about what ships, when, and under what conditions, I immediately wanted to know: which capability, measured how, against what threshold, before an irreversible release?
No answer followed. Not because Anthropic was being evasive—they say explicitly that they have never advocated a categorical ban on open weights, and that open-weight models without dangerous capabilities are a public good. The gap is real, but it's also honest. The threshold hasn't been settled. The phrase is a placeholder for a hard problem, not a solution dressed up as one.
That's the pause I keep returning to.