nonethewiser
9 hours ago
We already have a window into the future.
- Anthropic told everyone Mythos was dangerous because it's proficiency with biologics and cyber security
- Anthropic didn't release Mythos like everything else. They released a neutered fable. They didn't get rid of Mythos
- Anthropic opens a lab in SF
There was always a quesiton of "will the labs stop releasing their models and start building around them instead?" Yes - they already have. Anthropic is a biologic and cyber security company, in addition to intelligience.
Personally I wonder if they've been holding back a lot. Opus 5.5 was a good release after a little stagnation. Open AI releases good models and everyone says Anthropic sucks and -- Oh would you look at that - a better model finally and all of a sudden.
WarmWash
9 hours ago
The way it is right now is to build giga-monster models, use those internally to boost yourself, and distill them down into lighter consumer models.
Apparently OAI is already building GPT7 and GPT8
pvab3
9 hours ago
with how much competition, benchmaxxing, and increased compute for inference, I can't imagine that they are intentionally nerfing their own public models for any reason other than that they can't figure out how to package it into something the public can use. The training data for all these frontier models goes well into 2026 at this point.
bentt
8 hours ago
I can totally see them nerfing their public models.
Just beat the current winner by enough to own the spotlight for a bit, then start prepping for the next go round.
_rutinerad
3 hours ago
The Mondo Duplantis strategy.
jbs789
9 hours ago
Either that or they are throwing spaghetti at the wall to see what sticks ahead of the IPO. After all, if solving all diseases is the “total addressable market” then that sure helps.
Time will tell.
downrightmike
8 hours ago
But insurance companies do not want to cure things, not profitable. So that revenue is just not going to work, insurance will not cover it.
ejj28
8 hours ago
It doesn't have to be profitable, it just has to get investors believing it might be.
siliconc0w
7 hours ago
It's pretty frustrating to do cyber security work and not have access to the best models. OAI is a little more liberal here and I was able to get access to daybreak-blue but I have to use the lesser last-generation models. Essentially this gives a small number of orgs a huge advantage in the 'application layer' for that domain (including OAI or ANT themselves).
bpodgursky
9 hours ago
Yes, both labs already have monstrous internal teacher models they don't sell for inference, this is generally acknowledged. They cut releases for the public just to keep revenues growing, it's not their actual frontier.
sowhat1
9 hours ago
Ok. Given this hypothesis, why is the software they release generally considered crappy by competitor standards, benchmarks, and open source standards?
Claude Code is an awful codebase, has leaked its own source code multiple times, and scores the worst on number of tokens burned vs pass rate percentages.
Is anyone even using their Figma competitor?
monocasa
8 hours ago
Probably because bad code that you create initially without thinking that it's a core piece of your stack becomes depended on for its crappy behavior, and then you can't change much without breaking workflows.
I seem to recall Fred Brooks talking about that experience with OS/360 JCL (maybe just straight up in The Mythical Man Month?).
CoolestBeans
8 hours ago
If agents really are superpowerful at programming tasks why not just have it rewrite the tool that the majority of your customers use and have it recreate the bugs? I mean presumably its the primary force behind the current version so what's the major cost there?
XenophileJKO
7 hours ago
I imagine it comes down to economics.. there isn't much upside to fixing the last 20% of issues that the dumber faster models are missing.
The cost to serve, latency profile ,and internal demand for a maximal intelligence model would probably keep it pointed at harder and more valuable problems most of the time.
_aavaa_
6 hours ago
Like front-running Millennium prize solutions?
addaon
9 hours ago
Because even beyond-frontier LLMs are bad at software.
IanCal
9 hours ago
How many people does it stop from using the software?
owebmaster
9 hours ago
> Is anyone even using their Figma competitor?
Yes. CC started to create canvases without me asking. The mockups look good (it's just html+css), the tool is vibecoded crap
TeMPOraL
8 hours ago
I mean, on the one hand the tool may be vibecoded crap - I don't know, haven't checked, taking it at your word.
On the other hand, Opus 5.5 cracked zero-shotting proper LCARS interfaces that near-perfectly adhere to the franchise "design language" even in tiny details, while simultaneously being 100% functional following my admonitions about Airbus cockpit design rules and nuclear reactor control room standards.
So yeah, why wouldn't I use it? It works spectacularly well.
throwaway27448
8 hours ago
Their cyber security seems to be a much more materially interesting (and likely profitable) business than "intelligence". Unfortunately they've created a sort of mutually-assured-destruction racket where they take payments from both "sides" of any secured boundary.
I'm not holding my breath for the biology side of things, but I suppose it's possible they find interesting things.
yesbabyyes
7 hours ago
> I'm not holding my breath for the biology side of things
Well, that's the thing--perhaps we should be.
throwaway27448
4 hours ago
Maybe. I haven't heard anything impressive from the bio side of my social network, nor can I see anything from the computational side (including my own understanding). Thankfully it'll be pretty obvious if they find something interesting.