Peter Gostev / @petergostev:
Mistral Large 4 was trained on 4,000 GPUs and Astra was trained on 100,000 of them. We'll see what the kitten can pull off, but the disparity is pretty stark
Peter Gostev / @petergostev:
Mistral Large 4 was trained on 4,000 GPUs and Astra was trained on 100,000 of them. We'll see what the kitten can pull off, but the disparity is pretty stark
x.com/petergostev...
x.com/petergostev...
Peter Gostev / @petergostev:
For real? SemiAnalysis: "Anthropic Subscriptions Offer 5x+ More Value Than OpenAI"
Peter Gostev / @petergostev:
For real? SemiAnalysis: "Anthropic Subscriptions Offer 5x+ More Value Than OpenAI"
Answer is "not really." They can't reach expert-level, they're stuck around competent club level play.
ailearningchess.ai-learning-chess.workers.dev#/
Answer is "not really." They can't reach expert-level, they're stuck around competent club level play.
ailearningchess.ai-learning-chess.workers.dev#/
Peter Gostev / @petergostev:
"this model has now resolved more than 100 long-standing open problems across most areas of mathematics" I'd appreciate a heads up which ones they are, just a list of 100 titles, no need for papers
Peter Gostev / @petergostev:
"this model has now resolved more than 100 long-standing open problems across most areas of mathematics" I'd appreciate a heads up which ones they are, just a list of 100 titles, no need for papers
Peter Gostev / @petergostev:
Can't wait for this to be cancelled in 18 months' time!
Peter Gostev / @petergostev:
Can't wait for this to be cancelled in 18 months' time!
Peter Gostev / @petergostev:
What I'm curious about pacing is what it would have meant in practice, e.g. was training Mythos a mistake? It clearly kicked the competition into another gear and it was 100% because of Anthropic
Peter Gostev / @petergostev:
What I'm curious about pacing is what it would have meant in practice, e.g. was training Mythos a mistake? It clearly kicked the competition into another gear and it was 100% because of Anthropic
Superficially, this looks like a game. But really, this is a good *start* to a game. It's nowhere near a *fun* game to play.
You'd have to demonstrate the steer-ability on fine-grained details to even have a fighting chance.
x.com/petergostev...
Superficially, this looks like a game. But really, this is a good *start* to a game. It's nowhere near a *fun* game to play.
You'd have to demonstrate the steer-ability on fine-grained details to even have a fighting chance.
x.com/petergostev...
Peter Gostev / @petergostev:
Astra - recreate the original Craig Federighi in Blender.
Peter Gostev / @petergostev:
Astra - recreate the original Craig Federighi in Blender.
Peter Gostev / @petergostev:
I don't particularly believe we are in an automated research era, but I totally believe that engineering powered by AI will be eating the world
Peter Gostev / @petergostev:
I don't particularly believe we are in an automated research era, but I totally believe that engineering powered by AI will be eating the world
GPT-5.6-Luna is HALF the price - $0.20/$1.20
GPT-5.6-Luna is HALF the price - $0.20/$1.20
Peter Gostev / @petergostev:
First time I'm hearing of Douglas [image]
Peter Gostev / @petergostev:
First time I'm hearing of Douglas [image]
Peter Gostev / @petergostev:
Interesting reporting from @alexeheath: "On its current trajectory, [Gemini 4] is not expected to push frontier AI forward the way Fable and Sol just did" [image]
Peter Gostev / @petergostev:
Interesting reporting from @alexeheath: "On its current trajectory, [Gemini 4] is not expected to push frontier AI forward the way Fable and Sol just did" [image]
Peter Gostev / @petergostev:
Kimi K3 License: MIT + these 2 points: 1) Large AI hosting companies earning over $20M/year need a separate agreement. 2) Products over 100M users or $20M/month revenue must display "Kimi K3."
Peter Gostev / @petergostev:
Kimi K3 License: MIT + these 2 points: 1) Large AI hosting companies earning over $20M/year need a separate agreement. 2) Products over 100M users or $20M/month revenue must display "Kimi K3."
Peter Gostev / @petergostev:
Kimi-K3 is above the previous trend, but is at around level of 1st Jan 2026 capability of the US models - GPT-5.2 and Opus 4.5 [image]
Peter Gostev / @petergostev:
Kimi-K3 is above the previous trend, but is at around level of 1st Jan 2026 capability of the US models - GPT-5.2 and Opus 4.5 [image]