So many great AI Video are available, most of them are on
@replicate.com
It's addictive, I want to try them ALL!
So many great AI Video are available, most of them are on
@replicate.com
It's addictive, I want to try them ALL!
噂通り発表された
Terraが消えるというのも噂通り
ArtificialAnalysisのベンチマークでは性能はそれほぼ変わっていない
一方、APIでは半額くらいになっていて、コストパフォーマンスがいいモデルとなっている
トークン使用量も大きくは変わっていなくて、APIの値下げは直接感じられるはず
Astraもでて、Terraの立ち位置が微妙になったのだろうが、このように命名法則がすぐに変わっていくと追っている側としてはやりづらい
Claude Opus 5.5も同日に発表されていて、Grok 4.7も先日発表されたし、競争が激化している
噂通り発表された
Terraが消えるというのも噂通り
ArtificialAnalysisのベンチマークでは性能はそれほぼ変わっていない
一方、APIでは半額くらいになっていて、コストパフォーマンスがいいモデルとなっている
トークン使用量も大きくは変わっていなくて、APIの値下げは直接感じられるはず
Astraもでて、Terraの立ち位置が微妙になったのだろうが、このように命名法則がすぐに変わっていくと追っている側としてはやりづらい
Claude Opus 5.5も同日に発表されていて、Grok 4.7も先日発表されたし、競争が激化している
ArtificialAnalysisのベンチマークを見ると、Opus 5から性能が大幅に向上し、Fableをも上回る結果を出している
API価格も20%引きとなったが、トークン使用量も増え結局は値段は同じくらいになってしまっている
今まではAPIの価格は上がる一方だったが、トップ2社が価格を下げる方向に進んでいて、コスパも注視しているのが分かる
個人的にはトップモデルが必要なことは多くなく、それなりの性能のモデルが安く使えるなら歓迎する
ArtificialAnalysisのベンチマークを見ると、Opus 5から性能が大幅に向上し、Fableをも上回る結果を出している
API価格も20%引きとなったが、トークン使用量も増え結局は値段は同じくらいになってしまっている
今まではAPIの価格は上がる一方だったが、トップ2社が価格を下げる方向に進んでいて、コスパも注視しているのが分かる
個人的にはトップモデルが必要なことは多くなく、それなりの性能のモデルが安く使えるなら歓迎する
A10B is going to feel real snappy vs Sonnet
A10B is going to feel real snappy vs Sonnet
numbers for subs from x.com/semianalysis..., cost per task is from artificialanalysis, ant says quotas are still extended but i didn't correct for that to give them the best fighting chance
numbers for subs from x.com/semianalysis..., cost per task is from artificialanalysis, ant says quotas are still extended but i didn't correct for that to give them the best fighting chance
ArtificialAnalysisのベンチマーク結果を見る限り、Solの性能はFableに匹敵と言ってよさそう
でも、個人的には、性能よりも、コスパの方が重要に感じる
Fableと同等のSolは、Fableの3分の1のコストで、ArtificialAnalysisのベンチマークを実行したという
性能も重要だけど、一般ユーザーは、それと同じくらいコスパが重要
最近のモデルは、性能は良くなっているものの、トークン使用量が増え、結局かかる費用が上がるという現象が起こっている
その中で、コスパのことを考えて作られたモデルは重要な意味を持つと思う
ArtificialAnalysisのベンチマーク結果を見る限り、Solの性能はFableに匹敵と言ってよさそう
でも、個人的には、性能よりも、コスパの方が重要に感じる
Fableと同等のSolは、Fableの3分の1のコストで、ArtificialAnalysisのベンチマークを実行したという
性能も重要だけど、一般ユーザーは、それと同じくらいコスパが重要
最近のモデルは、性能は良くなっているものの、トークン使用量が増え、結局かかる費用が上がるという現象が起こっている
その中で、コスパのことを考えて作られたモデルは重要な意味を持つと思う
It’s always been a composite over lots of benchmarks. They upgraded several, dropped some, added some, and generally skewed more toward agentic tasks
artificialanalysis.ai/methodology/...
It’s always been a composite over lots of benchmarks. They upgraded several, dropped some, added some, and generally skewed more toward agentic tasks
artificialanalysis.ai/methodology/...
#artificialanalysis
https://links.madcoolsocial.com/eXhAE2F
#artificialanalysis
https://links.madcoolsocial.com/eXhAE2F
#Anthropic #ArtificialAnalysis #FreeCAD
Read more on Simon Willison's Weblog: https://postreads.co/feed-item/90502/click?source=bluesky
#Anthropic #ArtificialAnalysis #FreeCAD
Read more on Simon Willison's Weblog: https://postreads.co/feed-item/90502/click?source=bluesky
[EN] In-Depth Analysis of GPT-6 Sol (Max): Intelligence, Performance, and Prici…
https://ai-minor.com/blog/en/2026-09-23-1790117983886-gpt_6_sol__max__intelligence__performance_and_pric
#GPT-6Sol #ベンチマーク #ArtificialAnalysis #AI #Tech
[EN] In-Depth Analysis of GPT-6 Sol (Max): Intelligence, Performance, and Prici…
https://ai-minor.com/blog/en/2026-09-23-1790117983886-gpt_6_sol__max__intelligence__performance_and_pric
#GPT-6Sol #ベンチマーク #ArtificialAnalysis #AI #Tech
ArtificialAnalysisのベンチマークを見る限り、ClaudeのFableとOpusの間の性能
価格はClaude系等の同等性能のモデルと比較すると相当安い(xhighですらGemini 3.8 FlashやGPT 5.6 Terra(max)の同じくらい)
また、スピードも速い(GeminiのFlashに匹敵するくらい)
1Mあたりの価格も性能にしては相当安く設定されている
Muse Codeも発表し、本格的にAnthropicやOpenAI2対抗してきた
性能も問題なく、その割に速度が速く価格が安い
トップモデルの競争にMetaが参入してきた
ArtificialAnalysisのベンチマークを見る限り、ClaudeのFableとOpusの間の性能
価格はClaude系等の同等性能のモデルと比較すると相当安い(xhighですらGemini 3.8 FlashやGPT 5.6 Terra(max)の同じくらい)
また、スピードも速い(GeminiのFlashに匹敵するくらい)
1Mあたりの価格も性能にしては相当安く設定されている
Muse Codeも発表し、本格的にAnthropicやOpenAI2対抗してきた
性能も問題なく、その割に速度が速く価格が安い
トップモデルの競争にMetaが参入してきた
This is exactly the direction #AI should take.
Users don’t care which #LLM is slightly better at arithmetic, programming, or isolated tasks. The real challenge is #multidisciplinary AI—models that can handle #real-world problems holistically.
x.com/ArtificialAn...
This is exactly the direction #AI should take.
Users don’t care which #LLM is slightly better at arithmetic, programming, or isolated tasks. The real challenge is #multidisciplinary AI—models that can handle #real-world problems holistically.
x.com/ArtificialAn...