x.com/cognition/st...
x.com/cognition/st...
On FrontierCode v1.1 it matches Fable 5, at half the price.
On FrontierCode v1.1 it matches Fable 5, at half the price.
(also, why no xhigh effort level? that’s sick…)
(also, why no xhigh effort level? that’s sick…)
/2
/2
Cognition has released SWE-2, a new coding model that matches Fable 5.1's performance on FrontierCode 1.1 Main1 at 64% lower cost, scaling RL to multi-trillion parameters.
It's a strong move towards making high-end coding assistance more accessible.
Cognition has released SWE-2, a new coding model that matches Fable 5.1's performance on FrontierCode 1.1 Main1 at 64% lower cost, scaling RL to multi-trillion parameters.
It's a strong move towards making high-end coding assistance more accessible.
Benchmarks GPT-6 Sol & Luna:
• FrontierCode: 48.4% vs. Fable 5.1’s 48.7%, both at xhigh. $1.37 vs. $9.27 per task, roughly 85% cheaper.
• DeepSWE: 68.8% at max vs. Fable 5’s best result of 69.9% at xhigh. $2.74 vs. $13.41 per tas...
https://x.com/i/web/status/2102462091299098640
Benchmarks GPT-6 Sol & Luna:
• FrontierCode: 48.4% vs. Fable 5.1’s 48.7%, both at xhigh. $1.37 vs. $9.27 per task, roughly 85% cheaper.
• DeepSWE: 68.8% at max vs. Fable 5’s best result of 69.9% at xhigh. $2.74 vs. $13.41 per tas...
https://x.com/i/web/status/2102462091299098640
FrontierCode is a new AI concept that has the potential to revolutionize the field of artificial intelligence. This innovative approach could lead to sig...
🌐 aitechcodex.uk/news/2206/?utm_source=bluesky&utm_medium=social&utm_campaign=daily_news 📡 t.me/AITechNewsUK
FrontierCode is a new AI concept that has the potential to revolutionize the field of artificial intelligence. This innovative approach could lead to sig...
🌐 aitechcodex.uk/news/2206/?utm_source=bluesky&utm_medium=social&utm_campaign=daily_news 📡 t.me/AITechNewsUK
Every benchmark you've seen measures: does this code run?
FrontierCode measures: would a senior maintainer accept this PR?
Scope discipline. Style consistency. Regression safety.
Every benchmark you've seen measures: does this code run?
FrontierCode measures: would a senior maintainer accept this PR?
Scope discipline. Style consistency. Regression safety.
iirc, FrontierCode is measuring "mergeability" based on a bunch of heuristics determined by maintainers, rather than "does it complete the task". Makes it really noisy.
iirc, FrontierCode is measuring "mergeability" based on a bunch of heuristics determined by maintainers, rather than "does it complete the task". Makes it really noisy.
النموذج المفتوح المصدر اللي كسر الدنيا وبقى رقم 1 على FrontierCode
قعدت أختبره في حاجات حقيقية، وشوفوا عمل إيه
#حسام_الدين_حسن #خبير_اونلاين #Kimi #Kimi_K3 #KimiK3 #مفتوح_المصدر #OpenSource #الذكاء_الاصطناعي #نماذج_الذكاء_الاصطناعي #ذكاء_اصطناعي_عربي #Fable_5
النموذج المفتوح المصدر اللي كسر الدنيا وبقى رقم 1 على FrontierCode
قعدت أختبره في حاجات حقيقية، وشوفوا عمل إيه
#حسام_الدين_حسن #خبير_اونلاين #Kimi #Kimi_K3 #KimiK3 #مفتوح_المصدر #OpenSource #الذكاء_الاصطناعي #نماذج_الذكاء_الاصطناعي #ذكاء_اصطناعي_عربي #Fable_5
#Google #Gemini
aidisruption.ai/p/musks-grok...
#Google #Gemini
aidisruption.ai/p/musks-grok...
- grok 4.5 (high): cheap + fast
- opus 5 (medium): best
- fable 5 (high): if you're feeling fancy
and forget the rest.
Now time to see if Opus 5 is really that good. Source: cognition.com/frontiercode
- grok 4.5 (high): cheap + fast
- opus 5 (medium): best
- fable 5 (high): if you're feeling fancy
and forget the rest.
Now time to see if Opus 5 is really that good. Source: cognition.com/frontiercode
Así estaba hasta ahora, casi a cero,y Fable 5 da el salto hasta el 30% de golpe!
Así estaba hasta ahora, casi a cero,y Fable 5 da el salto hasta el 30% de golpe!
【コード生成AI「SWE-2」徹底検証:Fable 5.1同等性能を64%コスト削減した Cognition …】
CognitionのコーディングモデルSWE-2を早速検証。Kimi K3をポストトレーニングした結果、FrontierCodeでFable 5.1並みの精度を叩き出しつつ、コストを64%も削減できて驚いた。実務のコード生成コスト改善に直結する。…
👇 詳細・一次ソース解説
https://labomaru.com/posts/20260913121123/
#AI速報 #らぼまる #AIツール
【コード生成AI「SWE-2」徹底検証:Fable 5.1同等性能を64%コスト削減した Cognition …】
CognitionのコーディングモデルSWE-2を早速検証。Kimi K3をポストトレーニングした結果、FrontierCodeでFable 5.1並みの精度を叩き出しつつ、コストを64%も削減できて驚いた。実務のコード生成コスト改善に直結する。…
👇 詳細・一次ソース解説
https://labomaru.com/posts/20260913121123/
#AI速報 #らぼまる #AIツール
->Startup Fortune | More on "AI coding agents cost efficiency" at BigEarthData.ai