* no tokenizer, completely gone, it directly interprets bytes
* performance on par with OLMo 3
* everything is open, as usual for @ai2.bsky.social
allenai.org/blog/bolmo
* no tokenizer, completely gone, it directly interprets bytes
* performance on par with OLMo 3
* everything is open, as usual for @ai2.bsky.social
allenai.org/blog/bolmo
wdym you have olmo, molmo, and fucking bolmo
wdym you have olmo, molmo, and fucking bolmo
📝 Blog: buff.ly/WjU6bXV
⬇️ Download Bolmo 7B: buff.ly/pkCdFWl | 1B: buff.ly/QRBm6hC
📄 Report: buff.ly/08ouIoy
📝 Blog: buff.ly/WjU6bXV
⬇️ Download Bolmo 7B: buff.ly/pkCdFWl | 1B: buff.ly/QRBm6hC
📄 Report: buff.ly/08ouIoy
Claude has no token for Bolmo (because Bolmo is new) and so can't directly encode that name, but it can *decode* it somehow - it might get split into multiple tokens in some sentences, but not others?
also extremely ironic in this context
Claude has no token for Bolmo (because Bolmo is new) and so can't directly encode that name, but it can *decode* it somehow - it might get split into multiple tokens in some sentences, but not others?
also extremely ironic in this context
Main Link | Techmeme Permalink
Main Link | Techmeme Permalink
i wanna see what bolmo looks like
i wanna see what bolmo looks like
Byteification: AI2's New Bolmo AI Model Cuts AI Training Costs by 99%
#AI #AI2 #LLMs #OpenSourceAI #AIResearch #MachineLearning #Bolmo #ByteLevelAI #Tokenization #ModelEfficiency #DeepLearning
Byteification: AI2's New Bolmo AI Model Cuts AI Training Costs by 99%
#AI #AI2 #LLMs #OpenSourceAI #AIResearch #MachineLearning #Bolmo #ByteLevelAI #Tokenization #ModelEfficiency #DeepLearning
Институт искусственного интеллекта Аллена (AI2) разработал Bolmo, новое семейство языковых моделей байтового уровня. Bolmo использует существующие модели Olmo 3 …
Telegram ИИ Дайджест
#ai #news
Институт искусственного интеллекта Аллена (AI2) разработал Bolmo, новое семейство языковых моделей байтового уровня. Bolmo использует существующие модели Olmo 3 …
Telegram ИИ Дайджест
#ai #news
(1) Improving Recursive Transformers with Mixture of LoRAs
(2) <a href="https://researchtrend.ai/papers/2512.15586" class="hover:underline text-blue-600 dark:text-sky-400 no-card-link" target="_blank" rel="noopener" data-link="bsky">Bolmo: Byteifying the Next Generation of Language Models
(3) Bolmo: Byteifying the Next Generation of Language Models
🔍 More at researchtrend.ai/communities/MoE
(1) Improving Recursive Transformers with Mixture of LoRAs
(2) <a href="https://researchtrend.ai/papers/2512.15586" class="hover:underline text-blue-600 dark:text-sky-400 no-card-link" target="_blank" rel="noopener" data-link="bsky">Bolmo: Byteifying the Next Generation of Language Models
(3) Bolmo: Byteifying the Next Generation of Language Models
🔍 More at researchtrend.ai/communities/MoE
We'll wait.
You can have some reading on a strategy that uses utf8 instead of "words". Plenty to look up from here including some that can indeed make images.
allenai.org/blog/bolmo
We'll wait.
You can have some reading on a strategy that uses utf8 instead of "words". Plenty to look up from here including some that can indeed make images.
allenai.org/blog/bolmo
Published on Tuesday, December 16, 2025
Published on Tuesday, December 16, 2025
🔗 aidailypost.com/news/bolmo-a...
🔗 aidailypost.com/news/bolmo-a...
@allen_ai:
Introducing Bolmo, a new family of byte-level language models built by "byteifying" our open Olmo 3--and to our knowledge, the first fully open byte-level LM to match or surpass SOTA subword models across a wide range of tasks. 🧵 [image]
@allen_ai:
Introducing Bolmo, a new family of byte-level language models built by "byteifying" our open Olmo 3--and to our knowledge, the first fully open byte-level LM to match or surpass SOTA subword models across a wide range of tasks. 🧵 [image]
https://allenai.org/blog/bolmo
https://allenai.org/blog/bolmo
Bolmo: Byteifying the Next Generation of Language Models
https://arxiv.org/abs/2512.15586
Bolmo: Byteifying the Next Generation of Language Models
https://arxiv.org/abs/2512.15586