Developers like to say their work is so difficult and unique that it must have a 500b parameter model...
Bruh that webhook can be done by granite4 tiny in a web browser on mobile.
Analogy: Pixar for most of their rendering farms used commodity CPU and not GPU.
Developers like to say their work is so difficult and unique that it must have a 500b parameter model...
Bruh that webhook can be done by granite4 tiny in a web browser on mobile.
Analogy: Pixar for most of their rendering farms used commodity CPU and not GPU.
That's why you want a small one with tuning. I need this thing to track HTTP requests and not roleplay a waifu trained on Nietzsche
One of the models I get the most work out of is granite4. This is not an endorsement. It's "dumb" enough to do work with minimal gremlin mode.
That's why you want a small one with tuning. I need this thing to track HTTP requests and not roleplay a waifu trained on Nietzsche
One of the models I get the most work out of is granite4. This is not an endorsement. It's "dumb" enough to do work with minimal gremlin mode.
www.the-main-thread.com/p/quarkus-la...
www.the-main-thread.com/p/quarkus-la...
Get them on Docker Hub:
https://hub.docker.com/r/ai/granite-4.0-nano
https://hub.docker.com/r/ai/granite-4.0-h-nano
#Docker #Granite4 #LLM #AI #DevTools #IBM
Get them on Docker Hub:
https://hub.docker.com/r/ai/granite-4.0-nano
https://hub.docker.com/r/ai/granite-4.0-h-nano
#Docker #Granite4 #LLM #AI #DevTools #IBM
ollama.com/library/gran...
#ollama #AI #LLM #ibm #granite4 #granite
If you're stuck with something based on ChatGPT that's a platform problem and why everything is shitting the bed
Or, of course, not have any of that
If you're stuck with something based on ChatGPT that's a platform problem and why everything is shitting the bed
Or, of course, not have any of that
(Yes, yes AI. I like tech and I like to play. It's running locally, in my kitchen on a box that was on 24/7. It's still unethical, but I've tried to be as ethical as I can be)
(Yes, yes AI. I like tech and I like to play. It's running locally, in my kitchen on a box that was on 24/7. It's still unethical, but I've tried to be as ethical as I can be)
Ex: little granite4 running in ram bouncing made up data and various fuzzing off of endpoints, but if you say "[AI, LLM] for testing" someone will say it stole Harry Potter and can't dream or run commands
Ex: little granite4 running in ram bouncing made up data and various fuzzing off of endpoints, but if you say "[AI, LLM] for testing" someone will say it stole Harry Potter and can't dream or run commands
#AI #IBM #EnterpriseAI #OpenSource #Mamba #Granite4 #AIModels #HybridModels
winbuzzer.com/2025/10/03/i...
#AI #IBM #EnterpriseAI #OpenSource #Mamba #Granite4 #AIModels #HybridModels
winbuzzer.com/2025/10/03/i...
You can use granite4 on 2G of RAM completely offline. It won't tell you to kill yourself. You don't have to kill wildlife.
One of the earliest LLM projects I did was reorienting solar panels :) It's not fantasy. It's at least a decade old.
You can use granite4 on 2G of RAM completely offline. It won't tell you to kill yourself. You don't have to kill wildlife.
One of the earliest LLM projects I did was reorienting solar panels :) It's not fantasy. It's at least a decade old.
granite4/1b won't pick one
smollm3 wont either
qwen3-vl:235b won't either
deepseek-v3.1:671b won't either
yawn. nothing to see here. "AI" needs to die, soon.
granite4/1b won't pick one
smollm3 wont either
qwen3-vl:235b won't either
deepseek-v3.1:671b won't either
yawn. nothing to see here. "AI" needs to die, soon.
詳しくはこちら↓↓↓
gamefi.co.jp/2025/10/06/i...
詳しくはこちら↓↓↓
gamefi.co.jp/2025/10/06/i...
Models like old granite4 (I have a love/hate with this thing) are enough to power docling and langextract on CPU only. If I had to do that again I'd start with Gemma3n instead.
Models like old granite4 (I have a love/hate with this thing) are enough to power docling and langextract on CPU only. If I had to do that again I'd start with Gemma3n instead.
Even if you moved that to a static site or spindown container, that still takes the same or more energy than it costs me to run IBM's granite4 in 2G of RAM.
Try it :)
Even if you moved that to a static site or spindown container, that still takes the same or more energy than it costs me to run IBM's granite4 in 2G of RAM.
Try it :)
Granite is super fast, better than llama imo, but the tool calling still gets a bit confused...might try the larger one. Qwen likes to think...a lot.
Granite is super fast, better than llama imo, but the tool calling still gets a bit confused...might try the larger one. Qwen likes to think...a lot.
#AI #LLM #OpenSourceAI #EnterpriseAI #PrivateAI #LocalAI #SpreadsheetAI #Ollama #UnPerplexedSpready
#AI #LLM #OpenSourceAI #EnterpriseAI #PrivateAI #LocalAI #SpreadsheetAI #Ollama #UnPerplexedSpready
matasoft.hr/qtrendcontro...
#AI #LLM #Ollama #OllamaCloud #Searxng
matasoft.hr/qtrendcontro...
#AI #LLM #Ollama #OllamaCloud #Searxng
#AI #LLM #Ollama
#AI #LLM #Ollama