www.reddit.com/r/LocalLLM/c...
The 1-bit model retains ~78.9% accuracy after we shrunk it from 1.56TB to 594GB (-62% size).
Run on a Mac Studio + 128GB RAM device.
Kimi K3 is the strongest open model to date.
Guide: unsloth.ai/docs/models/...
GGUF: huggingface.co/unsloth/Kimi...
www.reddit.com/r/LocalLLM/c...
Man, this is fucking cool. It's so fast! And the author trained it themselves from TLDR pages and command examples. Best of all, you don't need to point this at a llama-server; it'll set one up for itself and keep it resident in the background, and RAM usage is small.
Man, this is fucking cool. It's so fast! And the author trained it themselves from TLDR pages and command examples. Best of all, you don't need to point this at a llama-server; it'll set one up for itself and keep it resident in the background, and RAM usage is small.
The competitive pressure rn is crazy
The competitive pressure rn is crazy
It lost at the one thing it was built for: asked something no tool could answer, it wrote six paragraphs ending "What's your Tuesday vibe?" Vs null
#LocalLLM #SelfHosted #Homelab #AIAgents #OpenSourceAI
It lost at the one thing it was built for: asked something no tool could answer, it wrote six paragraphs ending "What's your Tuesday vibe?" Vs null
#LocalLLM #SelfHosted #Homelab #AIAgents #OpenSourceAI
The result? 🤯 On open-source LLMs (Ollama), it actually beats today's NPUs and €2000 AI laptops.
#LocalLLM #Ollama #HardwareHacking #AI #TechSalvage
The result? 🤯 On open-source LLMs (Ollama), it actually beats today's NPUs and €2000 AI laptops.
#LocalLLM #Ollama #HardwareHacking #AI #TechSalvage
www.reddit.com/r/LocalLLM/c...
www.reddit.com/r/LocalLLM/c...
reddit randos: *slaps sticker on cix armv9 machine* sovereign.
reddit randos: *slaps sticker on cix armv9 machine* sovereign.
It can now generate its own code and use existing project files as context.
All running locally. Real local.
No API keys. No third-party bills.
#AI #LocalLLM #Agents #Programming #IndieDev #SelfHosted #BuildInPublic
It can now generate its own code and use existing project files as context.
All running locally. Real local.
No API keys. No third-party bills.
#AI #LocalLLM #Agents #Programming #IndieDev #SelfHosted #BuildInPublic
apfel is a open-source local llm. CLI and OpenAI-compatible local server for Apple's on-device foundation model.
→ Full source available on the repo link.
→ Released under o...
#opensource #selfhosted #localllm
apfel is a open-source local llm. CLI and OpenAI-compatible local server for Apple's on-device foundation model.
→ Full source available on the repo link.
→ Released under o...
#opensource #selfhosted #localllm
Test post, curious if anyone cares. Trained a LoRA for an open source LLM, fully local, no cloud dependency. Two days everything broke, then it finally worked. Video coming if there's interest.
Test post, curious if anyone cares. Trained a LoRA for an open source LLM, fully local, no cloud dependency. Two days everything broke, then it finally worked. Video coming if there's interest.
Following @garymarcus.bsky.social is a good start...
• It's not smarter than earlier models, just trained more cheaply
• It doesn't solve hallucinations or problems with reliability.
Rest of post at open.substack.com/pub/garymarc...
Following @garymarcus.bsky.social is a good start...
www.reddit.com/r/LocalLLM/s...
www.reddit.com/r/LocalLLM/s...
Run LLMs locally on Cloud Workstations. Uses:
Quantized models from 🤗
llama-cpp-python's webserver"
github.com/GoogleCloudP...
Run LLMs locally on Cloud Workstations. Uses:
Quantized models from 🤗
llama-cpp-python's webserver"
github.com/GoogleCloudP...
The automation potential of UE5.8 MCP is unreal! 🛠️🔥
#UE5 #UnrealEngine5 #ModelContextProtocol #LMStudio #Qwen25 #GameDev #IndieDev #LocalLLM #AI @UnrealEngine
The automation potential of UE5.8 MCP is unreal! 🛠️🔥
#UE5 #UnrealEngine5 #ModelContextProtocol #LMStudio #Qwen25 #GameDev #IndieDev #LocalLLM #AI @UnrealEngine
Ollama is an AI tool that allows you to run Large Language Models (LLMs) on device (local LLM).
Ollama blog: ollama.com/blog/functio...
#ollama #python #LLMs #ml #ai #localLLM #localai
Ollama is an AI tool that allows you to run Large Language Models (LLMs) on device (local LLM).
Ollama blog: ollama.com/blog/functio...
#ollama #python #LLMs #ml #ai #localLLM #localai
No SaaS.
No GPU chaos.
No fragile scripts.
This tutorial shows how to:
→ Run models in containers with RamaLama
→ Expose an OpenAI-compatible API
→ Connect it to Quarkus via LangChain4j
Clean. Reproducible. Production-aware.
buff.ly/lHMEpTl
#Java #AI #Quarkus #LocalLLM
No SaaS.
No GPU chaos.
No fragile scripts.
This tutorial shows how to:
→ Run models in containers with RamaLama
→ Expose an OpenAI-compatible API
→ Connect it to Quarkus via LangChain4j
Clean. Reproducible. Production-aware.
buff.ly/lHMEpTl
#Java #AI #Quarkus #LocalLLM