You can basically use any <10B LLM directly in the browser using #WebLLM. .
You can basically use any <10B LLM directly in the browser using #WebLLM. .
WebLLM : un moteur Open Source permettant d’exécuter des LLM entièrement dans le navigateur, sans serveur d’inférence, grâce à l’accélération GPU fournie par WebGPU.
WebLLM : un moteur Open Source permettant d’exécuter des LLM entièrement dans le navigateur, sans serveur d’inférence, grâce à l’accélération GPU fournie par WebGPU.
https://github.com/mlc-ai/web-llm
📬 Recevoir ma veille
https://github.com/mlc-ai/web-llm
📬 Recevoir ma veille
• Tiny 1.7B LLM running at 88 tokens / second ⚡
• Powered by MLC/WebLLM on WebGPU 🔥
• JSON Structured Generation entirely in the browser 🤏
• Tiny 1.7B LLM running at 88 tokens / second ⚡
• Powered by MLC/WebLLM on WebGPU 🔥
• JSON Structured Generation entirely in the browser 🤏
Baris Guler takes that idea seriously in our first-ever guest post on the Mozilla.ai blog:
🧱 WebLLM + WASM + WebWorkers
💻 Rust, Go, Python, JS
🔒 Fully local. No API calls.
Read the post here:
blog.mozilla.ai/3w-for-in-br...
Baris Guler takes that idea seriously in our first-ever guest post on the Mozilla.ai blog:
🧱 WebLLM + WASM + WebWorkers
💻 Rust, Go, Python, JS
🔒 Fully local. No API calls.
Read the post here:
blog.mozilla.ai/3w-for-in-br...
- No signup required
- 100% data privacy
- No API keys 🔒
Supported Models:
• #SmolLM2
• #Qwen
• #Llama
• #DeepSeek R1
• #Phi
• #TinyLlama
• & More
🔗 zalt.me/tools/free-a...
#FreeAI #Free #WebGPU #AI #WebLLM #ChatGPT #JS #OpenSource #LLM #Chrome
WebLLM runs LLMs directly in your browser with WebGPU acceleration, offering OpenAI API compatibility and enabling local, private AI tasks.
WebLLM runs LLMs directly in your browser with WebGPU acceleration, offering OpenAI API compatibility and enabling local, private AI tasks.
Still at a very early stage of course, but making some good progress!
Thanks to WebLLM, which brings hardware accelerated language model inference onto web browsers, via WebGPU 🚀
Still at a very early stage of course, but making some good progress!
Thanks to WebLLM, which brings hardware accelerated language model inference onto web browsers, via WebGPU 🚀
It's a chatbot pretending to be a human pretending to be an artist pretending to be a crayfish. No data gets sent anywhere - everything happens locally using Phi-3.5-mini in your browser using WebLLM technology.
It's a chatbot pretending to be a human pretending to be an artist pretending to be a crayfish. No data gets sent anywhere - everything happens locally using Phi-3.5-mini in your browser using WebLLM technology.
#aiTech #buildInPublic #webDev - 1/3
#aiTech #buildInPublic #webDev - 1/3
https://gigazine.net/news/20260214-on-device-browser-agent/
https://gigazine.net/news/20260214-on-device-browser-agent/
Shouts to @una.im (fire blazer, btw) & @matthiasrohmer.bsky.social !
Just updated my Antigravity & added Modern Web Guidance.
Going to try that "Ask Gemini" with a #WebMCP demo I made. Currently it uses Prompt API and WebLLM.
Shouts to @una.im (fire blazer, btw) & @matthiasrohmer.bsky.social !
Just updated my Antigravity & added Modern Web Guidance.
Going to try that "Ask Gemini" with a #WebMCP demo I made. Currently it uses Prompt API and WebLLM.
Today, I'm dealing with WebLLM and how to use it to do "AI" stuff in my app without blowing my bootstrapped budget for the app I'm building, Messijo.
Today, I'm dealing with WebLLM and how to use it to do "AI" stuff in my app without blowing my bootstrapped budget for the app I'm building, Messijo.
chat.webllm.ai
But 3% of browsers still don't support WebAssembly, 18% don't WebGPU.
chat.webllm.ai
But 3% of browsers still don't support WebAssembly, 18% don't WebGPU.
An open-source project runs full LLMs entirely inside your browser tab, no server or API call needed.
An open-source project runs full LLMs entirely inside your browser tab, no server or API call needed.
#WebLLM Chat brings #AI conversations to your browser using #WebGPU - runs locally for complete #privacy and #offline use. Built on #opensource tech, supports #customLLM and image analysis. More at chat.webllm.ai
#WebLLM Chat brings #AI conversations to your browser using #WebGPU - runs locally for complete #privacy and #offline use. Built on #opensource tech, supports #customLLM and image analysis. More at chat.webllm.ai
Covered:
✨ JupyterLab 4.4, Notebook 7.4
🧪 JupyterLite 0.6
🌐 In-browser Python/R
🖥️ Terminal w/ Vim
🧠 AI (WebLLM, on-device)
⚡ Hybrid kernels
Thanks Bloomberg & all who joined!
youtu.be/7kS_xfKEOmM
Covered:
✨ JupyterLab 4.4, Notebook 7.4
🧪 JupyterLite 0.6
🌐 In-browser Python/R
🖥️ Terminal w/ Vim
🧠 AI (WebLLM, on-device)
⚡ Hybrid kernels
Thanks Bloomberg & all who joined!
youtu.be/7kS_xfKEOmM
* As part of the browser's API, obviously
the new innovation is Per-Layer Embeddings, which let it consume dramatically less memory
it was created for phones, and is being rolled out to Android phones soon
developers.googleblog.com/en/introduci...
* As part of the browser's API, obviously
#WebAI #PromptAPI #WebLLM #GenAI
#WebAI #PromptAPI #WebLLM #GenAI