“Llama Masters 3” by Scott Hammack
[llama3/llama3.zzt] - “Llama Masters 3”
https://museumofzzt.com/file/play/llama3/
“Llama Masters 3” by Scott Hammack
[llama3/llama3.zzt] - “Llama Masters 3”
https://museumofzzt.com/file/play/llama3/
Inkling is a general-purpose model, designed to be adapted via fine-tuning to whatever task you need.
It currently ranks among the best open-weight models.
Good work from the Thinking Machines team.
Try it out here: github.com/mukel/llama3...
@stephanjanssen.be @mukel.bsky.social
#VDCERN #VoxxedDaysCERN
Try it out here: github.com/mukel/llama3...
@stephanjanssen.be @mukel.bsky.social
#VDCERN #VoxxedDaysCERN
✅ Llama3.java
✅ Vector API, FFM API
✅ Apple Silicon ❤️
— `git clone github.com/mukel/llama3... `
— `sdk install java 25.ea.17-graal`
— `make native` (optionally preload a model for zero overhead)
— Profit!🚀
#Java #GraalVM #LLM #LLama
✅ Llama3.java
✅ Vector API, FFM API
✅ Apple Silicon ❤️
— `git clone github.com/mukel/llama3... `
— `sdk install java 25.ea.17-graal`
— `make native` (optionally preload a model for zero overhead)
— Profit!🚀
#Java #GraalVM #LLM #LLama
Graal compiler: +10% faster inference with the latest early access build.
New features: batched prompt processing & AVX512 support.
Graal compiler: +10% faster inference with the latest early access build.
New features: batched prompt processing & AVX512 support.
Nicely done!👏
Nicely done!👏
Implement llama3 from scratch using jax in just 100 lines of code. Why jax? Because jax looks like a NumPy wrapper but it has cool features like xla; a linear algebra accelerator, jit, vmap, pmap etc., which makes your training go brr brr.
Implement llama3 from scratch using jax in just 100 lines of code. Why jax? Because jax looks like a NumPy wrapper but it has cool features like xla; a linear algebra accelerator, jit, vmap, pmap etc., which makes your training go brr brr.
1. found a podcast for me
2. sent a note making sure i'm connecting with people
3. ran an experiment on llama3 3b
4. came up with a new theory about MoE arch providing "live-ness"
5. updated it's chicken-scratch notes for an upcoming blog post
1. found a podcast for me
2. sent a note making sure i'm connecting with people
3. ran an experiment on llama3 3b
4. came up with a new theory about MoE arch providing "live-ness"
5. updated it's chicken-scratch notes for an upcoming blog post
Want an agent that can run entirely locally? Check out this tutorial that combines Adaptive RAG, Corrective RAG, and Self-RAG
Blog: www.elastic.co/search-labs/...
#ai #langchain
Want an agent that can run entirely locally? Check out this tutorial that combines Adaptive RAG, Corrective RAG, and Self-RAG
Blog: www.elastic.co/search-labs/...
#ai #langchain
Spring Boot + Llama3-java + @graalvm.org native-image
Java-native LLMs in an instant 🍃🐰🦙
Spring Boot + Llama3-java + @graalvm.org native-image
Java-native LLMs in an instant 🍃🐰🦙
arxiv.org/abs/2502.09992
Significant progress towards language diffusion models. Reportedly on par with LLaMA3 on many benchmarks.
arxiv.org/abs/2502.09992
Significant progress towards language diffusion models. Reportedly on par with LLaMA3 on many benchmarks.
#quarkus #Java #langchain4j #llama
#quarkus #Java #langchain4j #llama
Graal compiler: +10% faster inference with the latest early access build.
New features: batched prompt processing & AVX512 support.
You can wait for the publications for the rest! (Yes I know you are being snarky)
You can wait for the publications for the rest! (Yes I know you are being snarky)
Built a custom Llama3 agent with Ollama! 🛠️ Using Modelfiles to set a "Senior Dev" persona that calls out my "technical debt." Real feedback, zero sycophancy, 100% local. 💻🚀
#Ollama #Llama3 #CustomAI #DevLife #BuildInPublic #LocalAI
llama up, chatgpt down
llama up, chatgpt down