I made a notebook that includes all the goodies: QLoRA, gradient accumulation, gradient checkpointing with explanations on how they work 💝
below snapshot is with bsz=4 with simulated bsz=16 on L4 🤠
github.com/huggingface/...
I made a notebook that includes all the goodies: QLoRA, gradient accumulation, gradient checkpointing with explanations on how they work 💝
below snapshot is with bsz=4 with simulated bsz=16 on L4 🤠
github.com/huggingface/...
... the artificial intelligence system may not be trained, modified, or fine-tuned, including through recursive self-improvement, except to remove superintelligence precursor characteristics or to render inoperative the covered [AI]
... the artificial intelligence system may not be trained, modified, or fine-tuned, including through recursive self-improvement, except to remove superintelligence precursor characteristics or to render inoperative the covered [AI]
Here is the reading list:
• learning from human preferences (PPO, DPO, SimPO, CPO, RRHF, ORPO, CTO)
• real-world LLM (Llama-3, Aya, Arena's)
• efficient LLM (MoMa, LoRA, QLoRA, LESS)
Here is the reading list:
• learning from human preferences (PPO, DPO, SimPO, CPO, RRHF, ORPO, CTO)
• real-world LLM (Llama-3, Aya, Arena's)
• efficient LLM (MoMa, LoRA, QLoRA, LESS)
QLoRA fine-tuning with 4-bit with bsz of 4 can be done with 32 GB VRAM and is very fast! ✨
github.com/merveenoyan/...
QLoRA fine-tuning with 4-bit with bsz of 4 can be done with 32 GB VRAM and is very fast! ✨
github.com/merveenoyan/...
But how does it actually work?
Check out the video to learn LoRA and friends (LoRA+, QLoRA, VeRA, and DoRA)!
youtu.be/U80tjcThl9Q
But how does it actually work?
Check out the video to learn LoRA and friends (LoRA+, QLoRA, VeRA, and DoRA)!
youtu.be/U80tjcThl9Q
www.biorxiv.org/content/10.1...
huggingface.co/collections/...
huggingface.co/collections/...
www.biorxiv.org/content/10.1...
huggingface.co/collections/...
huggingface.co/collections/...
We’ve compared quantized Llama 3.2 1B QLoRA and the full precision model.
Results:
⚡Quantized model: 9.56s
⌛Full precision model: 19.14s
It’s 2x faster with quantization! 📈
Try it yourself with our example app ⭐
github.com/software-man...
We’ve compared quantized Llama 3.2 1B QLoRA and the full precision model.
Results:
⚡Quantized model: 9.56s
⌛Full precision model: 19.14s
It’s 2x faster with quantization! 📈
Try it yourself with our example app ⭐
github.com/software-man...
#lora #qlora #finetuning #llm
Origin | Interest | Match
Guide Link: medium.com/@techlatest....
#Finetuning #LLMs #AI #Agents #AImodels #LoRA #QLoRA
Guide Link: medium.com/@techlatest....
#Finetuning #LLMs #AI #Agents #AImodels #LoRA #QLoRA
But still.
But still.
¿Querés un LLM que escriba tu documentación con tu estilo? Cómo fine-tunearlo con QLoRA en una GPU de 8 GB, qué corpus necesitás y qué errores evitar
#finetuning #llm #documentacióntécnica #qlora #modeloslocales
Liquid AI's LFM2 model fine tuning using QLoRA and DPO has become a standardized workflow for generating aligned checkpoints. This...
#DeepLearning #GenerativeAI #AIInference #AIPulse
Liquid AI's LFM2 model fine tuning using QLoRA and DPO has become a standardized workflow for generating aligned checkpoints. This...
#DeepLearning #GenerativeAI #AIInference #AIPulse
It's a common theme with AI bros. They lie and they steal. Every single time. It's a grift, all of it.
It's a common theme with AI bros. They lie and they steal. Every single time. It's a grift, all of it.
Hands-on work across LLM engineering, RAG pipelines, QLoRA fine-tuning, and agent-based systems. Built real-world AI apps including multi-modal assistants, knowledge workers, and optimization tools.
#AI #LLM #RAG #Agents
· Pranjul Rathour · pranjulrathour41@gmail.com
· Pranjul Rathour · pranjulrathour41@gmail.com
Gotta say this stuff is fascinating
I have no business doing any of this but I can cause the little genie in the CLI is teaching me
Gotta say this stuff is fascinating
I have no business doing any of this but I can cause the little genie in the CLI is teaching me