#PrunaAI
PVideo Avatar — генератор видео с ИИ-аватарами Загружаешь

фото и аудио (или описываешь голосом текстом) — получаешь видео с говорящим аватаром. Поддерживает русский язык. Работает через платформу Replicate, где можно протестировать модель.

Проверить: https://replicate.com/prunaai/p-video-avatar
August 18, 2026 at 11:29 AM
Two seconds per image is seriously fast. P-Image is now on Magnific with 13 aspect ratios and 1K or 2K output. Change one detail, see it, adjust again while the idea is still fresh.

x.com/magnific/st...
Magnific (@magnific) on X
P-Image by @ideogram_ai x @PrunaAI is now on Magnific → Images in just 2 seconds → 13 different aspect ratios → 1K or 2K resolution Available now on Magnific
x.com
August 5, 2026 at 6:41 PM
Same prompt. Four AI image models. Four completely different results.

Nano Banana 2 Lite, PrunaAI/P-Image-Edit, Ideogram V4 Quality, and GPT-Image-2 each gave the same monster claw + energy drink scene a unique look.

Which one wins? Comment below.

#AI #PromptEngineering
July 24, 2026 at 12:07 PM
i like rotating them around but current roster is hitting replicate.com api with:

prunaai/hidream-l1-fast
bytedance/seedream-4.5
black-forest-labs/flux-2-max
openai/gpt-image-2

im a huge fan of seedream rn. I've never really done loras much, i like getting the image models as "base" as i can
July 13, 2026 at 4:15 PM
Pruna v0.3.4 is out! 🚀

👉 Read the notes: buff.ly/CElIET2

Here are a few highlights 👇
Release v0.3.4 · PrunaAI/pruna
The juiciest bits 🧃 New algorithms joining the garden 🌱 We added a fresh batch of algorithm support this release, and this one deserves an extra round of applause: all of these new algorithm…
buff.ly
June 25, 2026 at 3:03 PM
Thank you for contributing to Pruna OSS!

⭐️ Star our OSS: buff.ly/Y9j29GO
📚 API & quickstart: buff.ly/VLgjLgw
🌐 Website: www.pruna.ai
GitHub - PrunaAI/pruna: Pruna is a model optimization framework built for developers, enabling you to deliver faster, more efficient models with minimal overhead.
Pruna is a model optimization framework built for developers, enabling you to deliver faster, more efficient models with minimal overhead. - PrunaAI/pruna
buff.ly
June 9, 2026 at 3:03 PM
💜 Simply make AI models faster, cheaper, smaller, greener!

The toolkit is designed with simplicity in mind - requiring just a few lines of code to optimize your models. It supports various model types, including LLMs, Diffusion, and more!

⭐️ Give us a star: github.com/PrunaAI/pruna
June 5, 2026 at 3:04 PM
Thank you for contributing to Pruna OSS!

⭐️ Star our OSS: buff.ly/Y9j29GO
📚 API & quickstart: buff.ly/VLgjLgw
🌐 Website: www.pruna.ai
GitHub - PrunaAI/pruna: Pruna is a model optimization framework built for developers, enabling you to deliver faster, more efficient models with minimal overhead.
Pruna is a model optimization framework built for developers, enabling you to deliver faster, more efficient models with minimal overhead. - PrunaAI/pruna
buff.ly
May 27, 2026 at 3:00 PM
This algorithm progressively merges similar tokens between the attention and MLP stages of each transformer block, reducing the number of tokens processed during inference. The result: faster inference with minimal impact on quality.
GitHub - PrunaAI/pruna: Pruna is a model optimization framework built for developers, enabling you to deliver faster, more efficient models with minimal overhead.
Pruna is a model optimization framework built for developers, enabling you to deliver faster, more efficient models with minimal overhead. - PrunaAI/pruna
buff.ly
May 27, 2026 at 3:00 PM
⭐️ Star our OSS: github.com/PrunaAI/pruna
📚 API & quickstart: docs.api.pruna.ai/guides/quick...
🌐 Website: www.pruna.ai
May 4, 2026 at 6:18 AM
🚀 Pruna 0.3.3 is out!

Here are a few highlights:
- New benchmarking and evaluation tools.
- Algorithm and compatibility upgrades.
- Support for Python 3.13.
- A wide range of bug fixes and reliability improvements.

Thank you all for your contribution!

👉 Read the notes: github.com/PrunaAI/prun...
GitHub - PrunaAI/pruna: Pruna is a model optimization framework built for developers, enabling you to deliver faster, more efficient models with minimal overhead.
Pruna is a model optimization framework built for developers, enabling you to deliver faster, more efficient models with minimal overhead. - PrunaAI/pruna
github.com
May 4, 2026 at 6:18 AM
First Prune is almost over 🔥

This is our last call for anyone who wants to contribute to Pruna OSS.
Join before April 30: pick an open issue, ask to be assigned, and submit your PR.

👉 Search for your issue: github.com/PrunaAI/prun...
👉 Read more about Pruna OSS: www.pruna.ai/blog/first-p...
April 24, 2026 at 3:02 PM
We launched two speedy diffusion language model endpoints. llada2.1-mini hits ~1000+ TPS; llada2.1-flash reaches ~800+ TPS for fast, high-volume language tasks. See them on Replicate: replicate.com/prunaai/llad...

⭐️ Star our OSS: github.com/PrunaAI/pruna
prunaai/llada2.1-mini – Replicate
The fastest diffusion language model with up to ~1000+ tps
replicate.com
April 20, 2026 at 3:04 PM
We just launched two new models on Replicate!

• ERNIE-Image → higher quality (~33s)
• ERNIE-Image-Turbo → faster iteration (~6s)

No prompt engineering required — our built-in prompt enhancer handles it.

Try them:
replicate.com/prunaai/erni...
replicate.com/prunaai/erni...
prunaai/ernie-image – Replicate
ERNIE-Image is an open text-to-image generation model developed by the ERNIE-Image team at Baidu
replicate.com
April 16, 2026 at 3:01 PM
First Prune is halfway 🚀 with something special:

- A recap blog post about our OSS journey: how we started, what we’ve built, and what’s next.
- And a surprise: now each merged PR earns 60 credits.

👉 Read the blog: www.pruna.ai/blog/first-p...
👉 Search for your issue: github.com/PrunaAI/prun...
April 15, 2026 at 10:02 AM
Pruna AI’s optimized fast endpoints for GPT-OSS-20B and 120B are now available. Faster model responses at replicate.com/prunaai/gpt-... and replicate.com/prunaai/gpt-....

Give them a try!

⭐️ Star our OSS: github.com/PrunaAI/pruna
prunaai/gpt-oss-20b-fast – Replicate
Advanced 20B open-weight reasoning models to customize for any use case and run anywhere.
replicate.com
April 14, 2026 at 3:02 PM
Qwen-3.5-35B-A3B-Fast performance up with MoE kernel tuning—throughput doubled and TTFT reduced by 30% compared to competitors. Ready for high-load AI serving!

Details at replicate.com/prunaai/qwen...
replicate.com
April 13, 2026 at 3:02 PM
First Prune is in progress 🌱

If you want to join, there’s still time! Each merged PR in April earns 30 Pruna API credits.

Don't hesitate to contact us if you have questions— we’ll be here to help.

👉 Explore the open issues here: buff.ly/qVPCFdb
⭐️ Star our OSS: github.com/PrunaAI/pruna
April 9, 2026 at 7:03 PM
Models are getting fast. We make them faster. 🚀

We just deployed optimized inference for Google's Gemma 4 26B on @replicate.com, and we've managed to squeeze performance a lot against other deployments:

⚡ +20% throughput
⏱️ -50% time to first token

👉 Try it yourself: replicate.com/prunaai/gemm...
prunaai/gemma-4-26b-a4b-fast – Replicate
This is a version of the MoE Gemma 4 26B optimised by Pruna AI.
replicate.com
April 9, 2026 at 2:20 PM
Want to join First Prune? 🎂

Pick an open issue, submit your PR, and mark it ready for review by April 30.

Each merged PR earns 30 credits redeemable on the Pruna Inference API.

Check them here: github.com/PrunaAI/prun...
April 3, 2026 at 7:01 AM
Ready to make AI more efficient? Our AI Efficiency Courses cover LLM architectures, compression, quantization, and finetuning with hands-on exercises. Perfect for anyone ready to optimize AI beyond just size.

Start learning here: github.com/PrunaAI/ai-e...
GitHub - PrunaAI/ai-efficiency-courses: Courses on building, compressing, evaluating, and deploying efficient AI models.
Courses on building, compressing, evaluating, and deploying efficient AI models. - PrunaAI/ai-efficiency-courses
github.com
April 1, 2026 at 3:26 PM