#flashmoe
🤞 flashmoe support coming to llama.cpp?
June 13, 2026 at 6:26 AM
A 397B LLM on a Mac? Flash-Moe makes it happen, pushing edge compute limits. See how AI helped build this groundbreaking system!

https://thepixelspulse.com/posts/flash-moe-on-mac-397b-model-reality/

#flashmoe #danwoods #applesilicon
March 22, 2026 at 2:18 PM
The iPhone 17 Pro just ran a 400B LLM! Forget cloud AI – this demo proves massive, private AI models can live on your phone, thanks to a clever SSD streaming trick.

https://thepixelspulse.com/posts/iphone-17-pro-llm-on-device-ai-breakthrough/

#iphone17pro #flashmoe #a19pro
March 23, 2026 at 4:06 PM
wouldn't you still be extremely limited by the memory bandwidth if it doesn't fit cleanly into VRAM?

the M2 series could be ran on 128GBs pretty comfortably, but for M3 you would need at least 256GBs or SSD streaming, and flashmoe won't help with either of those afaik
June 13, 2026 at 6:34 AM
Alibaba Cloud AI Potential Conference: MoE Models Rise, AI Infrastructure Booms
If computing power becomes as ubiquitous as electricity, can Alibaba Cloud become the "power grid" of the AI era? #CloudComputing #AIInfrastructure
#Alibaba
aidisruption.ai/p/alibaba-cl...
Alibaba Cloud AI Potential Conference: MoE Models Rise, AI Infrastructure Booms
Alibaba Cloud advances AI infrastructure with FlashMoE, optimized clusters & In-DB AI, boosting efficiency for next-gen AI models.
aidisruption.ai
April 10, 2025 at 1:22 PM
397B AI model on a laptop? Flash-MoE makes it happen on Apple Silicon, but the real breakthrough isn't just scale—it's about *useful* quality for agentic tasks.

https://thepixelspulse.com/posts/flash-moe-laptop-model/

#flashmoe #applesilicon #qwen35397ba17b
March 22, 2026 at 6:11 PM