mlc.ai/modern-gpu-p...
mlc.ai/modern-gpu-p...
"First understand the GPU hardware, then learn the programming model we will use, and finally build state-of-the-art kernels step by step. Our main target is the Blackwell generation, and our main running examples are
"First understand the GPU hardware, then learn the programming model we will use, and finally build state-of-the-art kernels step by step. Our main target is the Blackwell generation, and our main running examples are
CFP is now out: 2025.nesyconf.org/call-for-pap...
🚨 Paper deadline: Feb 28 (abstract), March 7 (full)
#neurosymbolic #NeSy2025
github.com/AIoT-MLSys-L...
github.com/AIoT-MLSys-L...
Mines an LLM's past attention into reusable document scores, rendering compressed evidence with far less overhead than per-query compressors.
📝 arxiv.org/abs/2609.11209
👨🏽💻 github.com/UIUC-MLSys/R...
Mines an LLM's past attention into reusable document scores, rendering compressed evidence with far less overhead than per-query compressors.
📝 arxiv.org/abs/2609.11209
👨🏽💻 github.com/UIUC-MLSys/R...
30% faster LLM prefill—no model changes needed. LAPS just routes prompts smarter.
#AI #AIResearch #MachineLearning
https://autonainews.com/new-laps-system-cuts-llm-prefill-latency-over-30-with-load-aware-deflection/
30% faster LLM prefill—no model changes needed. LAPS just routes prompts smarter.
#AI #AIResearch #MachineLearning
https://autonainews.com/new-laps-system-cuts-llm-prefill-latency-over-30-with-load-aware-deflection/
📌 MLSys'25 top scores (5/5/5/4) - battle-tested at scale
📄 Paper: arxiv.org/abs/2502.19811
📦 Code: github.com/bytedance/fl...
📌 MLSys'25 top scores (5/5/5/4) - battle-tested at scale
📄 Paper: arxiv.org/abs/2502.19811
📦 Code: github.com/bytedance/fl...
Read more: https://ow.ly/pPUi50Zunxi
#UCDavisEngineering
Read more: https://ow.ly/pPUi50Zunxi
#UCDavisEngineering
boredmle.blogspot.com/2026/08/adva...
boredmle.blogspot.com/2026/08/adva...
Origin | Interest | Match
#gpu #sparse-attention #llm #machine-learning #deepseek
Origin | Interest | Match
#gpu #sparse-attention #llm #machine-learning #deepseek
Origin | Interest | Match
https://www.hpcwire.com/off-the-wire/uc-san-diego-packs-a-punch-of-ai-research-power-with-a-gift-from-nvidia/
Result Details
https://www.hpcwire.com/off-the-wire/uc-san-diego-packs-a-punch-of-ai-research-power-with-a-gift-from-nvidia/
Result Details
Discussion | hackernews | Author: crowwork
Discussion | hackernews | Author: crowwork