#llmsteering
90 SAEs on three LLMs gave a modest rank‑correlation (tau‑b ≈ 0.298) between interpretability and steering, and Delta Token Confidence boosted performance by ~52.5%. Read more: https://getnews.me/interpretability-vs-utility-in-sparse-autoencoders-for-llm-steering/ #sparseautoencoders #llmsteering
October 7, 2025 at 5:57 PM
REAL, an inference‑time steering framework, was tested on eight Llama and Qwen LLMs, achieving up to 81.5% gains and an average 20% improvement. Published October 2025. Read more: https://getnews.me/real-framework-boosts-inference-time-steering-of-llms/ #real #llmsteering #ai
October 3, 2025 at 8:21 AM
Researchers use sparse autoencoders to refine LLM steering vectors via denoising and augmentation, improving performance on limited data. Submitted 28 Sep 2025 (arXiv:2509.23799). Read more: https://getnews.me/refining-llm-steering-vectors-with-sparse-autoencoders/ #llmsteering #sparseautoencoder
September 30, 2025 at 2:30 PM