#llmsys
A curated, regularly updated list of academic papers on large language model systems, showing how research has shifted from serving single requests to optimizing whole agent trajectories.

Explore it here:
osp.fyi/llmsys-pape...
September 26, 2026 at 3:30 PM
Poe、phind、llmsys、vercelはもうGPT-4o対応してるんだ
May 14, 2024 at 8:53 AM
oh i see so the scheme was manipulating llmsys not automating prediction market trading
lmao
April 29, 2025 at 3:08 AM
Catching up on this. So glad there's still people at the top of AI like Demis Hassabis who are making sure its application to science is given the air time it deserves and even pushes the latest LLMSys leaderboard out of the news for 24 hrs.

www.youtube.com/watch?v=nQKm...
AI for Science with Sir Paul Nurse, Demis Hassabis, Jennifer Doudna, and John Jumper
YouTube video by Google DeepMind
www.youtube.com
November 25, 2024 at 10:01 AM
May 29, 2026 at 10:39 PM
@alexandr_wang SEAL Leaderboard Update🥇

We added 3 new models to the SEAL Leaderboards:

- GPT-4o-latest (gpt-4o-2024-08-06)
- Gemini 1.5 Pro (Aug 27, 2024) (gemini-1.5-pro-exp-0827)
- Mistral Large 2 (mistral-large-2407)

More detailed results in thread 🧵
#LLM #GenAI #LLMSYS

x.com/alexandr_wan...
x.com
x.com
September 4, 2024 at 8:24 PM
@worldai #llmsys #gpt2 #gpt5 #genai

Introducing GPT-5?

Mysterious GPT2-Chatbot Outperforms GPT-4!

youtu.be/u16ipSeYH7U?...

(Ed : Who did this #OpenAi #MSFT #Apple feels like some used a higher model to train a GPT2 🤔)
Introducing GPT-5? Mysterious GPT2-Chatbot Outperforms GPT-4!
Random Mysterious GPT-2 chatbot popped up on my feed! Out-performs gpt-4 on certain benchmarks and mostly beats every opensoruce model in every category. Spe...
youtu.be
April 29, 2024 at 11:23 PM
Siyu Wu, Yulong Ye, Zezhen Xiang, Pengzhou Chen, Gangda Xiong, Tao Chen: LLMSYS-HPOBench: Hyperparameter Optimization Benchmark Suite for Real-World LLM Systems https://arxiv.org/abs/2605.08305 https://arxiv.org/pdf/2605.08305 https://arxiv.org/html/2605.08305
May 12, 2026 at 6:43 AM