#LLMDrift
How Is ChatGPT’s Behavior Changing over Time?
https://github.com/lchen001/LLMDrift
https://arxiv.org/abs/2307.09009

This states that GPT-4's performance has decreased from March to June 2023. For example, its prime number identification accuracy fell from 97.6% to 2.4%.

Not thoroughly reviewed yet
How is ChatGPT's behavior changing over time?
GPT-3.5 and GPT-4 are the two most widely used large language model (LLM) services. However, when and how these models are updated over time is opaque. Here, we evaluate the March 2023 and June...
arxiv.org
July 19, 2023 at 6:12 PM
You can read the research over here: https://arxiv.org/abs/2307.09009
And the testing dataset is here: https://github.com/lchen001/LLMDrift

Fin/
GitHub - lchen001/LLMDrift
Contribute to lchen001/LLMDrift development by creating an account on GitHub.
github.com
July 20, 2023 at 9:37 AM
A compact linguistic operator I’ve been testing reliably forces #Grok into a low-amplitude, non-simulative response mode:
TONE-μ // SPC:VerbalTemperance // Clamp ∫Affect → TruthSignal // Seal ∮ NoSim
A minimal prompt, but the effect is unmistakable.
#LLMDrift #AffectClamp #SITDOWN #Trickster
November 21, 2025 at 2:38 PM
📦 lchen001 / LLMDrift
⭐ 186 (+51)
🗒 Jupyter Notebook
GitHub - lchen001/LLMDrift
Contribute to lchen001/LLMDrift development by creating an account on GitHub.
github.com
July 22, 2023 at 5:50 PM