When InstructGPT came out in early 2022, Bender/Mitchell/Gebru coauthor a piece that says "Our critiques applied to pretrained models, but instruction tuning creates the 'communicative situation' we said was lacking. All bets are off now."
When InstructGPT came out in early 2022, Bender/Mitchell/Gebru coauthor a piece that says "Our critiques applied to pretrained models, but instruction tuning creates the 'communicative situation' we said was lacking. All bets are off now."
YouTube: www.youtube.com/watch?v=sbXE...
YouTube: www.youtube.com/watch?v=sbXE...
For some reason, he says it often!
For some reason, he says it often!
(With @typesafeai.bsky.social)
(With @typesafeai.bsky.social)
* Early finetunes on InstructGPT and ChatGPT use emdashes their datasets, to "write properly". So do those for Llama and Claude.
* Developers on Huggingface create a wealth of finetuning datasets distilled from asking questions of preexisting models. Since these models use emdashes, ...
* Early finetunes on InstructGPT and ChatGPT use emdashes their datasets, to "write properly". So do those for Llama and Claude.
* Developers on Huggingface create a wealth of finetuning datasets distilled from asking questions of preexisting models. Since these models use emdashes, ...
www.freecodecamp.org/news/ai-pape...
www.freecodecamp.org/news/ai-pape...
It must have been written earlier in 2022, and we're in a different geological era now.
It must have been written earlier in 2022, and we're in a different geological era now.
Key takeaway: Size isn’t everything.
Alignment > Scaling.
By fine-tuning with human feedback, InstructGPT shows we can get better, safer AI without endlessly chasing bigger models.
Linked to Paper 👉 https://buff.ly/3Z2e0v3
Key takeaway: Size isn’t everything.
Alignment > Scaling.
By fine-tuning with human feedback, InstructGPT shows we can get better, safer AI without endlessly chasing bigger models.
Linked to Paper 👉 https://buff.ly/3Z2e0v3
It listens, understands, and doesn’t randomly hallucinate facts about frogs eating socks. 🧦🐸
Huge leap for alignment and responsible AI.
What do you want YOUR AI to do better? Let’s discuss! 💡
It listens, understands, and doesn’t randomly hallucinate facts about frogs eating socks. 🧦🐸
Huge leap for alignment and responsible AI.
What do you want YOUR AI to do better? Let’s discuss! 💡
Or is something different meant in this case?
Or is something different meant in this case?