🌎EWoK (Elements of World Knowledge)🌎: A cognition-inspired framework for evaluating basic world knowledge in language models
tl;dr: LLMs learn basic social concepts way easier than physical&spatial concepts
Paper: direct.mit.edu/tacl/article...
Website: ewok-core.github.io
🌎EWoK (Elements of World Knowledge)🌎: A cognition-inspired framework for evaluating basic world knowledge in language models
tl;dr: LLMs learn basic social concepts way easier than physical&spatial concepts
Paper: direct.mit.edu/tacl/article...
Website: ewok-core.github.io
(1) No APC. Supports green or diamond OA.
(2) High-quality and relatively fast peer review.
(3) Editors who provide real curational value and exercise purposeful agency in the decision process.
direct.mit.edu/tacl/article...
direct.mit.edu/tacl/article...
If you want to learn how targeting cohesin to defined loci in the genome affects the local chromatin environment and transcription, look no further!
rdcu.be/eLiT5
If you want to learn how targeting cohesin to defined loci in the genome affects the local chromatin environment and transcription, look no further!
rdcu.be/eLiT5
Computational Linguistics
TACL
NEJLT
(1) No APC. Supports green or diamond OA.
(2) High-quality and relatively fast peer review.
(3) Editors who provide real curational value and exercise purposeful agency in the decision process.
Computational Linguistics
TACL
NEJLT
1-0
1-0
IssueBench, our attempt to fix this, is accepted at TACL, and I will be at #EMNLP2025 next week to talk about it!
New results 🧵
We just released IssueBench – the largest, most realistic benchmark of its kind – to answer this question more robustly than ever before.
Long 🧵with spicy results 👇
IssueBench, our attempt to fix this, is accepted at TACL, and I will be at #EMNLP2025 next week to talk about it!
New results 🧵
We present an empirical evaluation and find that language models partially converge towards representations isomorphic to those of vision models. #EMNLP
📃 direct.mit.edu/tacl/article...
We present an empirical evaluation and find that language models partially converge towards representations isomorphic to those of vision models. #EMNLP
📃 direct.mit.edu/tacl/article...
doi.org/10.1162/tacl...
Always a pleasure to work with Abhinav Patil and @jumelet.bsky.social , who deserve the real credit
doi.org/10.1162/tacl...
Always a pleasure to work with Abhinav Patil and @jumelet.bsky.social , who deserve the real credit
Language models (LMs) are remarkably good at generating novel well-formed sentences, leading to claims that they have mastered grammar.
Yet they often assign higher probability to ungrammatical strings than to grammatical strings.
How can both things be true? 🧵👇
Language models (LMs) are remarkably good at generating novel well-formed sentences, leading to claims that they have mastered grammar.
Yet they often assign higher probability to ungrammatical strings than to grammatical strings.
How can both things be true? 🧵👇
🎉 1×ACL Main + 1×TACL
🎥 What Is That Talk About? A Video-to-Text Summarization Dataset for Scientific Presentations (ACL)
🧠 Explanatory Summarization with Discourse-Driven Planning (TACL)
Thanks to my amazing co-authors! 🙏
#LLMs #NLProc #TACL #ACL #ACL2025
👉 aclanthology.org/2025.tacl-1.14/
(2/6)
👉 aclanthology.org/2025.tacl-1.14/
(2/6)
Paper: arxiv.org/abs/2503.03044
Slides/video/poster: underline.io/lecture/1315...
Paper: arxiv.org/abs/2503.03044
Slides/video/poster: underline.io/lecture/1315...
github.com/primeqa/clapnq
arxiv.org/abs/2404.02103
github.com/primeqa/clapnq
arxiv.org/abs/2404.02103
Was telling my 17 year old son who’s been testing the coding waters about how I used/programmed TACL, VFP and SQL, C++, etc (he joked from the 1900’s, mom 😜), and I asked (dared) him to create a simple decision tree. It was hilarious, and now he’s looking at old school tutorials. 🙂
Was telling my 17 year old son who’s been testing the coding waters about how I used/programmed TACL, VFP and SQL, C++, etc (he joked from the 1900’s, mom 😜), and I asked (dared) him to create a simple decision tree. It was hilarious, and now he’s looking at old school tutorials. 🙂
🐡data contains cases where the "bad" response is just as good as chosen one
🐟model rankings can feel off (claude ranks lower than expected)
led by @cmalaviya.bsky.social, we study underspecified queries & detrimental effect on model evals; accepted to TACL 2025
🐡data contains cases where the "bad" response is just as good as chosen one
🐟model rankings can feel off (claude ranks lower than expected)
led by @cmalaviya.bsky.social, we study underspecified queries & detrimental effect on model evals; accepted to TACL 2025
Our #TACL 📄 connects behavior and internals:
💠 LMs amplify toxicity beyond humans
💠 Information about toxicity peaks in lower layers
💠 Bypassing these layers increases toxicity
More details👇 #NLProc #interpretability (1/🧵)
Our #TACL 📄 connects behavior and internals:
💠 LMs amplify toxicity beyond humans
💠 Information about toxicity peaks in lower layers
💠 Bypassing these layers increases toxicity
More details👇 #NLProc #interpretability (1/🧵)
We find that giving language models human-like memory decay *improves* language learning, while, unexpectedly, impairing human reading time prediction
Follow up results soon!
We find that giving language models human-like memory decay *improves* language learning, while, unexpectedly, impairing human reading time prediction
Follow up results soon!