direct.mit.edu/tacl/article...
direct.mit.edu/tacl/article...
🌎EWoK (Elements of World Knowledge)🌎: A cognition-inspired framework for evaluating basic world knowledge in language models
tl;dr: LLMs learn basic social concepts way easier than physical&spatial concepts
Paper: direct.mit.edu/tacl/article...
Website: ewok-core.github.io
🌎EWoK (Elements of World Knowledge)🌎: A cognition-inspired framework for evaluating basic world knowledge in language models
tl;dr: LLMs learn basic social concepts way easier than physical&spatial concepts
Paper: direct.mit.edu/tacl/article...
Website: ewok-core.github.io
(1) No APC. Supports green or diamond OA.
(2) High-quality and relatively fast peer review.
(3) Editors who provide real curational value and exercise purposeful agency in the decision process.
If you want to learn how targeting cohesin to defined loci in the genome affects the local chromatin environment and transcription, look no further!
rdcu.be/eLiT5
If you want to learn how targeting cohesin to defined loci in the genome affects the local chromatin environment and transcription, look no further!
rdcu.be/eLiT5
Computational Linguistics
TACL
NEJLT
(1) No APC. Supports green or diamond OA.
(2) High-quality and relatively fast peer review.
(3) Editors who provide real curational value and exercise purposeful agency in the decision process.
Computational Linguistics
TACL
NEJLT
IssueBench, our attempt to fix this, is accepted at TACL, and I will be at #EMNLP2025 next week to talk about it!
New results 🧵
We just released IssueBench – the largest, most realistic benchmark of its kind – to answer this question more robustly than ever before.
Long 🧵with spicy results 👇
IssueBench, our attempt to fix this, is accepted at TACL, and I will be at #EMNLP2025 next week to talk about it!
New results 🧵
We present an empirical evaluation and find that language models partially converge towards representations isomorphic to those of vision models. #EMNLP
📃 direct.mit.edu/tacl/article...
We present an empirical evaluation and find that language models partially converge towards representations isomorphic to those of vision models. #EMNLP
📃 direct.mit.edu/tacl/article...
doi.org/10.1162/tacl...
Always a pleasure to work with Abhinav Patil and @jumelet.bsky.social , who deserve the real credit
doi.org/10.1162/tacl...
Always a pleasure to work with Abhinav Patil and @jumelet.bsky.social , who deserve the real credit
Language models (LMs) are remarkably good at generating novel well-formed sentences, leading to claims that they have mastered grammar.
Yet they often assign higher probability to ungrammatical strings than to grammatical strings.
How can both things be true? 🧵👇
Language models (LMs) are remarkably good at generating novel well-formed sentences, leading to claims that they have mastered grammar.
Yet they often assign higher probability to ungrammatical strings than to grammatical strings.
How can both things be true? 🧵👇
🎉 1×ACL Main + 1×TACL
🎥 What Is That Talk About? A Video-to-Text Summarization Dataset for Scientific Presentations (ACL)
🧠 Explanatory Summarization with Discourse-Driven Planning (TACL)
Thanks to my amazing co-authors! 🙏
#LLMs #NLProc #TACL #ACL #ACL2025
👉 aclanthology.org/2025.tacl-1.14/
(2/6)
👉 aclanthology.org/2025.tacl-1.14/
(2/6)
Paper: arxiv.org/abs/2503.03044
Slides/video/poster: underline.io/lecture/1315...
Paper: arxiv.org/abs/2503.03044
Slides/video/poster: underline.io/lecture/1315...
github.com/primeqa/clapnq
arxiv.org/abs/2404.02103
github.com/primeqa/clapnq
arxiv.org/abs/2404.02103
Was telling my 17 year old son who’s been testing the coding waters about how I used/programmed TACL, VFP and SQL, C++, etc (he joked from the 1900’s, mom 😜), and I asked (dared) him to create a simple decision tree. It was hilarious, and now he’s looking at old school tutorials. 🙂
Was telling my 17 year old son who’s been testing the coding waters about how I used/programmed TACL, VFP and SQL, C++, etc (he joked from the 1900’s, mom 😜), and I asked (dared) him to create a simple decision tree. It was hilarious, and now he’s looking at old school tutorials. 🙂
🐡data contains cases where the "bad" response is just as good as chosen one
🐟model rankings can feel off (claude ranks lower than expected)
led by @cmalaviya.bsky.social, we study underspecified queries & detrimental effect on model evals; accepted to TACL 2025
🐡data contains cases where the "bad" response is just as good as chosen one
🐟model rankings can feel off (claude ranks lower than expected)
led by @cmalaviya.bsky.social, we study underspecified queries & detrimental effect on model evals; accepted to TACL 2025
Our #TACL 📄 connects behavior and internals:
💠 LMs amplify toxicity beyond humans
💠 Information about toxicity peaks in lower layers
💠 Bypassing these layers increases toxicity
More details👇 #NLProc #interpretability (1/🧵)
Our #TACL 📄 connects behavior and internals:
💠 LMs amplify toxicity beyond humans
💠 Information about toxicity peaks in lower layers
💠 Bypassing these layers increases toxicity
More details👇 #NLProc #interpretability (1/🧵)
We find that giving language models human-like memory decay *improves* language learning, while, unexpectedly, impairing human reading time prediction
Follow up results soon!
We find that giving language models human-like memory decay *improves* language learning, while, unexpectedly, impairing human reading time prediction
Follow up results soon!