#TextAsData/#NLP
🚨 #Data alert 📣 We just released #ParlLawSpeech – full texts of more than 40k bills, 28k laws, and 3 mio. parliamentary speeches from 7 countries (AT, CZ, DE, DK, ES, HR, HU) and the EU! If you study democracy with #TextAsData / #NLP methods, this is for you! A short 🧵 (1/3) #PoliSkyData #polisky
February 13, 2025 at 5:29 PM
📡The next term of the #TaDa Speaker Series starts next week!! We have six amazing speakers on all things #TextAsData/#NLP lined up for you

👇
October 11, 2023 at 8:49 AM
🤔 The #EU 🇪🇺 aspires to be a global actor 🌍 — but do other states recognize it as such?

My new study in @intlinteractions.bsky.social develops targeted #NLP / #TextAsData tools to analyze 50 years of foreign policy discourse in the annual #UN General Debate (1970–2020).

A thread (1/n)
July 16, 2025 at 7:57 AM
If you feel uneasy using LLMs for data annotation, you are right (if not, you should). It offers new chances for research that is difficult with traditional #NLP/#textasdata methods, but the risk of false conclusions is high!

Experiment + *evidence-based* mitigation strategies in this preprint 👇
🚨 New paper alert 🚨 Using LLMs as data annotators, you can produce any scientific result you want. We call this **LLM Hacking**.

Paper: arxiv.org/pdf/2509.08825
September 15, 2025 at 1:05 PM
If you do research with/about LLMs, image or audio models, or classic NLP methods, you should check out this call 👇

#textasdata #NLP #llm #commsky #polsky
Submission for #COMPTEXT2026 is still open 𝐮𝐧𝐭𝐢𝐥 𝐅𝐫𝐢𝐝𝐚𝐲.
We are looking forward to receiving your proposals for papers and panels.
We also appreciate it if you can spread the word and circulate shorturl.at/gRg0p!
See you in Birmingham, 23–25 April 2026!
January 12, 2026 at 8:54 AM
"Potenziale von #KI für die Personalisierung von (Weiter-)Bildung" - der Vortrag von @afischer1985.bsky.social
ist nun im @zpid.bsky.social Mediacenter frei verfügbar!

👉 zpid.cloud.panopto.eu/Panopto/Page...

#TextAsData #Embeddings #LLM #NLP #Education
April 2, 2024 at 7:54 AM
This scales corruption measurement to the universe of audited municipalities. It's transparent, replicable, & cost-effective (unlike hand-coding or LLMs).
📄 Paper: papers.ssrn.com/sol3/papers....
#TextAsData #Corruption #NLP #MachineLearning #Research #Methodology #OpenScience End/
Measuring Corruption from Text Data
Using Brazilian municipal audit reports, I construct an automated corruption index that combines a dictionary of audit irregularities with principal component a
papers.ssrn.com
December 9, 2025 at 1:12 PM
➡️ Predoctoral Research Fellow / PhD Candidate
👉 75%, 3 years, @wzb.bsky.social
👉 Focus: Quantitative text analysis ( #NLP | #CSS | #TextAsData ) on global crisis perceptions and dissertation research
👉 Full job call: www.wzb.eu/en/jobs/pred...
Predoctoral Research Fellow – PhD candidate (f/m/x) (ID 364)
For its VARICRIS project at the research department Global Governance, the WZB Berlin Social Science Center is looking for A Predoctoral Research Fellow – PhD candidate (f/m/x) to be employed...
www.wzb.eu
October 22, 2025 at 2:40 PM
🚨 LLMs+research=beware of #lookahead #bias 🚨

On Jan 16, 2026, join the #AMLEDS webinar w/ @amanela.bsky.social (WashU) on Lookahead Bias with #LLMs!

How do we do credible #forecasting or #TextAsData analysis when models have seen the future?

🕚 11am ET-5pm CET

#AI #EconTwitter #EconSky #NLP #ML
January 13, 2026 at 2:45 PM
➡️ Predoctoral Research Fellow / PhD Candidate
👉 75%, 3 years, @WZB Berlin
👉 Focus: Quantitative text analysis ( #NLP | #CSS | #TextAsData ) on global crisis perceptions and dissertation research
👉 Full job call: www.wzb.eu/en/jobs/pred...
Predoctoral Research Fellow – PhD candidate (f/m/x) (ID 364)
For its VARICRIS project at the research department Global Governance, the WZB Berlin Social Science Center is looking for A Predoctoral Research Fellow – PhD candidate (f/m/x) to be employed...
www.wzb.eu
October 22, 2025 at 2:06 PM
Suppose you have *3 hours* to tell lawyers about main ideas in #textanalytics, #TextAsData. Structured vs unstructured text, bag of words, stop-words, n-grams, word frequency, creating variables from words, dictionaries, sentiment analysis, NLP. 🎓📄💬
Anything else?
December 2, 2024 at 1:10 AM