Sign in

Dirk Hovy

@dirkhovy.bsky.social
676 followers 337 following 57 posts

Professor @milanlp.bsky.social for #NLProc, compsocsci, #ML Also at dirkhovy.com

PostsRepliesMedia
Reposted by Dirk Hovy
ACL 2027 @aclmeeting.bsky.social · 30/09/2026
🔔CFP for the #ACL2027NLP main conference is up! 2027.aclweb.org/calls/main/ 📝Submission deadline to ARR is Jan 4, 2027.
2027.aclweb.org
Main Conference Papers
ACL 2027 Call for Papers.
075
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 03/09/2026
We are back with #TBT #NLProc Best Paper at NeurIPS D&B 2024: Hanna Rose Kirk, @paul-rottger.bsky.social et al. introduce 'The PRISM Alignment Dataset' arxiv.org/pdf/2404.16019
arxiv.org
062
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 07/09/2026
#MemoryModay #NLProc Outstanding Paper at ACL 2024! @paul-rottger.bsky.social et al. evaluate LLM values and opinions in 'Political Compass or Spinning Arrow?' aclanthology.org/2024.acl-lon...
aclanthology.org
Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models
Paul Röttger, Valentin Hofmann, Valentina Pyatkin, Musashi Hinck, Hannah Kirk, Hinrich Schuetze, Dirk Hovy. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics…
0103
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 10/09/2026
#TBT #NLProc 'Understanding Political Discourse on Twitter through Election Manifestos' Maurer et al. (2024) present a method for predicting party positioning from tweets aclanthology.org/2024.finding...
arxiv.org
https://arxiv.org/pdf/2403.04445
042
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 14/09/2026
#MemoryModay #NLProc #AI #MachineLearning #SafetyFirst 'Safety-Tuned LLaMAs: Improving LLMs Safety' by Bianchi et al. explores training LLMs for safe refusals, warns of over-tuning. arxiv.org/pdf/2309.07875
openreview.net
Verifying your browser | OpenReview
Please complete the verification above.
052
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 17/09/2026
#TBT #NLProc 'Detecting Misogynous Memes with Text & Image Modalities' by Attanasio, @deboranozza.bsky.social, Bianchi. Their novel system uses Perceiver IO, surpassing all previous benchmarks. #AI #ScienceUpdate aclanthology.org/2022.semeval...
aclanthology.org
MilaNLP at SemEval-2022 Task 5: Using Perceiver IO for Detecting Misogynous Memes with Text and Image Modalities
Giuseppe Attanasio, Debora Nozza, Federico Bianchi. Proceedings of the 16th International Workshop on Semantic Evaluation (SemEval-2022). 2022.
052
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 21/09/2026
#MemoryModay #NLProc 'SocioProbe: What, When, and Where Language Models Learn about Sociodemographics' - Lauscher et al. delve into language models' grasp of sociodemographics. Their findings? Models learn, but don't apply it. aclanthology.org/2022.emnlp-m...
aclanthology.org
SocioProbe: What, When, and Where Language Models Learn about Sociodemographics
Anne Lauscher, Federico Bianchi, Samuel R. Bowman, Dirk Hovy. Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing. 2022.
062
Reposted by Dirk Hovy
Arianna Muti @arimuti.bsky.social · 22/09/2026
Our TACL paper "Let Guidelines Guide You: A Prescriptive Guideline-Centered Data Annotation Methodology" is finally out, only six months after acceptance! 😅🎉 direct.mit.edu/tacl/article...
1113
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 24/09/2026
#TBT #NLProc 'Data-Efficient Strategies for Expanding Hate Speech Detection into Under-Resourced Languages' funnels limited target-language data to fine-tune models & enhance effectiveness. By @paul-rottger.bsky.social et al. aclanthology.org/2022.emnlp-m...
aclanthology.org
Data-Efficient Strategies for Expanding Hate Speech Detection into Under-Resourced Languages
Paul Röttger, Debora Nozza, Federico Bianchi, Dirk Hovy. Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing. 2022.
052
Dirk Hovy @dirkhovy.bsky.social · 23/09/2026
Just arrived in lovely León. Tomorrow I’ll give a keynote at SEPLN, the Spanish Society for Natural Language Processing, on some of the most interesting problems ahead for NLP. But first: architecture and tapas! sepln2026.org/cuadro-de-po...
sepln2026.org
Ponentes principales – XLII Congreso Internacional de la Sociedad Española para el Procesamiento del Lenguaje Natural – SEPLN 2026
072
Reposted by Dirk Hovy
Joachim Baumann @joachimbaumann.bsky.social · 04/09/2026
After Science News, our recent ICML paper also got featured in @science.org! Article: www.science.org/content/arti... Paper: arxiv.org/pdf/2605.03202 @dirkhovy.bsky.social @sanmikoyejo.bsky.social #science #PeerReview #research
Screenshot of an article in Science Magazine on the opportunities and risks of AI used for peer review featuring an ICML conference paper by Joachim Baumann et al., 2026.
1104
Reposted by Dirk Hovy
Debora Nozza @deboranozza.bsky.social · 07/08/2026
Very happy that #IC2S22027 will be in Milan at Università Bocconi! I’m really looking forward to welcoming the computational social science community to my university and to Milan, and excited to be part of the team organizing it. See you in 2027! 🇮🇹✨
15212
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 27/07/2026
#MemoryModay #NLProc 'My Answer is C' by Wang et al. (2024) underscores the scrutiny needed for full text responses in LLMs multi-choice evaluations. aclanthology.org/2024.finding...
aclanthology.org
“My Answer is C”: First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models
Xinpeng Wang, Bolei Ma, Chengzhi Hu, Leon Weber-Genzel, Paul Röttger, Frauke Kreuter, Dirk Hovy, Barbara Plank. Findings of the Association for Computational Linguistics: ACL 2024. 2024.
083
Reposted by Dirk Hovy
EACL 2027 @eaclmeeting.bsky.social · 04/08/2026
4134 submissions have been submitted to the ARR August cycle, a 14% increase compared to the ARR October 2025 cycle corresponding to EACL🔥. #EACL2027 in Athens is likely to be the biggest EACL ever 👀 #NLProc @aclrollingreview.bsky.social
01712
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 30/07/2026
#TBT #NLProc "Narratives at Conflict" by Sinelnik and @dirkhovy.bsky.social looks at hidden tactics of disinformation campaigns! They analyzed 8,000 news articles across 4 languages to reveal how disinformation campaigns adapt narratives for different audiences. 🕵️‍♀️ aclanthology.org/2024.acl-srw...
aclanthology.org
Narratives at Conflict: Computational Analysis of News Framing in Multilingual Disinformation Campaigns
Antonina Sinelnik, Dirk Hovy. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 4: Student Research Workshop). 2024.
062
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 23/07/2026
#TBT #NLProc 'Language is Scary when Over-Analyzed...' by @arimuti.bsky.social et al. explores argumentative reasoning in misogyny detection (2024). Detecting implicit misogyny proves challenging for language models. aclanthology.org/2024.emnlp-m...
aclanthology.org
Language is Scary when Over-Analyzed: Unpacking Implied Misogynistic Reasoning with Argumentation Theory-Driven Prompts
Arianna Muti, Federico Ruggeri, Khalid Al Khatib, Alberto Barrón-Cedeño, Tommaso Caselli. Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing. 2024.
032
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 13/07/2026
#MemoryModay #NLProc Fornaciari et al.'s 2022 'Hard and Soft Evaluation of NLP models with BOOtSTrap SAmpling - BooStSa' is a useful Python tool for benchmarking NLP predictions with bootstrapped sampling. aclanthology.org/2022.acl-dem...
aclanthology.org
Hard and Soft Evaluation of NLP models with BOOtSTrap SAmpling - BooStSa
Tommaso Fornaciari, Alexandra Uma, Massimo Poesio, Dirk Hovy. Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics: System Demonstrations. 2022.
043
Reposted by Dirk Hovy
Dallas Card @dallascard.bsky.social · 15/07/2026
Text as Data has been a wonderful, long-running, non-archival workshop for empirical research at the intersection of AI and social science (especially work involving text). After a few years off, it will be happening again this year as a one-day event in early October!
2268
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 13/07/2026
🏆 Congrats to @zeerak.bsky.social and @dirkhovy.bsky.social on the 2016 Test of Time Award! 🎉 "Hateful Symbols or Hateful People? Predictive Features for Hate Speech Detection on Twitter" has shaped research on #hatespeech! Paper: aclanthology.org/N16-2013/ #NLProc
0216
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 18/06/2026
#TBT #NLProc 'Universal Joy: A Data Set and Results for Classifying Emotions Across Languages' by Lamprinidis et al. (2021) explores how AI research affects our planet. Tech can be green too! #SustainableTech aclanthology.org/2021.wassa-1.7
aclanthology.org
Universal Joy A Data Set and Results for Classifying Emotions Across Languages
Sotiris Lamprinidis, Federico Bianchi, Daniel Hardt, Dirk Hovy. Proceedings of the Eleventh Workshop on Computational Approaches to Subjectivity, Sentiment and Social Media Analysis. 2021.
032
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 25/06/2026
#TBT #NLProc 'Exploring challenges in Zero-shot Cross-lingual Hate Speech Detection, @deboranozza.bsky.social (2021) reveals how current models may inaccurately label non-hateful, language-specific interjections as hate speech signals.' aclanthology.org/2021.acl-sho...
aclanthology.org
Exposing the limits of Zero-shot Cross-lingual Hate Speech Detection
Debora Nozza. Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 2: Short…
032
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 29/06/2026
#MemoryMonday #NLProc 'BERTective: Language Models and Contextual Information for Deception Detection' by Fornaciari, T. et al. (2021) explores AI's ability to detect deceit through context. www.aclweb.org/anthology/20...
aclweb.org
BERTective: Language Models and Contextual Information for Deception Detection
Tommaso Fornaciari, Federico Bianchi, Massimo Poesio, Dirk Hovy. Proceedings of the 16th Conference of the European Chapter of the Association for Computational Linguistics: Main Volume. 2021.
042
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 02/07/2026
#TBT #NLProc Hessenthaler et al.'s 2022 work delves into AI's link with fairness & energy reduction in English NLP models, challenging bias reduction theories. #AI #NLP #sustainability aclanthology.org/2022.emnlp-m...
aclanthology.org
Bridging Fairness and Environmental Sustainability in Natural Language Processing
Marius Hessenthaler, Emma Strubell, Dirk Hovy, Anne Lauscher. Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing. 2022.
082
Reposted by Dirk Hovy
Debora Nozza @deboranozza.bsky.social · 02/07/2026
Very happy to have given my first keynote at an #NLProc related conference! This morning at CORIA-TALN 2026 in Nantes, I talked about emerging risks and research directions around the everyday use of LLMs. coria-taln-2026.ls2n.fr/conferences-...
1322
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 06/07/2026
#MemoryMonday #NLProc 'Entropy-based Attention Regularization Frees Unintended Bias Mitigation from Lists' by Attanasio et al. redefines bias reduction in #AI, sans prior term knowledge. #2022Publication aclanthology.org/2022.finding...
aclanthology.org
Entropy-based Attention Regularization Frees Unintended Bias Mitigation from Lists
Giuseppe Attanasio, Debora Nozza, Dirk Hovy, Elena Baralis. Findings of the Association for Computational Linguistics: ACL 2022. 2022.
052
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 09/07/2026
#TBT #NLProc Bergman et al.'s 'Guiding the Release of Safer E2E Conversational AI through Value Sensitive Design' explores AI launch with a value-sensitive lens. aclanthology.org/2022.sigdial...
aclanthology.org
Guiding the Release of Safer E2E Conversational AI through Value Sensitive Design
A. Stevie Bergman, Gavin Abercrombie, Shannon Spruit, Dirk Hovy, Emily Dinan, Y-Lan Boureau, Verena Rieser. Proceedings of the 23rd Annual Meeting of the Special Interest Group on Discourse and…
042
Reposted by Dirk Hovy
Joachim Baumann @joachimbaumann.bsky.social · 10/07/2026
Thank you, @punarpuli.bsky.social, for a great article! Find our paper and more insights from Jiaxin Pei, Sanmi Koyejo, @dirkhovy.bsky.social, and me in this thread👇 bsky.app/profile/joac...
051
Dirk Hovy @dirkhovy.bsky.social · 11/07/2026
Could not be prouder of @zeerak.bsky.social for this accomplishment: from first paper ever (based on an MSc thesis) to 10-year ToT award. We could not have anticipated the lasting impact of this paper, but it's a great honor – and a wonderful story for any student/advisor team (see Zee's thread).
1222
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 15/06/2026
#MemoryMonday #NLProc Uma, A. N. et al. examine AI model training in 'Learning from Disagreement: A Survey'. Disagreement-handling methods' performance is shaped by evaluation methods & dataset traits. www.jair.org/index.php/ja...
jair.org
032
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 12/06/2026
@taniseceron.bsky.social is presenting her work about political content in pre-training and post-training data at the AI & Society conference. #AIandSociety #NLProc
0175
Dirk Hovy @dirkhovy.bsky.social · 12/06/2026
On the heels of a fantastic Dagstuhl seminar on Social Intelligence in AI (thx, @jennhu.bsky.social, @maartensap.bsky.social, @tomerullman.bsky.social, & Lucie Flek), a callback to how we thought about this 5 years ago.
062
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 07/05/2026
#TBT #NLProc 'Predictive Biases in Natural Language Processing Models: A Conceptual Framework and Overview' by Shah, Schwartz, Dirk Hovy (2020). Dive into their mathematical NLP bias framework, its origins, impacts, and types. #AIEthics aclanthology.org/2020.acl-mai...
aclanthology.org
052
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 14/05/2026
#TBT #NLProc Check out 'Helpful or Hierarchical?' by Rashid, F. et al advocating better non-binary representation in tech. #Inclusion
aclanthology.org
Helpful or Hierarchical? Predicting the Communicative Strategies of Chat Participants, and their Impact on Success
Farzana Rashid, Tommaso Fornaciari, Dirk Hovy, Eduardo Blanco, Fernando Vega-Redondo. Findings of the Association for Computational Linguistics: EMNLP 2020. 2020.
032
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 18/05/2026
aclanthology.org/2021.finding...
aclanthology.org
“We will Reduce Taxes” - Identifying Election Pledges with Language Models
Tommaso Fornaciari, Dirk Hovy, Elin Naurin, Julia Runeson, Robert Thomson, Pankaj Adhikari. Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021. 2021.
042
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 18/05/2026
#MemoryModay #NLProc 'We will Reduce Taxes' - Identifying Election Pledges with Language Models' by Fornaciari et al. makes election promise tracking effortless with neural models.
142
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 21/05/2026
#TBT #NLProc Bianchi, Terragni & Dirk Hovy present a smarter technique for topic modeling. Their method uses contextual embeddings for clear, meaningful word clusters, transforming how we interpret large text collections.
aclanthology.org
Pre-training is a Hot Topic: Contextualized Document Embeddings Improve Topic Coherence
Federico Bianchi, Silvia Terragni, Dirk Hovy. Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language…
042
Reposted by Dirk Hovy
Anna Rogers @annarogers.bsky.social · 08/05/2026
I see lots of posts about the automated reviewing at AAAI. Just to be contrarian: here's a counter from @dirkhovy.bsky.social's team: openreview.net/forum?id=cJh...
openreview.net
Stop Automating Peer Review Without Rigorous Evaluation
As AI systems increasingly generate scientific knowledge, the human ability to critically evaluate research becomes more important, not less. Yet large language models offer a tempting solution to...
1144
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 11/05/2026
#MemoryModay #NLProc 'Visualizing Regional Language Variation Across Europe on Twitter' by Dirk Hovy et al. uncovers language differences across Europe with stunning visuals. Language is art! #Linguistics link.springer.com/referencewor...
link.springer.com
Visualizing Regional Language Variation Across Europe on Twitter
Geotagged Twitter data allows us to investigate correlations of geographic language variation, both at an interlingual and intralingual level. Based on data-driven studies of such relationships, this…
043
Reposted by Dirk Hovy
EACL 2027 @eaclmeeting.bsky.social · 06/05/2026
📢⚠️ IMPORTANT DATE CORRECTION: the ARR deadline for EACL 2027 is Aug 3, 2026 (not Aug 6 as previously announced). EACL is earlier than usual in '27, so this is the only viable ARR cycle! 📚 All areas of CL/NLP + related fields welcome. Full CfP coming soon. #NLProc #EACL2027
01512
Reposted by Dirk Hovy
Joachim Baumann @joachimbaumann.bsky.social · 01/05/2026
Can you boost your AI review scores by asking an LLM to rewrite your paper? Yes! We call it paper laundering Our @icmlconf.bsky.social spotlight paper argues current AI reviewers aren't ready to automate peer review, and outlines what a science of peer review automation should look like 🧵👇 #ICML2026
First page of the ICML 2026 spotlight paper "Stop Automating Peer Review Without Rigorous Evaluation" by Joachim Baumann, Jiaxin Pei, Sanmi Koyejo, and Dirk Hovy (Stanford University and Bocconi University). The abstract argues that today's AI systems should not be used to produce paper reviews, grounded in two empirical findings: a "hivemind effect" where AI reviewers show excessive agreement and reduce perspective diversity, and "paper laundering," where prompting an LLM to rewrite a paper trivially increases AI reviewer scores through stylistic changes rather than scientific improvements. The paper calls for a science of peer review automation rather than wholesale deployment of general-purpose LLMs.
44314
Reposted by Dirk Hovy
ACL Rolling Review (ARR) @aclrollingreview.bsky.social · 16/04/2026
🗓️ The ARR March review deadline is approaching: April 20 AoE. Finishing up your review? Run it through REVAS, a peer review assistant that makes your suggestions more actionable, flags unsupported claims, and grounds your feedback in the paper. 👉 revas.mbzuai.ac.ae
revas.mbzuai.ac.ae
REVAS — AI-Powered Peer Review Feedback for Academics
REVAS analyzes the weakness section of your peer review, scoring each paragraph on actionability, helpfulness, grounding, and verifiability.
034
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 06/04/2026
#MemoryModay #NLProc Uma et al. (2020) highlights 'A Case for Soft Loss Functions' efficacy using soft labels & crowd annotations in AI tasks, outshining top-tier methods.
ojs.aaai.org
053
Reposted by Dirk Hovy
nlpandcss.bsky.social @nlpandcss.bsky.social · 06/04/2026
To accommodate ACL decisions, we are further extending the commitment deadline for pre-reviewed ARR submissions to April 7!
044
Reposted by Dirk Hovy
ACL 2027 @aclmeeting.bsky.social · 06/04/2026
The paper acceptance notifications will be out by the 6th of April, AoE. The PCs are working hard throughout the holiday season to finalize the decisions. Apologies for the delay!
046
Reposted by Dirk Hovy
David Lazer @davidlazer.bsky.social · 06/04/2026
The deadline for submission to the Political Networks conference is this Friday. It's taking place Aug 4-7, in Manchester. sites.google.com/view/confpol...
sites.google.com
2026 Manchester
Application and registration
032
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 02/04/2026
#TBT #NLProc '[MASK]? Making Sense of Language-Specific BERT Models' by @deboranozza.bsky.social, Bianchi & @dirkhovy.bsky.social (2020), explores language-specific vs universal BERT models.
arxiv.org
052
Dirk Hovy @dirkhovy.bsky.social · 31/03/2026
I realized how much DMing is like being a professor/chairing a committee. You: - make a brilliant plan for 2+ hours of fun - prep lots of material - immediately get derailed by questions/arguments/etc. - keep it together to make the most of the time together - end up not using most of the material
161
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 30/03/2026
#MemoryModay #NLProc 'Hey Siri. Ok Google. Alexa: A topic modeling of user reviews for smart speakers,' by Nguyen & @dirkhovy.bsky.social decodes speaker reviews for user preferences using topic models. Domain knowledge needed for market analysis.
aclanthology.org
Hey Siri. Ok Google. Alexa: A topic modeling of user reviews for smart speakers
Hanh Nguyen, Dirk Hovy. Proceedings of the 5th Workshop on Noisy User-generated Text (W-NUT 2019). 2019.
052
Reposted by Dirk Hovy
Alexander Hoyle @alexanderhoyle.bsky.social · 26/03/2026
I wrote a blog post on my experience using AI for slide generation Basic idea: write your lecture notes first, then prompt the LLM to produce corresponding slides in reveal.js (h/t @chenhaotan.bsky.social). I'm picky about my slides but was happy with the results! alexanderhoyle.com/posts/ai-sli...
A slide showing that the posterior is proportional to the likelihood times the prior
4638
Reposted by Dirk Hovy
MilaNLP Lab @milanlp.bsky.social · 26/03/2026
#TBT #NLProc Fornaciari, @dirkhovy.bsky.social's 'Identifying Linguistic Areas for Geolocation' explores using social media writing for geolocation via Point-to-City (P2C).
aclanthology.org
Identifying Linguistic Areas for Geolocation
Tommaso Fornaciari, Dirk Hovy. Proceedings of the 5th Workshop on Noisy User-generated Text (W-NUT 2019). 2019.
042