Reposted by Maria ValentiniBoulder NLP @bouldernlp.bsky.social · 04/10/2026Boulder NLP is heading to COLM with new work on human-chatbot interaction, translationese, and data mixing for pretraining. See you in San Francisco! 🌉 🚋 🍞 0144
Maria Valentini @mvalentini.bsky.social · 20/07/2026Excited to be presenting our new paper at #CogSci2026 this coming week, which asks: Does Contextual Informativeness Predict Preschoolers’ Word Learning from Stories? I will be in poster session 3 on Friday evening. Full paper will be published in proceedings at the conclusion of the conference! :) 083
Reposted by Maria ValentiniBoulder NLP @bouldernlp.bsky.social · 09/07/2026Congratulations to @mginn.bsky.social, @covetedfish.bsky.social, Ali Marashian, @mvalentini.bsky.social, @alexispalmer.bsky.social & friends for receiving an Outstanding Paper award at #ACL2026 for the paper "Massively Multilingual Joint Segmentation and Glossing". aclanthology.org/2026.acl-lon...aclanthology.orgMassively Multilingual Joint Segmentation and GlossingMichael Ginn, Lindia Tjuatja, Enora Rice, Ali Marashian, Maria Valentini, Jasmine Xu, Graham Neubig, Alexis Palmer. Proceedings of the 64th Annual Meeting of the Association for Computational Linguist... 1173
Reposted by Maria Valentinimichael ginn @mginn.bsky.social · 06/04/2026Excited to announce that the PolyGloss paper has been accepted to @aclmeeting.bsky.social! Previously, we trained models to help in endangered language documentation workflows by automatically predicting interlinear glosses. But real-world user studies revealed crucial issues... 153
Reposted by Maria ValentiniCarl T. Bergstrom @carlbergstrom.com · 14/06/2025This is the third story I've read in a month about how AI chatbots are leading people into psychological crises. Gift linknytimes.comThey Asked an A.I. Chatbot Questions. The Answers Sent Them Spiraling. 18363121
Reposted by Maria ValentiniMark Riedl @markriedl.bsky.social · 21/01/2025I don’t really have the energy for politics right now. So I will observe without comment: Executive Order 14110 was revoked (Safe, Secure, and Trustworthy Development and Use of Artificial Intelligence) 29336
Maria Valentini @mvalentini.bsky.social · 16/01/2025We focus on automatically evaluating contextual informativeness relative to multiple target words in child-directed text, with implications for improving the automatic generation of educational stories for early childhood vocabulary intervention. Can’t wait to share, and learn about others’ work :) 010
Maria Valentini @mvalentini.bsky.social · 16/01/2025Excited to be presenting my work with @teaywright.bsky.social at #COLING2025 next week in Abu Dhabi! Find us in poster session 6/E on Jan 22nd (11 AM in the atrium). Paper: arxiv.org/abs/2412.17427arxiv.orgMeasuring Contextual Informativeness in Child-Directed TextTo address an important gap in creating children's stories for vocabulary enrichment, we investigate the automatic evaluation of how well stories convey the semantics of target vocabulary words, a tas... 1103
Reposted by Maria ValentiniMaria Antoniak @mariaa.bsky.social · 27/11/20241. Can you stop companies from training generative AI using your data? No, not currently. 2. Is this dataset meant for training generative AI? 🤷♀️ but more likely for research and statistical analysis. 3. Is it ok to duplicate and distribute people’s data without agency to opt out? I’d argue no. 2397
Reposted by Maria ValentiniCarl T. Bergstrom @carlbergstrom.com · 17/11/2024So many people, CS researchers included, think that you can explore how an LLM works by simply asking it to tell you what it is doing or "thinking". Here @jennhu.bsky.social provides an excellent illustration of how that approach fails even at the most basic level. 1836880