Daniel Fried @daniel-fried.bsky.social · 03/10/2025I'm excited about Andy's work -- generating scenarios that force LLMs to choose between conflicting values, allowing us to see which values they prioritize. Might be used for training in the future! We also show the importance of open-ended (vs multiple choice) evaluation. 040
Reposted by Daniel FriedIVADO @ivado.bsky.social · 19/08/2025The IVADO #Bootcamp marked the launch of the Thematic Semester on Autonomous #LLM Agents last week at the MIL Campus of @umontreal.ca. Over 4 days, researchers, experts, and #AI enthusiasts gathered for conferences, tutorials, and rich discussions, laying the groundwork for our next two workshops. 112
Daniel Fried @daniel-fried.bsky.social · 16/07/2025Excited to be presenting two of our papers at #ICML2025 and workshops, today through Saturday! Topics are memory for agents, and constructing coding environments for training & evaluation. See links below: 110
Reposted by Daniel FriedRobert Hawkins @rdhawkins.bsky.social · 28/05/2025Happy to announce the first workshop on Pragmatic Reasoning in Language Models — PragLM @ COLM 2025! 🎉 How do LLMs engage in pragmatic reasoning, and what core pragmatic capacities remain beyond their reach? 🌐 sites.google.com/berkeley.edu/praglm/ 📅 Submit by June 23rdsites.google.comPragLM @ COLM '25IMPORTANT DATES 13918
Reposted by Daniel FriedLucy Li @lucy3.bsky.social · 05/05/2025I'm joining Wisconsin CS as an assistant professor in fall 2026!! There, I'll continue working on language models, computational social science, & responsible AI. 🌲🧀🚣🏻♀️ Apply to be my PhD student! Before then, I'll postdoc for a year in the NLP group at another UW 🏔️ in the Pacific Northwest 1614514
Reposted by Daniel FriedChris Donahue @chrisdonahue.com · 05/03/2025Inaugurating new acct to share work from my PhD student! Wayne et al have been running a live eval platform Copilot Arena - a VSCode extension serving code completions from AI systems to real developers. See 🧵 for findings and preprint Excited to be evaluating human-AI *workflows* holistically! 0103
Reposted by Daniel FriedPranjal @pranjal2041.bsky.social · 26/02/2025What if AI agents did software engineering like humans—seeing the screen & using any developer tool? Introducing Programming with Pixels: an SWE environment where agents control VSCode via screen perception, typing & clicking to tackle diverse tasks. programmingwithpixels.com 🧵 184
Reposted by Daniel FriedNouha Dziri @nouhadziri.bsky.social · 23/01/2025Interested in knowing more about LLMs agents and in contributing to this topic?🚀 📢We're thrilled to announce REALM: The first Workshop for Research on Agent Language Models 🤖 #ACL2025NLP in Vienna 🎻 We have an exciting lineup of speakers 🗓️ Submit your work by *March 1st* @aclmeeting.bsky.social 1134
Reposted by Daniel FriedKush Jain @kjain14.bsky.social · 19/12/2024Thrilled to announce our new work TestGenEval, a benchmark that measures unit test generation and test completion capabilities. This work was done in collaboration with the FAIR CodeGen team. Preprint: arxiv.org/abs/2410.00752 Leaderboard: testgeneval.github.io/leaderboard.... 1177
Reposted by Daniel FriedMaarten Sap @maartensap.bsky.social · 07/01/2025CMU LTI is hosting predoc interns this summer, centered around "Language Technologies for All"! Please apply and circulate! lti.cs.cmu.edu/news-and-eve...lti.cs.cmu.eduCMU LTI Language Technology for All Internship 2025 - Language Technologies Institute - School of Computer Science - Carnegie Mellon UniversityThe LTI is currently seeking applicants for the summer 2025 Language Technology for All Internship 1198
Reposted by Daniel FriedConference on Language Modeling @colmweb.org · 17/12/2024Announcement #1: our call for papers is up! 🎉 colmweb.org/cfp.html And excited to announce the COLM 2025 program chairs @yoavartzi.com @eunsol.bsky.social @ranjaykrishna.bsky.social and @adtraghunathan.bsky.social 06624