Sign in

Daniel Fried

@daniel-fried.bsky.social
2.3K followers 194 following 10 posts

Assistant prof at LTI CMU. Working on NLP: language interfaces, applied pragmatics, language-to-code, grounding. dpfried.github.io

PostsRepliesMedia
Daniel Fried @daniel-fried.bsky.social · 03/10/2025
I'm excited about Andy's work -- generating scenarios that force LLMs to choose between conflicting values, allowing us to see which values they prioritize. Might be used for training in the future! We also show the importance of open-ended (vs multiple choice) evaluation.
040
Reposted by Daniel Fried
IVADO @ivado.bsky.social · 19/08/2025
The IVADO #Bootcamp marked the launch of the Thematic Semester on Autonomous #LLM Agents last week at the MIL Campus of @umontreal.ca. Over 4 days, researchers, experts, and #AI enthusiasts gathered for conferences, tutorials, and rich discussions, laying the groundwork for our next two workshops.
112
Daniel Fried @daniel-fried.bsky.social · 16/07/2025
Paper: arxiv.org/abs/2503.07358 Code: github.com/yiqingxyq/Re... Work led by Yiqing Xie, with Alex Xie, Divyanshu Sheth, Pengfei Liu, @daniel-fried.bsky.social and @carolynrose.bsky.social
arxiv.org
RepoST: Scalable Repository-Level Coding Environment Construction with Sandbox Testing
We present RepoST, a scalable method to construct environments that provide execution feedback for repository-level code generation for both training and evaluation. Unlike existing works that aim to ...
022
Daniel Fried @daniel-fried.bsky.social · 16/07/2025
2) RepoST. We automatically create executable environments from real GitHub repos, allowing us to train and evaluate models for function generation in real-world contexts. Presenting at the CODEML workshop on Fri Jul 18th. Also accepted to COLM, upcoming!
100
Daniel Fried @daniel-fried.bsky.social · 16/07/2025
Paper: arxiv.org/abs/2409.07429 Code: github.com/zorazrw/agen... Work led by @zorazrw.bsky.social, with Jiayuan Mao and @gneubig.bsky.social
arxiv.org
Agent Workflow Memory
Despite the potential of language model-based agents to solve real-world tasks such as web navigation, current methods still struggle with long-horizon tasks with complex action trajectories. In contr...
100
Daniel Fried @daniel-fried.bsky.social · 16/07/2025
1) Agent Workflow Memory. Allow agents to adapt online to carry out new tasks more accurately by inducing workflows for common sub-tasks. Today (Wed 7/17): 4:30-7pm. West Exhibition Hall B2-B3 W-202): Also at the CUA workshop, morning of Sat 7/19.
100
Daniel Fried @daniel-fried.bsky.social · 16/07/2025
Excited to be presenting two of our papers at #ICML2025 and workshops, today through Saturday! Topics are memory for agents, and constructing coding environments for training & evaluation. See links below:
110
Reposted by Daniel Fried
Robert Hawkins @rdhawkins.bsky.social · 28/05/2025
Happy to announce the first workshop on Pragmatic Reasoning in Language Models — PragLM @ COLM 2025! 🎉 How do LLMs engage in pragmatic reasoning, and what core pragmatic capacities remain beyond their reach? 🌐 sites.google.com/berkeley.edu/praglm/ 📅 Submit by June 23rd
sites.google.com
PragLM @ COLM '25
IMPORTANT DATES
13918
Daniel Fried @daniel-fried.bsky.social · 10/05/2025
Congrats Lucy!!
040
Reposted by Daniel Fried
Lucy Li @lucy3.bsky.social · 05/05/2025
I'm joining Wisconsin CS as an assistant professor in fall 2026!! There, I'll continue working on language models, computational social science, & responsible AI. 🌲🧀🚣🏻‍♀️ Apply to be my PhD student! Before then, I'll postdoc for a year in the NLP group at another UW 🏔️ in the Pacific Northwest
Wisconsin-Madison's tree-filled campus, next to a big shiny lake A computer render of the interior of the new computer science, information science, and statistics building. A staircase crosses an open atrium with visibility across multiple floors
1614514
Reposted by Daniel Fried
Chris Donahue @chrisdonahue.com · 05/03/2025
Inaugurating new acct to share work from my PhD student! Wayne et al have been running a live eval platform Copilot Arena - a VSCode extension serving code completions from AI systems to real developers. See 🧵 for findings and preprint Excited to be evaluating human-AI *workflows* holistically!
0103
Reposted by Daniel Fried
Pranjal @pranjal2041.bsky.social · 26/02/2025
What if AI agents did software engineering like humans—seeing the screen & using any developer tool? Introducing Programming with Pixels: an SWE environment where agents control VSCode via screen perception, typing & clicking to tackle diverse tasks. programmingwithpixels.com 🧵
184
Reposted by Daniel Fried
Nouha Dziri @nouhadziri.bsky.social · 23/01/2025
Interested in knowing more about LLMs agents and in contributing to this topic?🚀 📢We're thrilled to announce REALM: The first Workshop for Research on Agent Language Models 🤖 #ACL2025NLP in Vienna 🎻 We have an exciting lineup of speakers 🗓️ Submit your work by *March 1st* @aclmeeting.bsky.social
1134
Daniel Fried @daniel-fried.bsky.social · 15/01/2025
Congrats Mohit!!
160
Reposted by Daniel Fried
Kush Jain @kjain14.bsky.social · 19/12/2024
Thrilled to announce our new work TestGenEval, a benchmark that measures unit test generation and test completion capabilities. This work was done in collaboration with the FAIR CodeGen team. Preprint: arxiv.org/abs/2410.00752 Leaderboard: testgeneval.github.io/leaderboard....
1177
Reposted by Daniel Fried
Maarten Sap @maartensap.bsky.social · 07/01/2025
CMU LTI is hosting predoc interns this summer, centered around "Language Technologies for All"! Please apply and circulate! lti.cs.cmu.edu/news-and-eve...
lti.cs.cmu.edu
CMU LTI Language Technology for All Internship 2025 - Language Technologies Institute - School of Computer Science - Carnegie Mellon University
The LTI is currently seeking applicants for the summer 2025 Language Technology for All Internship
1198
Daniel Fried @daniel-fried.bsky.social · 06/01/2025
You can execute each generated function on a set of possible inputs to the function, group the functions according to the outputs, then choose the largest group: arxiv.org/abs/2204.11454 and Sec 4.6 of arxiv.org/abs/2203.07814, although I'm not sure what was done in these plots
arxiv.org
Natural Language to Code Translation with Execution
Generative models of code, pretrained on large corpora of programs, have shown great success in translating natural language to code (Chen et al., 2021; Austin et al., 2021; Li et al., 2022, inter ali...
070
Daniel Fried @daniel-fried.bsky.social · 02/01/2025
So sorry to hear this, what a loss - such a kind and fun guy and his work is so creative.
010
Reposted by Daniel Fried
Conference on Language Modeling @colmweb.org · 17/12/2024
Announcement #1: our call for papers is up! 🎉 colmweb.org/cfp.html And excited to announce the COLM 2025 program chairs @yoavartzi.com @eunsol.bsky.social @ranjaykrishna.bsky.social and @adtraghunathan.bsky.social
06624