Sign in

Noam Dahan

@dahannoam.bsky.social
175 followers 459 following 2 posts

Research fellow @ MPI-SWS | CS MSc @nlphuji.bsky.social| Researching NLP | Former news editor @ Haaretz | edahanoam.github.io | Looking for PhD 26fall

PostsRepliesMedia
Noam Dahan @dahannoam.bsky.social · 22/12/2025
Life update: I've joined Max Planck Institute for Software Systems as a research fellow (pre-phd), working with @lasha.bsky.social on factuality and nuanced forms of misinformation. Cheers from Germany!
060
Reposted by Noam Dahan
Daria Lioubashevski @darialioub.bsky.social · 11/11/2025
🚨 New preprint! One idea, many ways to say it – but does your brain track those options while you speak? Using LLMs, we put this to the test. www.biorxiv.org/content/10.1... We show for the 1st time that the brain represents multiple alternatives simultaneously in both listening and speaking. 🧵
1155
Reposted by Noam Dahan
Gabi Stanovsky @gabistanovsky.bsky.social · 03/02/2025
There's a lot of talk about regulating AI, but do regulators know the technology well enough? In our new paper, we survey major reg efforts & find they rely on benchmarking, which we know to be problematic. How did this happen & what can we do about it? arxiv.org/pdf/2501.15693
102
Reposted by Noam Dahan
asaf-yehudai.bsky.social @asaf-yehudai.bsky.social · 13/12/2024
New preprint! ✨ Interested in LLM-as-a-Judge? Want to get the best judge for ranking your system? our new work is just for you: "JuStRank: Benchmarking LLM Judges for System Ranking" 🕺💃 arxiv.org/abs/2412.09569
arxiv.org
JuStRank: Benchmarking LLM Judges for System Ranking
Given the rapid progress of generative AI, there is a pressing need to systematically compare and choose between the numerous models and configurations available. The scale and versatility of such eva...
195