Sign in

Cyrus Rashtchian

@cyroid.bsky.social
314 followers 109 following 13 posts

Researcher at Google. Improving LLM factuality, RAG and multimodal alignment and evaluation. San Diego. he/him ☀️🌱🧗🏻🏐 Prev UCSD, MSR, UW, UIUC.

PostsRepliesMedia
Cyrus Rashtchian @cyroid.bsky.social · 28/04/2025
ICLR 2025 was so much fun!
010
Cyrus Rashtchian @cyroid.bsky.social · 25/04/2025
Curious about fine-grained text-to-image model evaluation? Come see our spotlight paper on Gecko 🦎 in the afternoon poster session at #ICLR25 🏆Hall 3 + Hall 2B #359 🎖️Friday 3pm ICLR: iclr.cc/virtual/2025... Paper: arxiv.org/abs/2404.16820 Prompts: github.com/google-deepm...
iclr.cc
ICLR Poster Revisiting text-to-image evaluation with Gecko: on metrics, prompts, and human ratingICLR 2025
000
Cyrus Rashtchian @cyroid.bsky.social · 25/04/2025
Why do LLMs hallucinate with RAG?! 🤔 Find out at my #ICLR25 poster on Sufficient Context! 👋🏼 📍Hall 3 + Hall 2B #230 ⏰ Fri 25 Apr 10 a.m. to 12:30 p.m.
000
Cyrus Rashtchian @cyroid.bsky.social · 25/04/2025
Happy to chat with anyone at ICLR about RAG, LLMs, Factuality!
000
Reposted by Cyrus Rashtchian
Hailey Joren @haileyjoren.bsky.social · 24/04/2025
When RAG systems hallucinate, is the LLM misusing available information or is the retrieved context insufficient? In our #ICLR2025 paper, we introduce "sufficient context" to disentangle these failure modes. Work w Jianyi Zhang, Chun-Sung Ferng, Da-Cheng Juan, Ankur Taly, @cyroid.bsky.social
1115
Reposted by Cyrus Rashtchian
Hossein Mobahi @thegradient.bsky.social · 20/12/2024
1/2 Just a reminder about Google Research Scholar Program, providing up to $60K unrestricted gifts to recognize early-career professors and support world-class research at institutions around the world. This year, we are particularly interested in the following research areas...
1125
Cyrus Rashtchian @cyroid.bsky.social · 13/12/2024
Longer thread about our new factuality decoding method SLED at NeurIPS 2024. Main idea: freeze the model, but be thoughtful about the decoding. With a small amount of extra inference-time compute, we increase accuracy by 3% on several benchmarks! SLED helps for all major open source models!
120
Cyrus Rashtchian @cyroid.bsky.social · 13/12/2024
First shameless plug -- our new factuality decoding method, SLED gets SOTA improvements on 14+ models (Llama 2/3, Gemma, Mistral) & 9 benchmarks! See our #NeurIPS2024 poster today (Friday) in the East Exhibit Hall A-C #3311
100
Reposted by Cyrus Rashtchian
Inês🩷 @loverines.swifties.social · 02/12/2024
Hi friends!🩷 I have never done this but i’m making a list so and i can keep in touch with all of you more easily🫶🏻 please like this or say hi if i can add you🥰 Thank🫶🏻
22541
Reposted by Cyrus Rashtchian
Pablo Samuel Castro @pcastr.bsky.social · 02/12/2024
Everyone I spoke to at @rl-conference.bsky.social last summer agreed on it being one of the best conferences ever for an RL researcher... So many great RL-focused papers! CFP is out, send your work here!
14513
Cyrus Rashtchian @cyroid.bsky.social · 02/12/2024
Excited to try out bluesky and chat about GenAI and ML theory!
170