Sign in

Max Bartolo

@maxbartolo.bsky.social
306 followers 27 following 15 posts

Building robust LLMs @Cohere

PostsRepliesMedia
Reposted by Max Bartolo
Lisa Alazraki @lisaalaz.bsky.social · 22/05/2025
Thrilled to share our new preprint on Reinforcement Learning for Reverse Engineering (RLRE) 🚀 We demonstrate that human preferences can be reverse engineered effectively by pipelining LLMs to optimise upstream preambles via reinforcement learning 🧵⬇️
191
Max Bartolo @maxbartolo.bsky.social · 27/03/2025
I'm excited to share the tech report for our @cohere.com @cohereforai.bsky.social Command A and Command R7B models. We highlight our novel approach to model training including self-refinement algorithms and model merging techniques at scale. Read more below! ⬇️
1104
Max Bartolo @maxbartolo.bsky.social · 19/03/2025
I really enjoyed my MLST chat with Tim @neuripsconf.bsky.social about the research we've been doing on reasoning, robustness and human feedback. If you have an hour to spare and are interested in AI robustness, it may be worth a listen 🎧 Check it out at youtu.be/DL7qwmWWk88?...
073
Max Bartolo @maxbartolo.bsky.social · 13/02/2025
Check out @lisaalaz.bsky.social's internship work with us @cohere.com questioning the rationale behind rationales 🔥
041
Max Bartolo @maxbartolo.bsky.social · 11/12/2024
Super excited to see PRISM recognised as a #NeurIPS2024 best paper. This was an incredible large-scale effort by @hannahrosekirk.bsky.social and fantastic collaborators. If you're interested in human feedback, check it out, there are 100+ pages of detailed insights! 🔥
091
Reposted by Max Bartolo
Adina Williams @adinawilliams.bsky.social · 11/12/2024
Our paper PRISM alignment won a best paper award at #neurips2024! All credits to @hannahrosekirk.bsky.social A.Whitefield, P.Röttger, A.M.Bean, K.Margatina, R.Mosquera-Gomez, J.Ciro, @maxbartolo.bsky.social H.He, B.Vidgen, S.Hale Catch Hannah tomorrow at neurips.cc/virtual/2024/poster/97804
blog.neurips
2679
Reposted by Max Bartolo
Tim Rocktäschel @handle.invalid · 04/12/2024
Excited to reveal Genie 2, our most capable foundation world model that, given a single prompt image, can generate an endless variety of action-controllable, playable 3D worlds. Fantastic cross-team effort by the Open-Endedness Team and many other teams at Google DeepMind! 🧞
39719
Max Bartolo @maxbartolo.bsky.social · 02/12/2024
Looking forward to @neuripsconf.bsky.social #NeurIPS #NeurIPS2024 in Vancouver next week! ❄️ Reach out (or pop by the @cohere.com booth) if you want to chat about human feedback, robustness and reasoning, prompt optimisation, adversarial data, glitch tokens, evaluation, or anything else!
media.tenor.com
an advertisement for vancouver in british columbia canada
ALT: an advertisement for vancouver in british columbia canada
0100
Max Bartolo @maxbartolo.bsky.social · 29/11/2024
Sparks of multi-hop reasoning ✨
082
Max Bartolo @maxbartolo.bsky.social · 24/11/2024
Fun to see Douwe's Dynabench plot continue to inspire new groundbreaking benchmarking work!
040
Max Bartolo @maxbartolo.bsky.social · 20/11/2024
🚨 LLMs can learn to reason from procedural knowledge in pretraining data! 🚨 I particularly enjoy research where the evidence contradicts our initial hypothesis. If you're interested in LLM reasoning, check out the 60+ pages of in-depth work at arxiv.org/abs/2411.12580
3677
Reposted by Max Bartolo
atla @atla-ai.bsky.social · 19/11/2024
We launched Judge Arena with @huggingface.bsky.social @clefourrier.bsky.social - a platform that lets you easily compare models as judges side-by-side and vote for the best evaluation Check out the live leaderboard and start voting now 🤗
0103