Reposted by Max BartoloLisa Alazraki @lisaalaz.bsky.social · 22/05/2025Thrilled to share our new preprint on Reinforcement Learning for Reverse Engineering (RLRE) 🚀 We demonstrate that human preferences can be reverse engineered effectively by pipelining LLMs to optimise upstream preambles via reinforcement learning 🧵⬇️ 191
Max Bartolo @maxbartolo.bsky.social · 27/03/2025I'm excited to share the tech report for our @cohere.com @cohereforai.bsky.social Command A and Command R7B models. We highlight our novel approach to model training including self-refinement algorithms and model merging techniques at scale. Read more below! ⬇️ 1104
Max Bartolo @maxbartolo.bsky.social · 19/03/2025I really enjoyed my MLST chat with Tim @neuripsconf.bsky.social about the research we've been doing on reasoning, robustness and human feedback. If you have an hour to spare and are interested in AI robustness, it may be worth a listen 🎧 Check it out at youtu.be/DL7qwmWWk88?... 073
Max Bartolo @maxbartolo.bsky.social · 13/02/2025Check out @lisaalaz.bsky.social's internship work with us @cohere.com questioning the rationale behind rationales 🔥 041
Max Bartolo @maxbartolo.bsky.social · 11/12/2024Super excited to see PRISM recognised as a #NeurIPS2024 best paper. This was an incredible large-scale effort by @hannahrosekirk.bsky.social and fantastic collaborators. If you're interested in human feedback, check it out, there are 100+ pages of detailed insights! 🔥 091
Reposted by Max BartoloAdina Williams @adinawilliams.bsky.social · 11/12/2024Our paper PRISM alignment won a best paper award at #neurips2024! All credits to @hannahrosekirk.bsky.social A.Whitefield, P.Röttger, A.M.Bean, K.Margatina, R.Mosquera-Gomez, J.Ciro, @maxbartolo.bsky.social H.He, B.Vidgen, S.Hale Catch Hannah tomorrow at neurips.cc/virtual/2024/poster/97804blog.neurips 2679
Reposted by Max BartoloTim Rocktäschel @handle.invalid · 04/12/2024Excited to reveal Genie 2, our most capable foundation world model that, given a single prompt image, can generate an endless variety of action-controllable, playable 3D worlds. Fantastic cross-team effort by the Open-Endedness Team and many other teams at Google DeepMind! 🧞 39719
Max Bartolo @maxbartolo.bsky.social · 02/12/2024Looking forward to @neuripsconf.bsky.social #NeurIPS #NeurIPS2024 in Vancouver next week! ❄️ Reach out (or pop by the @cohere.com booth) if you want to chat about human feedback, robustness and reasoning, prompt optimisation, adversarial data, glitch tokens, evaluation, or anything else!media.tenor.coman advertisement for vancouver in british columbia canadaALT: an advertisement for vancouver in british columbia canada 0100
Max Bartolo @maxbartolo.bsky.social · 24/11/2024Fun to see Douwe's Dynabench plot continue to inspire new groundbreaking benchmarking work! 040
Max Bartolo @maxbartolo.bsky.social · 20/11/2024🚨 LLMs can learn to reason from procedural knowledge in pretraining data! 🚨 I particularly enjoy research where the evidence contradicts our initial hypothesis. If you're interested in LLM reasoning, check out the 60+ pages of in-depth work at arxiv.org/abs/2411.12580 3677
Reposted by Max Bartoloatla @atla-ai.bsky.social · 19/11/2024We launched Judge Arena with @huggingface.bsky.social @clefourrier.bsky.social - a platform that lets you easily compare models as judges side-by-side and vote for the best evaluation Check out the live leaderboard and start voting now 🤗 0103