J Rosser @jrosseruk.bsky.social · 27/06/2025Quiz time: I shrank my pretrained LLM's VRAM usage from 1TB+ to 50GB with zero drop in performance. What did I do? 1) 🪄Quantization🪄 2) ⚡FlashAttention⚡ 3) 🏁Gradient checkpointing🏁 4) 💾KV cache to CPU💾 020
J Rosser @jrosseruk.bsky.social · 17/06/2025🎉It's happened! AgentBreeder got its first citation!🎉 arxiv.org/abs/2506.04572 000
J Rosser @jrosseruk.bsky.social · 12/06/2025Excited to share that I've joined @spotify.bsky.social as a Research Scientist PhD Intern this summer investigating Mechanistic Interpretability for long context reasoning in LLMs! #Spotify #InternAtSpotify #LifeAtSpotify 000
J Rosser @jrosseruk.bsky.social · 12/06/2025Had an awesome time presenting mine and @jfoerst.bsky.social's "AgentBreeder" paper at #ICLR2025 to a room of over a hundred researchers! Our approach addresses pre-deployment Al safety risks in multi-agent systems by evolving thousands of "scaffolds" from base LLMs in a multi-objective setting. 120