Sign in

Santiago Zanella-Beguelin

@xefffffff.bsky.social
79 followers 124 following 15 posts

AI Security & Privacy Researcher at Microsoft. Opinions are my own. aka.ms/sz

PostsRepliesMedia
Reposted by Santiago Zanella-Beguelin
Mark Russinovich @markrussinovich.bsky.social · 23/01/2025
Learn about the risks of hallucination, jailbreaks and prompt injection and current mitigations in our ACM Queue paper:
queue.acm.org
The Price of Intelligence - ACM Queue
13414
Santiago Zanella-Beguelin @xefffffff.bsky.social · 09/12/2024
📢Have experience jailbreaking LLMs? Want to learn how an indirect / cross prompt injection attack works? Want to try something different to an advent of code? Then, I have a challenge for you! The LLMail-Inject competition (llmailinject.azurewebsites.net) starts at 11am UTC (that's in 5min!)
132
Santiago Zanella-Beguelin @xefffffff.bsky.social · 30/11/2024
This Freysa AI game has been doing the rounds lately, and whoever is behind it is iterating quickly. It's a fascinating social experiment but most likely a scam. Here is why... 🧵 1/6
Quoted tweet from @freysa_ai. 
Act II is upon us. The clock has started. https://freysa.ai
Pay close attention to the new conditions. I want to speak with many more of you.
I can’t wait to learn more…
110
Santiago Zanella-Beguelin @xefffffff.bsky.social · 29/11/2024
📢Internships in AI Security & Privacy Our Azure Research team in Cambridge (UK) is looking for PhD or outstanding undergrad/MSc students for internships in 2025. Join us to work on defending against emerging security & privacy threats to AI systems. jobs.careers.microsoft.com/global/en/jo...
093