Sign in

Stratis Tsirtsis

@stratiss.bsky.social
102 followers 143 following 17 posts

Postdoc @ Hasso Plattner Institute working on machine learning. Previously @ Max Planck Institute, Meta, Stanford, NTUA. 💻 stsirtsis.github.io

PostsRepliesMedia
Stratis Tsirtsis @stratiss.bsky.social · 22/10/2025
What if AI agents aren't here to replace us, but to facilitate our decisions? In a study with 1600 participants, we show that a human with action choices narrowed by an AI makes better sequential decisions than an AI or a human alone. 📜 arxiv.org/abs/2510.16097
111
Reposted by Stratis Tsirtsis
Tobias Gerstenberg @tobigerstenberg.bsky.social · 17/10/2025
The Causality in Cognition Lab at Stanford University is recruiting PhD students this cycle! We are a supportive team who happened to wear bluesky appropriate colors for the lab photo (this wasn't planned). 💙 Lab info: cicl.stanford.edu Application details: psychology.stanford.edu/admissions/p...
0427
Reposted by Stratis Tsirtsis
Yatong Chen @yatongchen.bsky.social · 22/09/2025
We (w/ Moritz Hardt, Olawale Salaudeen and @joavanschoren.bsky.social) are organizing the Workshop on the Science of Benchmarking & Evaluating AI @euripsconf.bsky.social 2025 in Copenhagen! 📢 Call for Posters: rb.gy/kyid4f 📅 Deadline: Oct 10, 2025 (AoE) 🔗 More info: rebrand.ly/bg931sf
1217
Reposted by Stratis Tsirtsis
Mariya Toneva @mtoneva.bsky.social · 04/09/2025
So excited and honored to receive an ERC Starting Grant for the project BrainAlign!! BrainAlign will bring LLMs closer to human understanding by directly aligning them with the human brain. Stay tuned for our findings, and multiple postdoc and PhD openings in the coming years!
4505
Stratis Tsirtsis @stratiss.bsky.social · 20/08/2025
todo: * thesis defense ✅ Grateful to the committee and reviewers Marius Kloft, @arkrause.bsky.social, @rupakmajumdar.bsky.social, and @tobigerstenberg.bsky.social for their time and support. No words are enough to thank my advisor @autreche.bsky.social for everything I’ve learned from him so far 🙏
250
Stratis Tsirtsis @stratiss.bsky.social · 01/08/2025
Last week I had the pleasure of presenting a 2.5-hour tutorial on "Counterfactuals in Minds and Machines" at UAI 2025 in Rio 🇧🇷, prepared together with @autreche.bsky.social and @tobigerstenberg.bsky.social. We've made all materials and references available here: learning.mpi-sws.org/counterfactu...
172
Stratis Tsirtsis @stratiss.bsky.social · 18/07/2025
In Athens 🇬🇷 for the Greeks in AI symposium. Super excited to present our work on "Counterfactual Token Generation in LLMs" (bit.ly/4nMibs2) and see all the amazing work Greek people all over the world are doing on AI! If you are in Athens, let's meet! Next, heading to👇
111
Reposted by Stratis Tsirtsis
uai2026 @auai.org · 04/06/2025
did you check our amazing list of tutorials in Rio? spanning - hyperparameter optimization - counterfactual reasoning - bayesian nonparametrics for causality - causal inference with deep generative models - modern variational inference 👉 www.auai.org/uai2025/tuto...
auai.org
Uncertainty in Artificial Intelligence
0145
Stratis Tsirtsis @stratiss.bsky.social · 30/05/2025
The LLM API you use returns (and charges you for) 5 tokens. Did the LLM actually generate 5 tokens? Or is the provider overcharging you? 🤔 In arxiv.org/abs/2505.21627, led by Ander Artola Velasco, we argue (game-theoretically) for a change from pay-per-token to pay-per-character.
101
Stratis Tsirtsis @stratiss.bsky.social · 28/04/2025
Presenting this today at 17:00 in Hall 4 #6
010
Stratis Tsirtsis @stratiss.bsky.social · 25/04/2025
In Singapore for #ICLR2025! I'll be presenting our work on a causal methodology for evaluating LLMs (arxiv.org/abs/2502.01754) at the "Building Trust in LLMs" workshop on Monday. If you are working on causality, game theory and/or LLMs, let's grab a ☕️ during the conference!
131
Reposted by Stratis Tsirtsis
Manuel Gomez Rodriguez @autreche.bsky.social · 05/02/2025
LLMs rely on randomization to respond to a prompt: they may respond differently to the same prompt if asked multiple times. In “Evaluation of LLMs via Coupled Token Generation” (arxiv.org/abs/2502.01754), we argue that the eval of LLMs should control for this randomization 1/
arxiv.org
Evaluation of Large Language Models via Coupled Token Generation
State of the art large language models rely on randomization to respond to a prompt. As an immediate consequence, a model may respond differently to the same prompt if asked multiple times. In this wo...
172
Stratis Tsirtsis @stratiss.bsky.social · 14/12/2024
Let's talk causality and LLMs! Come find us at the posters in East Hall C. 11:30-12:00 & 14:30-15:00. #neurips2024
010
Stratis Tsirtsis @stratiss.bsky.social · 27/11/2024
What would an LLM have said, counterfactually? Here is a short video illustrating our method for counterfactual token generation. We will present this work at the CaLM workshop at #neurips2024. See you in Vancouver! 📜 arxiv.org/abs/2409.17027 💻 made with manim in python
052
Stratis Tsirtsis @stratiss.bsky.social · 20/11/2024
Hey there 🦋 Let's start with an intro. I'm a final-year PhD student at the Max Planck Institute for Software Systems, working on machine learning, decision making and social aspects of AI. Currently on the academic job market, looking for tenure-track positions👇 💻 stsirtsis.github.io
041