Sign in

Thaddäus Wiedemer

@thwiedemer.bsky.social
54 followers 107 following 22 posts

Research Scientist at Google Deepmind Zürich | PhD student in ML at Max Planck Institute Tübingen and University of Tübingen.

PostsRepliesMedia
Thaddäus Wiedemer @thwiedemer.bsky.social · 30/07/2026
When an LLM is underperforming, we tweak the prompt. Why not do the same for video models solving visual reasoning tasks? Visual prompt engineering changes the start frame appearance but not the task, improving performance. 📄 arxiv.org/pdf/2607.25537 🌐 visual-prompt-engineering.github.io
100
Thaddäus Wiedemer @thwiedemer.bsky.social · 03/02/2026
How useful are self-generated 'mental images' (visual aids) in MLLM/UMM reasoning? Turns out: currently not very. Visualizations have small errors that compound in multi-step problems, and models often ignore correct visual aids in their decision making.
121
Reposted by Thaddäus Wiedemer
Wieland Brendel @wielandbrendel.bsky.social · 15/12/2025
🚀 We're hiring! The @ellisinsttue.bsky.social leads the AI development for Germany’s new open-source nationwide Adaptive Intelligent System learning platform for schools (as part of a consortium led by Assecor & KI macht Schule, and mandated by the FWU). 👉 Apply now: forms.gle/XmLkwEDD45fY...
153
Reposted by Thaddäus Wiedemer
A. Sophia Koepke @askoepke.bsky.social · 21/10/2025
🎉 Excited to present our paper VGGSounder: Audio‑Visual Evaluations for Foundation Models today at #ICCV2025! 🕦 Poster Session 1 | 11:30–13:30 📍 Poster #88 Come by if you're into audio-visual learning and want to know whether multiple modalities actually help or hurt.
161
Thaddäus Wiedemer @thwiedemer.bsky.social · 25/09/2025
Are we experiencing a 'GPT moment' in vision? In our new preprint, we show that generative video models can solve a wide range of tasks across the entire vision stack without being explicitly trained for it. 🌐 video-zero-shot.github.io 1/n
241
Thaddäus Wiedemer @thwiedemer.bsky.social · 18/02/2025
Check out our newest paper! As always, it was super fun working on this with @prasannamayil.bsky.social
051
Reposted by Thaddäus Wiedemer
Andreas Hochlehnert @ahochlehnert.bsky.social · 17/02/2025
CuratedThoughts: Data Curation for RL Datasets 🚀 Since DeepSeek-R1 introduced reasoning-based RL, datasets like Open-R1 & OpenThoughts emerged for fine-tuning & GRPO. Our deep dive found major flaws — 25% of OpenThoughts needed elimination by data curation. Here's why 👇🧵
1139
Reposted by Thaddäus Wiedemer
Wieland Brendel @wielandbrendel.bsky.social · 11/02/2025
🚀 We’re hiring! Join Bernhard Schölkopf & me at @ellisinsttue.bsky.social to push the frontier of #AI in education! We’re building cutting-edge, open-source AI tutoring models for high-quality, adaptive learning for all pupils with support from the Hector Foundation. 👉 forms.gle/sxvXbJhZSccr...
Hiring announcement: ELLIS Institute Tübingen is looking for ML Researchers & Engineers for Open-Source AI Tutoring (m/f/d). The image features a white background with bold black text and the colorful ELLIS logo at the bottom.
1914