Sign in

tommymarto.bsky.social

@tommymarto.bsky.social
8 followers 30 following 0 posts
PostsRepliesMedia
Reposted by @tommymarto.bsky.social
Johannes Schusterbauer @joh-schb.bsky.social · 26/05/2026
Diffusion models treat every part of an image equally. → Same number of steps. Same compute. But images aren’t uniform. 🤔 Some regions are easy, others are hard. So why force the model to treat them the same? 🧵
12812
Reposted by @tommymarto.bsky.social
Stefan Baumann @stefanabaumann.bsky.social · 13/04/2026
You don't imagine the future by mentally rendering a movie. You trace how things move -- abstractly, sparsely, step by step. We built a model that does exactly this. It predicts motion, not pixels -- and it's 3,000× faster than video world models. Myriad, accepted at @cvprconference.bsky.social
2259
Reposted by @tommymarto.bsky.social
Stefan Baumann @stefanabaumann.bsky.social · 15/10/2025
🤔 What happens when you poke a scene — and your model has to predict how the world moves in response? We built the Flow Poke Transformer (FPT) to model multi-modal scene dynamics from sparse interactions. It learns to predict the 𝘥𝘪𝘴𝘵𝘳𝘪𝘣𝘶𝘵𝘪𝘰𝘯 of motion itself 🧵👇
1248