Sign in

Gabriele Goletto

@gabrigole.bsky.social
542 followers 285 following 4 posts

Research Scientist @ Microsoft. 👨‍💻 gabrielegoletto.github.io

PostsRepliesMedia
Reposted by Gabriele Goletto
Dima Damen @ECCV 2026 @dimadamen.bsky.social · 15/12/2025
Preprint now on ArXiv 📢 The N-Body Problem: Parallel Execution from Single-Person Egocentric Video Input: Single-person egocentric video 👤 Out: imagine how these tasks can be performed faster by N > 1 people, correctly e.g. N=2 👥 📎 arxiv.org/abs/2512.11393 👀 zhifanzhu.github.io/ego-nbody/ 1/4
176
Reposted by Gabriele Goletto
Dima Damen @ECCV 2026 @dimadamen.bsky.social · 10/04/2025
Now on ArXiv our @cvprconference.bsky.social #CVPR2025 paper Learning from Streaming Video with Orthogonal Gradients Instead of shuffling clips, can we learn from videos fed sequentially, where you see a clip once, in order? How to deal with the correlation of gradients over training? 1/3
1172
Reposted by Gabriele Goletto
tommiekerssies.bsky.social @tommiekerssies.bsky.social · 31/03/2025
Image segmentation doesn’t have to be rocket science. 🚀 Why build a rocket engine full of bolted-on subsystems when one elegant unit does the job? 💡 That’s what we did for segmentation. ✅ Meet the Encoder-only Mask Transformer (EoMT): tue-mps.github.io/eomt (CVPR 2025) (1/6)
184
Reposted by Gabriele Goletto
Gabriele Berton @berton-gabri.bsky.social · 14/02/2025
Excited to release the first worldwide aerial image localization method (and demo!) Take an aerial or satellite image from anywhere in the world, and AstroLoc can (probably) find its location, and provide a precise footprint! Links to paper, demo and full-length (5 min) video ⬇️
191
Reposted by Gabriele Goletto
Dima Damen @ECCV 2026 @dimadamen.bsky.social · 07/02/2025
🛑📢 HD-EPIC: A Highly-Detailed Egocentric Video Dataset hd-epic.github.io arxiv.org/abs/2502.04144 New collected videos 263 annotations/min: recipe, nutrition, actions, sounds, 3D object movement &fixture associations, masks. 26K VQA benchmark to challenge current VLMs 1/N
2346
Reposted by Gabriele Goletto
Dima Damen @ECCV 2026 @dimadamen.bsky.social · 05/12/2024
Now on ArXiv ShowHowTo: Generating Scene-Conditioned Step-by-Step Visual Instructions arxiv.org/abs/2412.01987 soczech.github.io/showhowto/ Given one real image &variable sequence of text instructions, ShowHowTo generates a multi-step sequence of images *conditioned on the scene in the REAL image* 🧵
1183