Reposted by Linus Härenstam-NielsenVladimir Yugay @vyuga3d.bsky.social · 07/10/2025📽️ Check out Visual Odometry Transformer! VoT is an end-to-end model for getting accurate metric camera poses from monocular videos. vladimiryugay.github.io/vot/ 1104
Reposted by Linus Härenstam-NielsenAndreas Geiger @andreasgeiger.bsky.social · 01/10/2025#TTT3R: 3D Reconstruction as Test-Time Training TTT3R offers a simple state update rule to enhance length generalization for #CUT3R — No fine-tuning required! 🔗Page: rover-xingyu.github.io/TTT3R We rebuilt @taylorswift13’s "22" live at the 2013 Billboard Music Awards - in 3D! 0384
Reposted by Linus Härenstam-NielsenChristoph Reich @christophreich.bsky.social · 09/07/2025🦖 We present “Feed-Forward SceneDINO for Unsupervised Semantic Scene Completion”. #ICCV2025 🌍: visinf.github.io/scenedino/ 📃: arxiv.org/abs/2507.06230 🤗: huggingface.co/spaces/jev-a... @jev-aleks.bsky.social @fwimbauer.bsky.social @olvrhhn.bsky.social @stefanroth.bsky.social @dcremers.bsky.social 12410
Linus Härenstam-Nielsen @linushn.bsky.social · 09/07/2025The code for our #CVPR2025 paper, PRaDA: Projective Radial Distortion Averaging, is now out! Turns out distortion calibration from multiview 2D correspondences can be fully decoupled from 3D reconstruction, greatly simplifying the problem arxiv.org/abs/2504.16499 github.com/DaniilSinits... 1125
Reposted by Linus Härenstam-NielsenDominik Schnaus @schnaus.bsky.social · 03/06/2025Can we match vision and language representations without any supervision or paired data? Surprisingly, yes! Our #CVPR2025 paper with @neekans.bsky.social and @dcremers.bsky.social shows that the pairwise distances in both modalities are often enough to find correspondences. ⬇️ 1/4 12712
Reposted by Linus Härenstam-NielsenFelix Wimbauer @fwimbauer.bsky.social · 13/05/2025Can you train a model for pose estimation directly on casual videos without supervision? Turns out you can! In our #CVPR2025 paper AnyCam, we directly train on YouTube videos and achieve SOTA results by using an uncertainty-based flow loss and monocular priors! ⬇️ 12410
Reposted by Linus Härenstam-Nielsenreqo.bsky.social @reqo.bsky.social · 03/04/2025Our paper, ”Semantic Library Adaptation: LoRA Retrieval and Fusion for Open-Vocabulary Semantic Segmentation”, has been accepted to #CVPR 2025. 📄 Paper: arxiv.org/abs/2503.21780 🧪 Code: github.com/rezaqorbani/...arxiv.orgSemantic Library Adaptation: LoRA Retrieval and Fusion for Open-Vocabulary Semantic SegmentationOpen-vocabulary semantic segmentation models associate vision and text to label pixels from an undefined set of classes using textual queries, providing versatile performance on novel datasets. Howeve... 111
Reposted by Linus Härenstam-NielsenSimon Weber @simwebertum.bsky.social · 24/03/2025Very glad to announce that our "Finsler Multi-Dimensional Scaling" paper, accepted at #CVPR2025, is now on Arxiv! arxiv.org/abs/2503.18010 183
Reposted by Linus Härenstam-NielsenZhenjun Zhao @ericzzj.bsky.social · 18/03/2025AnyCalib: On-Manifold Learning for Model-Agnostic Single-View Camera Calibration Javier Tirado-Garín, @jcivera.bsky.social tl;dr: image->ViT+DPT->Field of View (FoV) fields->bijective rays and corresponding image coordinates->closed-form model-agnostic intrinsics arxiv.org/abs/2503.12701 052
Reposted by Linus Härenstam-NielsenDaniel Cremers @dcremers.bsky.social · 13/03/2025We are thrilled to have 12 papers accepted to #CVPR2025. Thanks to all our students and collaborators for this great achievement! For more details check out cvg.cit.tum.de 13612
Reposted by Linus Härenstam-NielsenZhenjun Zhao @ericzzj.bsky.social · 04/03/2025MUSt3R: Multi-view Network for Stereo 3D Reconstruction Yohann Cabon, Lucas Stoffl, Leonid Antsfeld, Gabriela Csurka, Boris Chidlovskii, Jerome Revaud, @vincentleroy.bsky.social tl;dr: make DUSt3R symmetric and iterative+multi-layer memory mechanism->multi-view DUSt3R arxiv.org/abs/2503.01661 1254
Reposted by Linus Härenstam-NielsenLu Sang @lu-sang.bsky.social · 03/03/2025🥳 Thrilled to announce that our work, "4Deform: Neural Surface Deformation for Robust Shape Interpolation," has been accepted to #CVPR2025 🙌 💻 Check our project page: 4deform.github.io 👏 Great thanks to my amazing co-authors. @ricmarin.bsky.social @dongliangcao.bsky.social @dcremers.bsky.social 093
Reposted by Linus Härenstam-NielsenLu Sang @lu-sang.bsky.social · 23/01/2025🥳Thrilled to share our work, "Implicit Neural Surface Deformation with Explicit Velocity Fields", accepted at #ICLR2025 👏 code is available at: github.com/Sangluisme/I... 😊Huge thanks to my amazing co-authors. @dongliangcao.bsky.social @dcremers.bsky.social 👏Special thanks to @ricmarin.bsky.social 0196
Reposted by Linus Härenstam-NielsenDaniel Cremers @dcremers.bsky.social · 16/01/2025Indeed - everyone had a blast - thank you all for the great talks, discussions and Ski/snowboarding! 1454
Linus Härenstam-Nielsen @linushn.bsky.social · 08/01/2025DiffCD: A Symmetric Differentiable Chamfer Distance for Neural Implicit Surface Fitting, presented at #ECCV2024 Paper: arxiv.org/abs/2407.17058 Code/project: github.com/linusnie/dif... 000
Linus Härenstam-Nielsen @linushn.bsky.social · 08/01/2025Reposting some of my prior works here on this site :) "Semidefinite Relaxations for Robust Multiview Triangulation" at #CVPR2023! paper: arxiv.org/abs/2301.11431 code: github.com/Linusnie/rob... 020