Hermann Blum @hermannblum.bsky.social · 28/08/2026A new result that I am very excited about: A self-supervised visuospatial representation with better 3D awareness. If you have RGB-D inputs (robots usually have), our model makes use of the depth to give you better semantic and geometric features than DINO. 📄 arxiv.org/abs/2608.27226 000
Hermann Blum @hermannblum.bsky.social · 21/10/2025CroCoDL Workshop at #ICCV, next talk coming up in Room 301B: 16.30 - 17.00: @sattlertorsten.bsky.social on Vision Localization Across Modalities full schedule: localizoo.com/workshop 010
Hermann Blum @hermannblum.bsky.social · 21/10/2025In 30 mins! CroCoDL Poster Session (posters #248-#257), during the official coffee break 3pm - 4pm. Contributed works cover Visual Localization, Visual Place Recognition, Room Layout Estimation, Novel View Synthesis, 3D Reconstruction 011
Hermann Blum @hermannblum.bsky.social · 21/10/2025CroCoDL Workshop at #ICCV, next talk coming up in Room 301B: 14.15 - 14.45: @ayoungk.bsky.social on Bridging heterogeneous sensors for robust and generalizable localization full schedule: localizoo.com/workshop/ 011
Hermann Blum @hermannblum.bsky.social · 21/10/2025CroCoDL Workshop at #ICCV, next talk coming up in Room 301B: 13.15 - 13.45: @gabrielacsurka.bsky.social on Privacy Preserving Visual Localization full schedule: buff.ly/kM1Ompf 052
Hermann Blum @hermannblum.bsky.social · 20/10/2025Attending #ICCV? Join the CroCoDL workshop this afternoon! localizoo.com/workshop Speakers: @gabrielacsurka.bsky.social, @ayoungk.bsky.social, David Caruso, @sattlertorsten.bsky.social w/ @zbauer.bsky.social @mihaidusmanu.bsky.social @linfeipan.bsky.social @marcpollefeys.bsky.social 011
Hermann Blum @hermannblum.bsky.social · 05/09/2025Today a delivery arrived that marks an exciting milestone for my lab: our first research grant! 010
Hermann Blum @hermannblum.bsky.social · 17/07/2025We just released code, models, and data for FrontierNet! Key idea 💡Instead of detecting froniers in a map, we directly predict them from images. Hence, FrontierNet can implicitly learn visual semantic priors to estimate information gain. That speeds up exploration compared to geometric heuristics. 120
Hermann Blum @hermannblum.bsky.social · 21/06/2025This is the first time in a while I am creating a new talk. This will be fun! I'll be up later today at the Visual SLAM workshop at @roboticsscisys.bsky.social buff.ly/ADHxPsX 020
Hermann Blum @hermannblum.bsky.social · 19/06/2025Finally arriving home today after attending @cvprconference.bsky.social . This was the first #CVPR that I could attend in person! I expected it to be super crowded but was surprised - lots of time and space for chats at the poster session and the 15min talks could really go into detail. 180
Reposted by Hermann BlumZuria Bauer @zbauer.bsky.social · 15/06/2025Do you want to learn more about our novel dataset for Cross-device localization? Come by poster 121 and meet CroCoDL 🐊 cc @marcpollefeys.bsky.social @hermannblum.bsky.social @mihaidusmanu.bsky.social @cvprconference.bsky.social @ethz.ch 051
Hermann Blum @hermannblum.bsky.social · 14/06/2025We just extended our submission deadline for 8-page paper submissions until June 30. Accepted submissions go into ICCV WS proceedings 📄 010
Hermann Blum @hermannblum.bsky.social · 11/06/2025all set for our poster lineup at the cv4mr.github.io #cvpr workshop 000
Hermann Blum @hermannblum.bsky.social · 09/06/2025I‘ll be at CVPR this week and I am actively looking for PhD students (job announcement will go out the week after). Just send me a message if you are interested to meet up. 000
Reposted by Hermann BlumHaofei Xu @haofeixu.bsky.social · 05/06/2025Excited to present our #CVPR2025 paper DepthSplat next week! DepthSplat is a feed-forward model that achieves high-quality Gaussian reconstruction and view synthesis in just 0.6 seconds. Looking forward to great conversations at the conference! 3267
Hermann Blum @hermannblum.bsky.social · 17/05/2025If you‘re watching #eurovision tonight, look out for the robots from ETH! Really cool to see something I could work with during my PhD featured as a swiss highlight 🤖 091
Hermann Blum @hermannblum.bsky.social · 13/05/2025We are organizing the 1st Workshop on Cross-Device Visual Localization at #ICCV #ICCV2025 Localizing multiple phones, headsets, and robots to a common reference frame is so far a real problem in mixed-reality applications. Our new challenge will track progress on this issue. ⏰ paper deadline: June 6 020
Reposted by Hermann BlumAndreas Geiger @andreasgeiger.bsky.social · 24/04/2025🏠 Introducing DepthSplat: a framework that connects Gaussian splatting with single- and multi-view depth estimation. This enables robust depth modeling and high-quality view synthesis with state-of-the-art results on ScanNet, RealEstate10K, and DL3DV. 🔗 haofeixu.github.io/depthsplat/ 13913
Hermann Blum @hermannblum.bsky.social · 02/04/2025Exciting news for LabelMaker! 1️⃣ ARKitLabelMaker, the largest annotated 3D dataset, was accepted to CVPR 2025! This was an amazing effort of Guangda Ji 👏 🔗 labelmaker.org 📄 arxiv.org/abs/2410.13924 2️⃣ Mahta Moshkelgosha extended the pipeline to generate 3D scene graphs: 👩💻 github.com/cvg/LabelMak...labelmaker.orgLabelMaker 🎨LabelMaker 010
Reposted by Hermann BlumNikhil Garg @nkgarg.bsky.social · 10/03/2025*Please repost* @sjgreenwood.bsky.social and I just launched a new personalized feed (*please pin*) that we hope will become a "must use" for #academicsky. The feed shows posts about papers filtered by *your* follower network. It's become my default Bluesky experience bsky.app/profile/pape... 23535299
Reposted by Hermann BlumAndrew Davison @ajdavison.bsky.social · 25/02/2025Open source code now available MASt3R-SLAM: the best dense visual SLAM system I've ever seen. Real-time and monocular, and easy to run with a live camera or on videos without needing to know the camera calibration. Brilliant work from Eric and Riku. 0316
Hermann Blum @hermannblum.bsky.social · 24/02/2025We have an excellent opportunity for a tenured, flagship AI professorship at @unibonn.bsky.social and lamarr-institute.org Application Deadline is End of March. www.uni-bonn.de/en/universit...uni-bonn.deFull Professorship (W3) in Artificial Intelligence and Machine LearningW3 Professorship 040
Hermann Blum @hermannblum.bsky.social · 05/02/2025Very proud of Boyang for this work, great to see first shoutouts! The gist is that exploration has always been treated as a geometric problem, but we show visual cues are really helpful to detect frontiers and predict their info gain. W/ FrontierNet, you can get RGB-only exploration/object search/+ 020
Hermann Blum @hermannblum.bsky.social · 03/12/2024Turns out aria-glasses are a very useful tool to demonstrate actions to robots: Based on egocentric video we track dynamic changes in a scene graph and use the representation to replay or plan interactions for robots 🔗 behretj.github.io/LostAndFound/ 📄 arxiv.org/abs/2411.19162 📺 youtu.be/xxMsaBSeMXo 1276
Reposted by Hermann BlumJane Wang @janexwang.bsky.social · 27/11/2024Applications are open for student researcher positions at Google DeepMind for 2025! Due date Dec 13 www.google.com/about/career...google.comStudent Researcher, 2025 — Google Careers 1317
Reposted by Hermann BlumMatías Mattamala @mmattamala.bsky.social · 24/11/2024Reposting the SLAM Handbook again for the robotics people arriving here :) 2338
Hermann Blum @hermannblum.bsky.social · 19/09/2024Are you also a bit exhausted after #ICRA submission week? Let us brighten your day with a real "SpotLight" 💡 🔗 timengelbracht.github.io/SpotLight/ 📄 arxiv.org/abs/2409.11870 We detect and generate interaction for almost any light switch and can then map which switch turns on which light #Robotics 040