CroCoDL Workshop at #ICCV, next talk coming up in Room 301B: 16.30 - 17.00: @sattlertorsten.bsky.social on Vision Localization Across Modalities full schedule: localizoo.com/workshop
Hermann Blum
@hermannblum.bsky.social
ML & CV for robot perception assistant professor @ Uni Bonn & Lamarr Institute interested in self-learning & autonomous robots, likes all the messy hardware problems of real-world experiments https://rpl.uni-bonn.de/ https://hermannblum.net/
In 30 mins! CroCoDL Poster Session (posters #248-#257), during the official coffee break 3pm - 4pm. Contributed works cover Visual Localization, Visual Place Recognition, Room Layout Estimation, Novel View Synthesis, 3D Reconstruction
CroCoDL Workshop at #ICCV, next talk coming up in Room 301B: 14.15 - 14.45: @ayoungk.bsky.social on Bridging heterogeneous sensors for robust and generalizable localization full schedule: localizoo.com/workshop/
CroCoDL Workshop at #ICCV, next talk coming up in Room 301B: 13.15 - 13.45: @gabrielacsurka.bsky.social on Privacy Preserving Visual Localization full schedule: buff.ly/kM1Ompf
Attending #ICCV? Join the CroCoDL workshop this afternoon! localizoo.com/workshop Speakers: @gabrielacsurka.bsky.social, @ayoungk.bsky.social, David Caruso, @sattlertorsten.bsky.social w/ @zbauer.bsky.social @mihaidusmanu.bsky.social @linfeipan.bsky.social @marcpollefeys.bsky.social
Today a delivery arrived that marks an exciting milestone for my lab: our first research grant!
We just released code, models, and data for FrontierNet! Key idea 💡Instead of detecting froniers in a map, we directly predict them from images. Hence, FrontierNet can implicitly learn visual semantic priors to estimate information gain. That speeds up exploration compared to geometric heuristics.
This is the first time in a while I am creating a new talk. This will be fun! I'll be up later today at the Visual SLAM workshop at @roboticsscisys.bsky.social buff.ly/ADHxPsX
Finally arriving home today after attending @cvprconference.bsky.social . This was the first #CVPR that I could attend in person! I expected it to be super crowded but was surprised - lots of time and space for chats at the poster session and the 15min talks could really go into detail.
Do you want to learn more about our novel dataset for Cross-device localization? Come by poster 121 and meet CroCoDL 🐊 cc @marcpollefeys.bsky.social @hermannblum.bsky.social @mihaidusmanu.bsky.social @cvprconference.bsky.social @ethz.ch
We just extended our submission deadline for 8-page paper submissions until June 30. Accepted submissions go into ICCV WS proceedings 📄
🌺 Excited to announce the 1st CroCoDL Workshop on Large-Scale Cross-Device Localization at #ICCV2025! Tackling real-world localization across smartphones, AR/VR, and robots with invited talks, paper track & competition on a novel dataset. 🔗 localizoo.com/workshop 🧵 1/2
I‘ll be at CVPR this week and I am actively looking for PhD students (job announcement will go out the week after). Just send me a message if you are interested to meet up.
Excited to present our #CVPR2025 paper DepthSplat next week! DepthSplat is a feed-forward model that achieves high-quality Gaussian reconstruction and view synthesis in just 0.6 seconds. Looking forward to great conversations at the conference!
🏠 Introducing DepthSplat: a framework that connects Gaussian splatting with single- and multi-view depth estimation. This enables robust depth modeling and high-quality view synthesis with state-of-the-art results on ScanNet, RealEstate10K, and DL3DV. 🔗 haofeixu.github.io/depthsplat/
If you‘re watching #eurovision tonight, look out for the robots from ETH! Really cool to see something I could work with during my PhD featured as a swiss highlight 🤖
We are organizing the 1st Workshop on Cross-Device Visual Localization at #ICCV #ICCV2025 Localizing multiple phones, headsets, and robots to a common reference frame is so far a real problem in mixed-reality applications. Our new challenge will track progress on this issue. ⏰ paper deadline: June 6
🌺 Excited to announce the 1st CroCoDL Workshop on Large-Scale Cross-Device Localization at #ICCV2025! Tackling real-world localization across smartphones, AR/VR, and robots with invited talks, paper track & competition on a novel dataset. 🔗 localizoo.com/workshop 🧵 1/2
🏠 Introducing DepthSplat: a framework that connects Gaussian splatting with single- and multi-view depth estimation. This enables robust depth modeling and high-quality view synthesis with state-of-the-art results on ScanNet, RealEstate10K, and DL3DV. 🔗 haofeixu.github.io/depthsplat/
Exciting news for LabelMaker! 1️⃣ ARKitLabelMaker, the largest annotated 3D dataset, was accepted to CVPR 2025! This was an amazing effort of Guangda Ji 👏 🔗 labelmaker.org 📄 arxiv.org/abs/2410.13924 2️⃣ Mahta Moshkelgosha extended the pipeline to generate 3D scene graphs: 👩💻 github.com/cvg/LabelMak...
LabelMaker 🎨
LabelMaker
labelmaker.org
*Please repost* @sjgreenwood.bsky.social and I just launched a new personalized feed (*please pin*) that we hope will become a "must use" for #academicsky. The feed shows posts about papers filtered by *your* follower network. It's become my default Bluesky experience bsky.app/profile/pape...
Open source code now available MASt3R-SLAM: the best dense visual SLAM system I've ever seen. Real-time and monocular, and easy to run with a live camera or on videos without needing to know the camera calibration. Brilliant work from Eric and Riku.
MASt3R-SLAM code release! github.com/rmurai0610/M... Try it out on videos or with a live camera Work with @ericdexheimer.bsky.social*, @ajdavison.bsky.social (*Equal Contribution)
We have an excellent opportunity for a tenured, flagship AI professorship at @unibonn.bsky.social and lamarr-institute.org Application Deadline is End of March. www.uni-bonn.de/en/universit...
Full Professorship (W3) in Artificial Intelligence and Machine Learning
W3 Professorship
uni-bonn.de
Very proud of Boyang for this work, great to see first shoutouts! The gist is that exploration has always been treated as a geometric problem, but we show visual cues are really helpful to detect frontiers and predict their info gain. W/ FrontierNet, you can get RGB-only exploration/object search/+
FrontierNet: Learning Visual Cues to Explore Boyang Sun, Hanzhi Chen, Stefan Leutenegger, Cesar Cadena, @marcpollefeys.bsky.social , Hermann Blum tl;dr: predict frontier (where we weren't yet) using RGBD and then make a map, and not otherwise. arxiv.org/abs/2501.04597
Turns out aria-glasses are a very useful tool to demonstrate actions to robots: Based on egocentric video we track dynamic changes in a scene graph and use the representation to replay or plan interactions for robots 🔗 behretj.github.io/LostAndFound/ 📄 arxiv.org/abs/2411.19162 📺 youtu.be/xxMsaBSeMXo
Applications are open for student researcher positions at Google DeepMind for 2025! Due date Dec 13 www.google.com/about/career...
Student Researcher, 2025 — Google Careers
google.com
Reposting the SLAM Handbook again for the robotics people arriving here :)
We are in the process of editing a SLAM handbook, to be published by Cambridge University Press, with many *stellar* contributors. Part 1 is available as an online draft for public comments. Help us find bugs/problems! Link to release repo is here: lnkd.in/gZhTkaxb
Are you also a bit exhausted after #ICRA submission week? Let us brighten your day with a real "SpotLight" 💡 🔗 timengelbracht.github.io/SpotLight/ 📄 arxiv.org/abs/2409.11870 We detect and generate interaction for almost any light switch and can then map which switch turns on which light #Robotics