Hermann Blum

@hermannblum.bsky.social

ML & CV for robot perception assistant professor @ Uni Bonn & Lamarr Institute interested in self-learning & autonomous robots, likes all the messy hardware problems of real-world experiments https://rpl.uni-bonn.de/ https://hermannblum.net/

In 30 mins! CroCoDL Poster Session (posters #248-#257), during the official coffee break 3pm - 4pm. Contributed works cover Visual Localization, Visual Place Recognition, Room Layout Estimation, Novel View Synthesis, 3D Reconstruction

Bild

We just released code, models, and data for FrontierNet! Key idea 💡Instead of detecting froniers in a map, we directly predict them from images. Hence, FrontierNet can implicitly learn visual semantic priors to estimate information gain. That speeds up exploration compared to geometric heuristics.

We just extended our submission deadline for 8-page paper submissions until June 30. Accepted submissions go into ICCV WS proceedings 📄

Zuria Bauer@zbauer.bsky.social · last yr.

🌺 Excited to announce the 1st CroCoDL Workshop on Large-Scale Cross-Device Localization at #ICCV2025! Tackling real-world localization across smartphones, AR/VR, and robots with invited talks, paper track & competition on a novel dataset.   🔗 localizoo.com/workshop 🧵 1/2

We are organizing the 1st Workshop on Cross-Device Visual Localization at #ICCV #ICCV2025 Localizing multiple phones, headsets, and robots to a common reference frame is so far a real problem in mixed-reality applications. Our new challenge will track progress on this issue. ⏰ paper deadline: June 6

Zuria Bauer@zbauer.bsky.social · last yr.

🌺 Excited to announce the 1st CroCoDL Workshop on Large-Scale Cross-Device Localization at #ICCV2025! Tackling real-world localization across smartphones, AR/VR, and robots with invited talks, paper track & competition on a novel dataset.   🔗 localizoo.com/workshop 🧵 1/2

Open source code now available MASt3R-SLAM: the best dense visual SLAM system I've ever seen. Real-time and monocular, and easy to run with a live camera or on videos without needing to know the camera calibration. Brilliant work from Eric and Riku.

Riku Murai@rmurai0610.bsky.social · last yr.

MASt3R-SLAM code release! github.com/rmurai0610/M... Try it out on videos or with a live camera Work with @ericdexheimer.bsky.social*, @ajdavison.bsky.social (*Equal Contribution)

Very proud of Boyang for this work, great to see first shoutouts! The gist is that exploration has always been treated as a geometric problem, but we show visual cues are really helpful to detect frontiers and predict their info gain. W/ FrontierNet, you can get RGB-only exploration/object search/+

Dmytro Mishkin@ducha-aiki.bsky.social · 2y ago

FrontierNet: Learning Visual Cues to Explore Boyang Sun, Hanzhi Chen, Stefan Leutenegger, Cesar Cadena, @marcpollefeys.bsky.social , Hermann Blum tl;dr: predict frontier (where we weren't yet) using RGBD and then make a map, and not otherwise. arxiv.org/abs/2501.04597