Excited to share our new paper: M2SVid: End-to-End Inpainting and Refinement for Monocular-to-Stereo Video Conversion! ACCEPTED by 3DV 2026!🎬 👉 Project: m2svid.github.io 📄 Paper: arxiv.org/abs/2505.16565 Done with Goutam Bhat, Prune Truong, @hildekuehne.bsky.social Federico Tombari 🧵👇
Nina Shvetsova
@ninashv.bsky.social
PhD student at the University of Tuebingen. Computer vision, video understanding, multimodal learning. https://ninatu.github.io/
ICCV 2025 🌺 Aloha from Hawaii! MPI-INF (D2) is presenting 4 papers this year (one Highlight). Thread 👇
🌍 11 ELLIS Members and Scholars from five countries have received ERC Starting Grants! Congratulations to all awardees! 👏 Last week @erc.europa.eu awarded 478 grants totaling €761M to support early-career researchers across Europe. 🔗 Learn more: ellis.eu/news/erc-awa...
🚀 UTD is now fully released! Code ✅ Models ✅ 2M video descriptions ✅ Debiased splits for 12 datasets ✅ Everything you need to benchmark video models more fairly is now public: 🔗 github.com/ninatu/utd-p... 🎥 Let’s make video understanding actually about video understanding.
Everybody misses the 1-page rebuttal. These lengthy forum style comments are a nightmare: a nightmare for the authors who spend way too much time writing them, a nightmare for the reviewers who spend too much time understanding them, a nightmare for the ACs who will have to summarize all. Stop it!
As a #NeurIPS AC I did a little stat. Paper reviews+responses in my batch contain 515k characters and 75k words. This is without including papers themselves or authors/reviewers discussions. So please be nice to your AC, we are doing our best. Also I miss the one page PDF response of old ICML.
Finishing your PhD or just defended? Apply to the #ICCV2025 Doctoral Consortium. Get feedback and mentorship from leading researchers in computer vision. Doctoral consortium info: iccv.thecvf.com/Conferences/...
Extended EPIC-SOUND paper was accepted at TPAMI arxiv.org/abs/2302.006... This follows ICASSP 2023 oral, extended for detection and further analysis epic-kitchens.github.io/epic-sounds/ work by @jaesunghuh.bsky.social Jacob Chalk @ekazakos.bsky.social @oxford-vgg.bsky.social @bristoluni.bsky.social
Update on hidden prompts in papers targeting LLM reviews: ICML 2025 PCs react. icml.cc/Conferences/...
Today, we release Franca, a new vision Foundation Model that matches and often outperforms DINOv2. The data, the training code and the model weights are open-source. This is the result of a close and fun collaboration @valeoai.bsky.social (in France) and @funailab.bsky.social (in Franconia)🚀
1/ Can open-data models beat DINOv2? Today we release Franca, a fully open-sourced vision foundation model. Franca with ViT-G backbone matches (and often beats) proprietary models like SigLIPv2, CLIP, DINOv2 on various benchmarks setting a new standard for open-source research.
Papers being presented from our group at #ICML2025! Congratulations to all the authors! To know more, visit us in the poster sessions! A 🧵with more details: @icmlconf.bsky.social @mpi-inf.mpg.de
Happening now! Check out the great work from Felix and Co. We improve video action grounding by >=10% on V-HICO and DALY(hope we didn't miss anyone)! Fri 13 Jun 10:30 a.m. CDT — 12:30 p.m. CDT ExHall D Poster #306 Paper: openaccess.thecvf.com/content/CVPR...
Thread: Workshop Papers from Our Lab at CVPR 2025! 🚀 👏 Huge congrats to our members on these workshop paper acceptances! Excited to see their work at #CVPR2025 🌟 #MPI-INF #D2 #Workshop #AI #ComputerVision #PhD @mpi-inf.mpg.de
Thread: Main Conference Papers from Our Lab at CVPR 2025! 🚀 👏 Big congrats to everyone! Keep an eye out at #CVPR2025 🌟 #MPI-INF #D2 #ComputerVision #AI #PhD #ML @mpi-inf.mpg.de
🎉 Exciting News #CVPR2025! We’re proud to announce that we have 5 papers accepted to the main conference and 7 papers accepted at various CVPR workshops this year! We’re looking forward to sharing our research with the community in Nashville! Stay tuned for more details! @mpi-inf.mpg.de
Do you want to present your recently accepted or ongoing work @cvprconference.bsky.social #CVPR2025 EgoVis workshop? Submit your abstract before DL of Fri 2 May, egovis.github.io/cvpr25/#cfp
a blue and white penguin is sitting on a yellow origami crane
ALT: a blue and white penguin is sitting on a yellow origami crane
media.tenor.com
🚀Excited to announce our CVPR 2025 paper: Unbiasing through Textual Descriptions! We release new descriptions for 1.9M(!) videos and object-debiased splits for 12 datasets! 🔗Project: utd-project.github.io by @ninashv.bsky.social et al 🧵👇 @cvprconference.bsky.social
UTD Dataset: Mitigating Representation Bias in Video Benchmarks
A dataset with textual descriptions and debiased splits for video benchmarks.
utd-project.github.io