#CVPR2026 paper: It's Never Too Late: Noise Optimization for Collapse Recovery in Trained Diffusion Models Text-to-image models often collapse to near-identical samples. Our fix: optimize the noise. Start from pink 🩷, not white noise. 🔗 akoepke.github.io/divgen/index... 1/6
A. Sophia Koepke
@askoepke.bsky.social
Currently at BAIR (Berkeley). Junior research group leader at TUM | University of Tübingen. Previously at VGG (Oxford). Interested in multi-modal learning. 🔗 https://akoepke.github.io/
New paper: Back into Plato’s Cave Are vision and language models converging to the same representation of reality? The Platonic Representation Hypothesis says yes. BUT we find the evidence for this is more fragile than it looks. Project page: akoepke.github.io/cave_umwelten/ 1/9
Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale
Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale
akoepke.github.io
🎉 Excited to present our paper VGGSounder: Audio‑Visual Evaluations for Foundation Models today at #ICCV2025! 🕦 Poster Session 1 | 11:30–13:30 📍 Poster #88 Come by if you're into audio-visual learning and want to know whether multiple modalities actually help or hurt.
Our #CVPR2025 workshop on Emergent Visual Abilities and Limits of Foundation Models (EVAL-FoMo) is taking place this afternoon (1-6pm) in room 210. Workshop schedule: sites.google.com/view/eval-fo...
EVAL-FoMo 2 - Schedule
Date: June 11 (1:00pm - 6:00pm)
sites.google.com
Our paper submission deadline for the EVAL-FoMo workshop @cvprconference.bsky.social has been extended to March 19th! sites.google.com/view/eval-fo... We welcome submissions (incl. published papers) on the analysis of emerging capabilities / limits in visual foundation models. #CVPR2025
Our 2nd Workshop on Emergent Visual Abilities and Limits of Foundation Models (EVAL-FoMo) is accepting submissions. We are looking forward to talks by our amazing speakers that include @saining.bsky.social, @aidanematzadeh.bsky.social, @lisadunlap.bsky.social, and @yukimasano.bsky.social. #CVPR2025
🔥 #CVPR2025 Submit your cool papers to Workshop on Emergent Visual Abilities and Limits of Foundation Models 📷📷🧠🚀✨ sites.google.com/view/eval-fo... Submission Deadline: March 12th!
Upcoming 𝗠𝘂𝗻𝗶𝗰𝗵 𝗔𝗜 𝗟𝗲𝗰𝘁𝘂𝗿𝗲 featuring Prof. Franca Hoffmann from California Institute of Technology and Prof. Holger Hoos from RWTH Aachen University: munichlectures.ai 🗓️ December 17, 2024 🕙 16:00 CET 🏫 Senatssaal, #LMU Munich
Kicking off our TUM AI - Lecture Series tomorrow with none other than Jiaming Song, CSO @LumaLabsAI. He'll be talking about "Dream Machine: Emergent Capabilities from Video Foundation Models". Live stream: youtu.be/oilWwsXZamA 7pm GMT+1 / 10am PST (Mon Dec 2nd)