Vicky Kalogeiton

@vickykalogeiton.bsky.social

Assistant Professor at Ecole Polytechnique, IP_Paris// Before: Oxford_VGG, Inria Grenoble // multimodality, genAI enthusiast // happy mum+dog_mum // opinions: mine

🎉 Our work MIRO is accepted to #ICML2026 @icmlconf.bsky.social We integrate human preferences directly during pretraining with multi-reward conditioning. ⚡MIRO is 19x faster than baselines and 370x cheaper at inference! 🤗 Try out the models: huggingface.co/spaces/nicol... See you in Seoul 🇰🇷 !

MIRO - a Hugging Face Space by nicolas-dufour

Multi-reward conditioned text-to-image diffusion (ICML 2026)

huggingface.co

Nicolas Dufour@nicolasdufour.bsky.social · 3mo ago

Thrilled to share that MIRO is accepted to ICML 2026 @icmlconf.bsky.social ! 🎉 By training on the reward scores, we can simply condition the model on high rewards at inference time to guarantee top-tier, aligned outputs. We’ve updated our paper with some additional results!

As an SAC for @neuripsconf.bsky.social, I don't agree with PCs approach to reject papers based on ranking. I ranked my accepted papers as requested and explicitly stated that I support the acceptance of all papers. I was not given an explanation of why papers at the end of the rank were rejected.

#ICCV2025 is deeply committed to promoting diversity, equity, and inclusion within our community. As part of this commitment, travel support is available to help broaden participation. Applications will be reviewed on a rolling basis until August 20, 2025 (anywhere on Earth).

Bild

Text-to-image models are trained on billions of data. But, is it necessary? Our "How far can we go with ImageNet for T2I generation?‬" @lucasdegeorge.bsky.social @arrijitghosh.bsky.social @nicolasdufour.bsky.social @davidpicard.bsky.social shows that no, if we are careful arxiv.org/abs/2502.21318

Bild
David Picard@davidpicard.eurosky.social · last yr.

🚨 New preprint! How far can we go with ImageNet for Text-to-Image generation? w. @arrijitghosh.bsky.social @lucasdegeorge.bsky.social @nicolasdufour.bsky.social @vickykalogeiton.bsky.social TL;DR: Train a text-to-image model using 1000 less data in 200 GPU hrs! 📜https://arxiv.org/abs/2502.21318 🧵👇