Jiaang Li

@jiaangli.bsky.social

PhD student at University of Copenhagen @belongielab.org | #nlp #computervision | ELLIS student @ellis.eu ๐ŸŒ https://jiaangli.github.io/

(1/n) Thrilled to share my first paper at Meta FAIR! "EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data" ๐Ÿ‘ถ Human infants learn language from sparse, noisy multimodal input. Today's VLMs can't. We built a benchmark + challenge to close that gap. ๐Ÿงต

Bild

๐ŸŽ‰ Excited to share our work "๐—ฅ๐—”๐—ฉ๐—˜๐—ก๐—˜๐—”: ๐—” ๐—•๐—ฒ๐—ป๐—ฐ๐—ต๐—บ๐—ฎ๐—ฟ๐—ธ ๐—ณ๐—ผ๐—ฟ ๐— ๐˜‚๐—น๐˜๐—ถ๐—บ๐—ผ๐—ฑ๐—ฎ๐—น ๐—ฅ๐—ฒ๐˜๐—ฟ๐—ถ๐—ฒ๐˜ƒ๐—ฎ๐—น-๐—”๐˜‚๐—ด๐—บ๐—ฒ๐—ป๐˜๐—ฒ๐—ฑ ๐—ฉ๐—ถ๐˜€๐˜‚๐—ฎ๐—น ๐—–๐˜‚๐—น๐˜๐˜‚๐—ฟ๐—ฒ ๐—จ๐—ป๐—ฑ๐—ฒ๐—ฟ๐˜€๐˜๐—ฎ๐—ป๐—ฑ๐—ถ๐—ป๐—ด", accepted at #ICLR2026! ๐Ÿ‡ง๐Ÿ‡ท I'll be attending ICLR in person โ€” would love to connect and chat there! ๐Ÿค ๐Ÿ—“๏ธ Sat, Apr 25, 2026, 10:30 AM โ€“ 1:00 PM GMT-03 ๐Ÿ“ Pavilion 4 P4-# 3618

Bild

Feeling overwhelmed by all the recent developments in video understanding? What used to require dozens of modular computational workflows involving SLAM, feature tracking, optical flow, camera calibration, multiview geometric constraints, and resnet backbones is now... (1/3)

Bild

Would you present your next NeurIPS paper in Europe instead of traveling to San Diego (US) if this was an option? Sรธren Hauberg (DTU) and I would love to hear the answer through this poll: (1/6)

NeurIPS participation in Europe

We seek to understand if there is interest in being able to attend NeurIPS in Europe, i.e. without travelling to San Diego, US. In the following, assume that it is possible to present accepted papers ...

docs.google.com

Check out our new preprint ๐“๐ž๐ง๐ฌ๐จ๐ซ๐†๐‘๐š๐ƒ. We use a robust decomposition of the gradient tensors into low-rank + sparse parts to reduce optimizer memory for Neural Operators by up to ๐Ÿ•๐Ÿ“%, while matching the performance of Adam, even on turbulent Navierโ€“Stokes (Re 10e5).

Bild

I am excited to announce our latest work ๐ŸŽ‰ "Cultural Evaluations of Vision-Language Models Have a Lot to Learn from Cultural Theory". We review recent works on culture in VLMs and argue for deeper grounding in cultural theory to enable more inclusive evaluations. Paper ๐Ÿ”—: arxiv.org/pdf/2505.22793

Paper title "Cultural Evaluations of Vision-Language Models
Have a Lot to Learn from Cultural Theory"

I wonโ€™t be attending #ICLR in person this year๐Ÿ˜ข. But feel free to check our paper โ€˜Revisiting the Othello World Model Hypothesisโ€™ with Anders Sรธgaard, accepted at ICLR world models workshop! Paper link arxiv.org/abs/2503.04421

Revisiting the Othello World Model Hypothesis

Li et al. (2023) used the Othello board game as a test case for the ability of GPT-2 to induce world models, and were followed up by Nanda et al. (2023b). We briefly discuss the original experiments, ...

arxiv.org

Forget just thinking in words. ๐Ÿ””Our New Preprint: ๐Ÿš€ New Era of Multimodal Reasoning๐Ÿšจ ๐Ÿ” Imagine While Reasoning in Space with MVoT Multimodal Visualization-of-Thought (MVoT) revolutionizes reasoning by generating visual "thoughts" that transform how AI thinks, reasons, and explains itself.

Bild

FGVC12 Workshop is coming to #CVPR 2025 in Nashville! Are you working on fine-grained visual problems? This year we have two peer-reviewed paper tracks: i) 8-page CVPR Workshop proceedings ii) 4-page non-archival extended abstracts CALL FOR PAPERS: sites.google.com/view/fgvc12/...

FGVC12 Workshop - Submission

Call for Papers Workshop - Date TBC (either June 11th or 12th 2025) FGVC12 will have two paper tracks and a nectar track: Proceedings track: 8-page papers that will appear in the official CVPR worksho...

sites.google.com

FGVC Workshop@fgvcworkshop.bsky.social ยท 2y ago

FGVC12 Workshop accepted to CVPR 2025, Nashville! CALL FOR PAPERS: sites.google.com/view/fgvc12/... We discuss domains where expert knowledge is typically required and investigate artificial systems that can efficiently distinguish a large number of very similar visual concepts. #CVPR #CVPR2025 #AI

FGVC12 Workshop at CVPR 2025 - Call for papers.

I'm recruiting 1-2 PhD students to work with me at the University of Colorado Boulder! Looking for creative students with interests in #NLP and #CulturalAnalytics. Boulder is a lovely college town 30 minutes from Denver and 1 hour from Rocky Mountain National Park ๐Ÿ˜Ž Apply by December 15th!

A photo of Boulder, Colorado, shot from above the university campus and looking toward the Flatirons.