Our #ECCV2026 Whareformer paper tackles long-term object tracking in Egocentric videos 🎓https://arxiv.org/abs/2607.08537 see Dima's thread below 🔽 joint work w Jacob Chalk, @saptarshisinha.bsky.social @dimadamen.bsky.social from Bristol & @skamalas.bsky.social from @naverlabseurope.bsky.social
Dima Damen
@dimadamen.bsky.social
Professor of Computer Vision, @BristolUni. Senior Research Scientist @GoogleDeepMind - passionate about the temporal stream in our lives. http://dimadamen.github.io
Thanks @elliot-eu.bsky.social for the short interview on why egocentric vision is the right place to start your exploration into understanding human actions youtu.be/DfbxlgWpKSQ
Why first-person AI matters, by Dima Damen
YouTube video by ELLIOT Project
youtu.be
How can AI learn everyday actions from a person’s point of view? 👀 In this ELLIOT interview, @dimadamen.bsky.social (University of Bristol) explains why first-person vision matters and why translating human actions into language remains a major challenge for multimodal AI. Link in comments 👇
Excited to share our #ECCV2026 paper introducing Whareformer, which learns to track what is where in long egocentric videos. Proud to collaborate with Jacob Chalk and @sinhasaptarshi.bsky.social and the inspiring, humbling forces that are @dlarlus.bsky.social and @dimadamen.bsky.social
Whareformer @eccv.bsky.social #ECCV2026 fantastic collab bw @compscibristol.bsky.social @naverlabseurope.bsky.social by Jacob Chalk w/ @sinhasaptarshi.bsky.social @skamalas.bsky.social & @dlarlus.bsky.social Paper, Code &models public: jacobchalk.github.io/Whareformer/ arxiv.org/abs/2607.08537 N/N
NEW Whareformer: Learning to Track What is Where in Long Egocentric Videos @eccvconf #ECCV2026 paper The first online learning approach to track dynamic objects in ego video reasoning on appearance &3D location w track memory & explicit new track token. jacobchalk.github.io/Whareformer/ 🧵
*NEW* Our #ECCV2026 @eccv.bsky.social paper Towards in-the-wild Egocentric 3D Hand-Object Pose Estimation Now on ArXiv w Dataset, Code&model sid2697.github.io/epic-contact/ arxiv.org/abs/2606.30598 Two contributions: 1. EPIC-Contact Dataset 2. HOPformer Method &Checkpoint 🧵 1/6
#ECCV2026 paper: A scalar per patch from pre-trained ViTs enables fast moving navigation in the real world 966 *REAL* nav episodes (!!) performed by S. Janny with Dino-v3, Dino-v2, DUNE, VC1, AM-RADIO encoders show that patch features can be bottlenecked to 1 value ➡️ affordances emerge. 1/8
For your Embodied AI task you want a recurrent model with constant complexity per step, but you don't want to lose the power of transformers (which store the full obs history and attend to it)? Do not despair, we have your back. We distill transformers into recurrent transformers 1/8
Get to know @hazeldoughty.bsky.social, Asst Prof at @unileiden.bsky.social 🇳🇱 and ELLIS Member. She researches computer vision, with a focus on fine-grained video understanding under limited supervision. Her goal is to enable a deeper understanding of video content and motion. #WomenInELLIS
Congratulations @zhifan-zhu.bsky.social who passed his viva today @compscibristol.bsky.social, examined by Siyu Tang (ETH Zurich) &Andrew Calway on: Understanding 3D Dynamics of Hands, Objects and Bodies in Long Egocentric Videos zhifanzhu.github.io He is on the job market for positions in industry
What can you do with a 4-hour delay at Denver after @cvprconference.bsky.social? Create a collage of this week... 🧵 to tag all those who helped shape these captured moments... it was lovely to see you all and hope to catch up again soon...
We are excited to welcome @dimadamen.bsky.social as a speaker at the Workshop on Multimodal Foundation Models!🌐🤖 Professor of Computer Vision at the University of Bristol and Senior Research Scientist at Google DeepMind, she will present: “Modelling in an Ego-Sensed World”. 🔗 Link in the comments.
Many congratulations Prof Jitendra Malik elected as a Fellow of the Royal Society @royalsociety.org in 2026 - well and truly deserved! Jitendra's contributions extend to inspiring, advising and mentoring many in our community, and just being a very nice person! royalsociety.org/news/2026/05...
In my team we have two open positions on AI for robotics and in particular on manipulation, for junior and senior researchers. With us, connect fundamental research questions with real operational problems and contribute to emerging AI-driven robotics services. careers.werecruit.io/en/naver-lab...
From the V-Jepa2 presentation: Action anticipation in EPIC-KITCHENS is "a very challenging task and still there’s head room for improvement on this task". V-Jepa2 establishes a new SOTA but still <65% on both verb and noun acc. Mido Assran at the @iclr-conf.bsky.social workshop on World Models
Attending @iclr-conf.bsky.social #ICLR2026? Join us for the 2nd workshop on world models - full day 202 A/B @meng-yue-yang.bsky.social opening the workshop right now
I so much enjoyed Jay McClelland's talk @iclr-conf.bsky.social #ICLR2026 workshop on New Frontiers in Associative Memory. Some quotes: * AI is staged until deployed and then they are frozen at inference... AI systems are not continual learners 1/3
For those attending @iclr-conf.bsky.social maybe we need a reminder that a panel moderator should be "like a great referee: if they did their job perfectly, you barely noticed they were there, but the game moved beautifully because of them."
Workshops are always more interesting than the main conference! @iclr-conf.bsky.social What's next in Next Token Prediction workshop @jcniebles.bsky.social starting the action,
Exciting 2 weeks ahead @csail.mit.edu thanks to B Freeman for hosting.. Sad I'll be missing Antonio - but what would be a better inspirational place to stay than his office!! Already caught up with @vincentsitzmann.bsky.social @sarameghanbeery.bsky.social ... will be a fun time!
At @brics-uob.bsky.social summit today @bristoluni.bsky.social. Happy faces of users all around. I &my team are proud users thanks to national DSIT grants. I took virtual pictures with the original Isambard and the new Isambard! Thanks @simonmcs.bsky.social, Sadaf and the team for the invitation.
Congrats Ahmad Darkhalil who passed his PhD today @compscibristol.bsky.social, examined by @csprofkgd.bsky.social & @chriswolfvision.bsky.social. Thesis includes his excellent work on: VISOR, EPIC Fields, HD-EPIC, EgoPoints, &Hand-Object Detection. Check works at: ahmaddarkhalil.github.io
Next week I’ll be in Bristol as a PhD viva examiner and to give a talk, hosted by @dimadamen.bsky.social. Looking forward to it 🤗
Giving a talk this Friday at @fau.de, hosted by @visionbernie.bsky.social. Looking forward to it 🤗
Thanks Bristol Centre for Supercomputing (BriCS) for featuring our work as users of the fantastic #IsambardAI w/ Michael Wray, J Chalk @prajwalgatti.bsky.social @sinhasaptarshi.bsky.social S Bansal, J Zhao @bristoluni.bsky.social @compscibristol.bsky.social www.youtube.com/watch?v=jxrn...
Case study - Isambard-AI: Assistive Technology
YouTube video by BriCS
youtube.com
anyone else keeps confusing overleaf with openreview when they type? I keep typing openreview when I mean overleaf, and overleaf when I mean openreview! #overleaf #openreview Ideas to solve my brain wiring are welcome!
Special thanks Ayush @ayusht.bsky.social @universitypress.cambridge.org for visiting us @bristoluni.bsky.social @compscibristol.bsky.social for a great #MaVi Seminar on "Physical Inductive Biases for World Models" and thoughtful 1-1s with the researchers. Have a good trip back &pls visit again soon,
Make sure your paper is recognised... 24 hours until the deadline for this year's EgoVis Distinguished Paper Awards #CVPR2026 @cvprconference.bsky.social #EgoVis #Paper #Award
Call for Nominations EgoVis 2024/2025 Distinguished Paper Awards. Published a paper contributing to Ego Vision in 2024/25? Innovative &advancing Ego Vision? Worthy of a prize? DL for nominations 20 Feb 2026 Awards announced @cvprconference.bsky.social #CVPR2026 egovis.github.io/awards/2024_...
UPDATE: We've updated the download process and now you can download the videos of our HowToGround1M dataset in addition to the iGround dataset. Also, we now provide access to the full HowTo100M dataset! Download our datasets, or HowTo100M at github.com/ekazakos/grove
GitHub - ekazakos/grove: Code implementation for the paper "Large-scale Pre-training for Grounded Video Caption Generation" (ICCV 2025)
Code implementation for the paper "Large-scale Pre-training for Grounded Video Caption Generation" (ICCV 2025) - ekazakos/grove
github.com
We have released the code, checkpoints and datasets for GROVE! Grab them at github.com/ekazakos/grove. We also provide grove-transformers — an inference-only interface for GROVE, implemented with 🤗 Transformers. You can load GROVE and try it on any video data in just a few lines of code!
If you missed this before NY... A reminder that our @compscibristol.bsky.social #MaVi Summer Program - for current PhD students from Europe inc. @ellis.eu unit students and internationally is open for applications - DL 29 Jan 2026. uob-mavi.github.io/Summer@MaVi....
Applications are open for visiting PhD @compscibristol.bsky.social @bristoluni.bsky.social in 2026 - DL 29 Jan Would you like to work with any of the Faculty working in Machine Learning and Computer Vision #mavi as part of our summer of research at Bristol program? uob-mavi.github.io/Summer@MaVi....
3rd Egocentric Vision (EgoVis) workshop will be held as a full day workshop @cvprconference.bsky.social #CVPR2026 egovis.github.io/cvpr26/ CFP and challenge deadlines after the NY Great lineup of 7 keynote speakers... See you in Denver!