Chris Offner

@chrisoffner3d.bsky.social

Student Researcher @ RAI Institute, MSc CS Student @ ETH Zurich visual computing, 3D vision, spatial AI, machine learning, robot perception. 📍Zurich, Switzerland

Huge props to Lord and Miller for stepping up and doing what directors like Nolan are too cowardly to do: be up front and give loud and generous credit to their VFX team. Great directors don’t need to lie about how their movies are made; the work speaks for itself.

Bild

All the "you need to learn AI skills or you'll get left behind" things are patently nonsense. It's easy to use and only becomes easier to use over time. If there's skill it's in knowing what it does well and what is does poorly

Looking forward to a busy #ICCV2025. I will give three (very different) talks at workshops and tutorials, see info below. We also present two papers, ACE-G and SCR Priors. And it's the 10th (!) anniversary of the R6D workshop, which we co-organize.

Bild

🚀 Europe’s first exascale supercomputer is here! JUPITER, launched in Germany, is the EU’s most powerful system and fourth fastest worldwide. 100% powered by renewables, it has also ranked first in energy efficiency. It will boost AI, science, and climate research. Read more - europa.eu/!vcWBqW

A futuristic corridor inside a data center with rows of tall, blue-lit server racks on both sides. Text overlaid at the bottom reads "JUPITER Supercomputer: Europe enters the exascale supercomputing league." In the lower right corner, there is a logo of the European Commission.

There is a lot to hate about the politics of the silicon valley right, but they do actually want to build stuff, and I would prefer if the left didn't cede "we should be able to build stuff" to the right.

I can't* fathom why the top picture, and not the bottom picture, is the standard diagram for an autoencoder. The whole idea of an autoencoder is that you complete a round trip and seek cycle consistency—why lay out the network linearly?

Bild

In general I think 3D vision would do well to take some inspiration from Bayesians. I guess these days they lost their glamour, but imo it's a very nice way of thinking that feels somewhat lost currently.

"It is beautiful. It is elegant. Does it work well in practice? Not really. This is often the caveat we face in research: the things that are beautiful don't work and the things that work are not beautiful." – Daniel Cremers

"As roboticists and computer vision people [outside of big tech], do we have to just wait for the next foundation model?" I share the frustration. It's disempowering when most major progress recently is downstream of "foundation models" that you don't have the compute or data to train yourself.

Yay, DINOv3 is out! SigLIP (VLMs) and DINO are two competing paradigms for image encoders. My intuition is that joint vision-language modeling works great for semantic problems but may be too coarse for geometry problems like SfM or SLAM. Most animals navigate 3D space perfectly without language.

BildBild

What are the best resources to learn about VLMs? Papers, tutorials, courses, blog posts, whatever is good. I can read the Kimi-VL or GLM tech reports and follow the breadcrumbs but I'd appreciate any and all recommendations towards a useful VLM curriculum! 🙏

I agree 100%. It's one thing to criticize corporate practices, the social impact, ethics, or future risks. But I watch in total awe how it writes in a few seconds a well documented program in a language/API I don't know, while they complain "but it might have a bug and requires a pass or two" O_o

gilbetron 60 days ago | next [=]
I get so confused on this. I play around, test, and mess with LLMs all the time and they are miraculous. Just amazing, doing things we dreamed about for decades. I mean, I can ask for obscure things with subtle nuance where I misspell words and mess up my question and it figures it out. It talks to me like a person. It generates really cool images. It helps me write code. And just tons of other stuff that astounds me.
And people just sit around, unimpressed, and complain that ... what ... it isn't a perfect superintelligence that understands everything perfectly? This is the most amazing technology I've experienced as a 50+ year old nerd that has been sitting deep in tech for basically my whole life. This is the stuff of science fiction, and while there totally are limitations, the speed at which it is progressing is insane. And people are like, "Wah, it can't write code like a Senior engineer with 20 years of experience!"
Crazy.

What happens when ETH Zurich teams up with Google to shape the future of #AugmentedReality? We're opening a new #Researchhub that brings together top minds to tackle one of the biggest challenges in tech: seamlessly blending the digital and physical worlds. Read more:

Making augmented reality suitable for society

ETH Zurich is establishing a new research hub for augmented reality that involves close collaboration with Google. One of the ETH-Co-heads, Christian Holz, explains the importance of networking in thi...

ethz.ch