Noah Snavely

@snavely.bsky.social

3D vision fanatic http://snavely.io

Sean Bell, Kavita Bala, and I received a SIGGRAPH Test-of-Time Award for our 12-year-old paper, Intrinsic Images in the Wild (IIW -- I like to call it "Double-I Double-U", but that name didn't catch on).

Xiangli et al., "Honey, I Shrunk the Arc de Triomphe!" Metric depth estimators aren't actually metric. With curated, scaled data, they can be adapted to be better.

Humans can watch tasks like cooking or assembly and reason about what happened, when, and between which parts. Can LVLMs do the same? We built Flat-Pack Bench to test this – and found there is still a long way to go. Accepted at #CVPR2026! 🎥🪑🧩(1/n)

Very excited about @nthngdy.bsky.social's new work! It really gets to the bottom (or top, depends where the head in LMs is 😜) and fundamentals of contemporary LLMs. A real treat of a paper: solid theory, and very cool experiments.

Nathan Godey@nthngdy.bsky.social · 5mo ago

🧵New paper: "Lost in Backpropagation: The LM Head is a Gradient Bottleneck" The output layer of LLMs destroys 95-99% of your training signal during backpropagation, and this significantly slows down pretraining 👇

Attention-grabbing idea for an academic paper: Somehow work your personal phone number into the title, like those old LifeLock ads featuring the CEO's social security number. I bet this kind of stunt would garner attention, but I haven't figured how to work it naturally into a CVPR paper.

I thought Nirvanna the Band the Show the Movie was a fantastic comedy film! I went in knowing nothing about the premise and had a great time. I have no idea how they filmed parts of it, especially on a budget of just $2M. A nice time in the theater.

The CVPR process seems driven by the goal of extracting 3 reviews for each paper, a goal that seems to lead to a lot of angst. Is there a reason why 3 is a magic number? Why not two, or even one (assuming the quality is high)?