Thomas Kipf

@tkipf.bsky.social

Research at Google DeepMind. Ex-Physicist. Controllable World Simulators (GNNs, Structured World Models, Neural Assets). TLM Veo Capabilities (Ingredients & more). 📍 San Francisco, CA

Recordings of the NeSy 2025 keynotes are now available! 🎥 Check out insightful talks from @guyvdb.bsky.social, @tkipf.bsky.social and D McGuinness on our new Youtube channel www.youtube.com/@NeSyconfere... Topics include using symbolic reasoning for LLM, and object-centric representations!

NeSy conference

The NeSy conference studies the integration of deep learning and symbolic AI, combining neural network-based statistical machine learning with knowledge representation and reasoning from symbolic appr...

youtube.com

Two life updates: 1) About a year ago I decided to join the Veo team to work on capabilities. It’s been a fun ride! Excited for what’s still to come. 2) I've been busy caring for a newborn the past couple of days 🥰 Excited for the incredible world he will grow up in. Veo's impression below:

Hot take: 90% of what ACs/SACs do could in principle already be automated (with the remaining 10% being process oversight and borderline decision making). At least right now, it seems like reviewers have the more important job for the most part.

Hilde Kuehne@hildekuehne.bsky.social · 2y ago

No worries. I also got a reviewer invite… I can highly recommend it… I already decided for ICLR to review instead of ACing. It’s actually a nice break. And people always forget that in the end it’s reviewers who are making calls on papers, not ACs.

Waymo deserves to be the number one tourist attraction in San Francisco right now, and it's not even close For like ~$11 you get to ride in a genuine self-driving car with up to four people! Wildly entertaining

Thrilled to announce Boltz-1, the first open-source and commercially available model to achieve AlphaFold3-level accuracy on biomolecular structure prediction! An exciting collaboration with Jeremy, Saro, and an amazing team at MIT and Genesis Therapeutics. A thread!

Bild

The world doesn’t live on a pixel grid and neither should vision models! Excited to share Moving off-the-Grid (MooG): a video model w/o grid-based representations. MooG learns detached “off-the-grid tokens” that bind to (and track) scene elements as camera & content move. 🧵