Nils Trost

@trostnils.bsky.social

LLMs can memorize even a phone number seen once in training.🔒 Google’s VaultGemma fixes that, being the first open-weight LLM trained from scratch with differential privacy, so rare secrets leave no trace. ☕ new video explaining Differential Privacy through VaultGemma 👇 🎥 youtu.be/UwX5zzjwb_g

What's up with Google's new VaultGemma model? – Differential Privacy explained

YouTube video by AI Coffee Break with Letitia

youtu.be

Ever wondered how Energy-Based Models (EBMs) work and how they differ from normal neural networks? ☕️ We go over EBMs and then dive into the Energy-Based Transformers paper to make LLMs that refine guesses, self-verify, and could adapt compute to problem difficulty.

Bild

🤖 Can we trust AI in science? I'm excited to be speaking at the final event of the Young Marsilius Fellows 2025, themed "Dancing with Right & Wrong?" – a title that feels increasingly relevant these days. I'll be joining a panel on "(How) can we trust AI in science?" to discuss questions like:

I'm very excited to finally share the main work of my PhD! We explored the evolutionary dynamics of gene regulation and expression during gonad development in primates. We cover among others: X chromosome dynamics (incl. in a developing XXY testis), gene regulatory networks and cell type evolution.

Kaessmann Lab@kaessmannlab.bsky.social · last yr.

We are delighted to share our new preprint “The evolution of gene regulatory programs controlling gonadal development in primates” www.biorxiv.org/content/10.1...

Excited to share that I’ll be joining the Summer School “AI and Human Values” this September at the Marsilius-Kolleg of Heidelberg University as a speaker. I'll be giving an introduction to how large language models actually work—before the summer school dives deeper into broader implications.

Bild

Long videos are a nightmare for language models—too many tokens, slow inference. ☠️ We explain STORM ⛈️, a new architecture that improves long video LLMs using Mamba layers and token compression. Reaches better accuracy than GPT-4o on benchmarks and up to 8× more efficiency. 📺 youtu.be/uMk3VN4S8TQ

Token-Efficient Long Video Understanding for Multimodal LLMs | Paper explained

YouTube video by AI Coffee Break with Letitia

youtu.be

New video about: REPA (Representation Alignment), a clever trick to align diffusion transformers’ representations with pretrained transformers like DINOv2. It accelerates training and improves the diff. model’s ability to do things other than image generation (like image classification).

REPA Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You ...

YouTube video by AI Coffee Break with Letitia

youtu.be