Pingchuan Ma

@pima-hyphen.bsky.social

PhD Student at Ommer Lab, Munich (Stable Diffusion) 🎯 Working on getting my first 3M.

Diffusion models treat every part of an image equally. → Same number of steps. Same compute. But images aren’t uniform. 🤔 Some regions are easy, others are hard. So why force the model to treat them the same? 🧵

🤔 What happens when you poke a scene — and your model has to predict how the world moves in response? We built the Flow Poke Transformer (FPT) to model multi-modal scene dynamics from sparse interactions. It learns to predict the 𝘥𝘪𝘴𝘵𝘳𝘪𝘣𝘶𝘵𝘪𝘰𝘯 of motion itself 🧵👇

Bild

I just wrapped up my bidding process for AAAI 26, which is always an enjoyable experience. This year, I came across submissions titled things like “000ASD”, “testy”, “123321241”, and even cases with identical titles and abstracts but different submission numbers.