Yash Bhalgat

@ysbhalgat.bsky.social

PhD at VGG, Oxford w/ Andrew Zisserman, Andrea Vedaldi, Joao Henriques, Iro Laina. Past: Senior RS Qualcomm #AI #Research, UMich, IIT Bombay. I occasionally post AI memes. yashbhalgat.github.io

"LLaDA: Large Language Diffusion Models" Nie et al. Just read this fascinating paper. Scaled up Masked Diffusion Language Models to 8B params, and show that it can match #LLMs (including Llama 3) while solving some key limitations! Let's dive in... 🧵 (1/8) #genai

Bild

Need to rig 3D models? 🦖 New work from UCSD and Adobe: "RigAnything: Template-Free Autoregressive Rigging for Diverse 3D Assets" Liu et al. tl;dr: reduces rigging time from 2 mins to 2 secs, works on any shape category & doesn't need predefined templates! 🚀

"Latent Radiance Fields with 3D-aware 2D Representations" Zhou et al., #ICLR2025 tl;dr: Novel framework that integrates 3D awareness into VAE latent space using correspondence-aware encoding, enabling high-quality rendered images with ~50% memory savings. (1/n) 🧵

"EdgeRunner" (#ICLR2025) from #Nvidia & PKU introduces an auto-regressive auto-encoder for mesh generation, supporting up to 4000 faces at 512³ resolution. 🤩 Their mesh tokenization algorithm (adapted from EdgeBreaker) achieves ~50% compression (4-5 tokens per face vs 9), making training efficient.

🚨🚨🚨 Reminder: closing in 3 weeks time 🚨🚨🚨 Please re-post! Note: Oxford recruits faculty at Associate Professor level - we have no Assistant Professor level.

Maurice Fallon@mauricefallon.bsky.social · 2y ago

Multiple faculty positions at University of Oxford in @oxengsci.bsky.social - Join Us! Recruiting for 3 Information Engineering faculty - including Robotics, Computer Vision, Machine Learning. Please repost! Faculty positions in Oxford are typically linked to a college. ⬇️ details in thread ⬇️

"NeuralSVG: An Implicit Representation for Text-to-Vector Generation" (1/2) Encodes SVGs as implicit neural representations using a small MLP trained with Score Distillation Sampling (SDS). Maps 2D coordinates to shape/color outputs. Dropout-like technique ensures ordered, layered structures.

BildBildBild

"AR4D: Autoregressive 4D Generation from Monocular Videos" *without* SDS. Autoregressively generate "3D frames" (aka 3DGS) starting from a canonical space, and using a local deformation field for each frame -- high-quality prompt-aligned generations. #ai #nerf #GenAI #video

Bild

Another gem from Bill Freeman, Katie Bouman & team 🌌 A differentiable rendering framework for direct #exoplanet imaging, leveraging wavefront sensing to refine starlight subtraction. Tested on JWST, it approaches noise limits and reveals faint planets like never before! 🚀 #ai #astronomy

BildBild