@jacobaustin123.bsky.social

Researcher at Google DeepMind. I make LLMs go fast. I also play piano and climb sometimes. Opinions my own

Training our most capable Gemini models relies heavily on our JAX software stack+Google's TPU hardware platforms. If you want to learn more, see this awesome book "How to Scale Your Model": jax-ml.github.io/scaling-book/ Put together by several of my Google DeepMind colleagues listed below 🎉.

@jacobaustin123.bsky.social · 2y ago

Making LLMs run efficiently can feel scary, but scaling isn’t magic, it’s math! We wanted to demystify the “systems view” of LLMs and wrote a little textbook called “How To Scale Your Model” which we’re releasing today. 1/n

Making LLMs run efficiently can feel scary, but scaling isn’t magic, it’s math! We wanted to demystify the “systems view” of LLMs and wrote a little textbook called “How To Scale Your Model” which we’re releasing today. 1/n

Bild