Samuel Müller

@sammuller.bsky.social

(Tab)PFNs, TrivialAugment etc.

The recent days have been horrific. We can't become numb to repeated instances of illegal and unconstitutional action by government agencies. It's even worse when public officials are blatantly lying in ways that contradict dozens of pieces of video evidence.

Aaron Rupar@atrupar.com · 7mo ago

Sen. Lankford blatantly lies about what the video of Renee Good's killing shows: "A classic law enforcement moment -- they have to fire their weapon and then when you see her car crash, law enforcement is running to her to provide aid. They're never looking to be able to take a life of individuals."

Compute is increasing much faster than data. How can we improve classical supervised learning long term (the underlying tech of most of GenAI)? Our ICML position paper's answer: simply train on a bunch of artificial data (noise) and only do inference on real-world data! 1/n

MiniMax-01 takeaways - 7 of 8 layers are linear att - implemented a flash-variant of linear attention + ring-att - post-norm is back in large models! (using deepnorm) - prob. wrong scaling laws, as lr schedule is not adapted (see Chinchilla) Let's see how it fares in the arena!

Los modelos preentrenados para datos tabulares (TabPFN) podrían ser el nuevo state of the art para regresión y clasificación. 🙄 Habrá que probarlo. El GitHub al final del hilo. Éste en concreto es enorme y si se comporta como dicen es un gran salto adelante en el state of the art del campo.

Samuel Müller@sammuller.bsky.social · 2y ago

This might be the first time after 10 years that boosted trees are not the best default choice when working with data in tables. Instead a pre-trained neural network is, the new TabPFN, as we just published in Nature 🎉

Groundbreaking work, congrats to the team!! 🎉 When I started my PhD 3 years ago, our tabular benchmark showed tree-based models miles ahead of neural networks. On the same benchmark, TabPFN v2 now reaches in 10s what CatBoost achieves in 4h of tuning 🤯

Bild
Samuel Müller@sammuller.bsky.social · 2y ago

This might be the first time after 10 years that boosted trees are not the best default choice when working with data in tables. Instead a pre-trained neural network is, the new TabPFN, as we just published in Nature 🎉

This might be the first time after 10 years that boosted trees are not the best default choice when working with data in tables. Instead a pre-trained neural network is, the new TabPFN, as we just published in Nature 🎉

Bild