Kimon Fountoulakis

@kfountou.bsky.social

Associate Professor at CS UWaterloo Machine Learning Lab: opallab.ca

The SIAM Conference on Optimization 2026 will be in Edinburgh! I don’t really work on optimization anymore (at least not directly), but it’s cool to see a major optimization conference taking place where I did my PhD.

Bild

Our new work on scaling laws that includes compute, model size, and number of samples. The analysis involves an extremely fine-grained analysis of online sgd built up over the last 8 years of understanding sgd on simple toy models (tensors, single index models, multi index model)

Eshaan Nichani@eshaannichani.bsky.social · last yr.

Excited to announce a new paper with Yunwei Ren, Denny Wu, @jasondeanlee.bsky.social! We prove a neural scaling law in the SGD learning of extensive width two-layer neural networks. arxiv.org/abs/2504.19983 🧵below (1/10)

ChatGPT gives me the ability to expand my search capabilities on topics that I can only roughly describe, or even illustrate with a figure, when I don’t know the exact keywords to use in a Google search.

I enjoyed reading the paper "A Generalized Neural Tangent Kernel for Surrogate Gradient Learning" (Spotlight, NeurIPS 2024). They extend the NTK framework to activation functions that have finitely many jumps.

BildBild

Shenghao's Ph.D Thesis "Perspectives of Graph Diffusion: Computation, Local Partitioning, Statistical Recovery, and Applications" is now available. Link: dspacemainprd01.lib.uwaterloo.ca/server/api/c... Relevant papers: 1) Local Graph Clustering with Noisy Labels (ICLR 2024)

Bild
Kimon Fountoulakis@kfountou.bsky.social · 2y ago

. Shenghao Yang passed his PhD defence today. Shenghao is the second PhD student to graduate from our group. I am very happy for Shenghao and the work that he has done!

GLOW is returning on 𝗠𝗮𝗿𝗰𝗵 𝟮𝟲𝘁𝗵, 𝟱𝗽𝗺 𝗖𝗘𝗧 with a special guest: @petar-v.bsky.social 🌟 He will lecture on LLMs as GNNs – a topic which received quite some attention at our last session. Specifically, we will learn how Graph ML tools can help understand LLM generalisation

Every time I try to use chatGPT or Gemini to write some text for me, I always get vague results. This is not useful. It can't be creative. I don't know why people talk about it as if it actually produces knowledge. Even if I initially think it's helpful, I end up rewriting the entire text because,