Linyang He

@linyanghe.bsky.social

PhD Student @ Mesgarani Lab, @zuckermanbrain.bsky.social, Columbia University Human Intelligence&Machine Intelligence https://linyanghe.github.io/

🌍Introducing BabyBabelLM: A Multilingual Benchmark of Developmentally Plausible Training Data! LLMs learn from vastly more data than humans ever experience. BabyLM challenges this paradigm by focusing on developmentally plausible data We extend this effort to 45 new languages!

Bild

In our new paper, we explore how we can build encoding models that are both powerful and understandable. Our model uses an LLM to answer 35 questions about a sentence's content. The answers linearly contribute to our prediction of how the brain will respond to that sentence. 1/6