Alberto Cazzaniga

@albecazzaniga.bsky.social

Geometry and deep learning @areasciencepark

🔥 Two PhD positions open @UniTrieste funded by @AreaSciencePark! 🔥 Join the Laboratory of Data Engineering to advance research in AI and its scientific applications. We’re looking for motivated students ready to dive into interdisciplinary research in deep learning and AI.

Really excited to share our latest interpretability work on multimodal models! The communication between image and text is localised in a single token in multimodal-output vision-language models. Paper: arxiv.org/html/2412.0664… Happy to discuss it at #NeurIPS2024 More below 👇

arxiv.org

Francesco Ortu@francescortu.bsky.social · 2y ago

🚨 🚨 Excited to share our latest paper, now on #arXiv! 🖼️ We studied how unified VLMs, trained to generate both text and images (e.g., Meta's Chameleon), exchange information between modalities, comparing them to standard VLMs. 📄 Paper: arxiv.org/abs/2412.06646 Deep dive: 👇

We will present our work tomorrow on "The representation landscape of few-shot learning and fine-tuning in LLMs" #NeurIPS2024 Poster Session East 1 Wednesday h. 11-14 Number #3303 Great work with @diegodoimo.bsky.social @alexpietroserra.bsky.social @ansuin arxiv.org/abs/2409.03662 More 👇

The representation landscape of few-shot learning and fine-tuning in large language models

In-context learning (ICL) and supervised fine-tuning (SFT) are two common strategies for improving the performance of modern large language models (LLMs) on specific tasks. Despite their different nat...

arxiv.org

@diegodoimo.bsky.social · 2y ago

Just landed in Vancouver to present @neuripsconf.bsky.social the results of our new work! Few-shot learning and fine-tuning change the layers inside LLMs in a dramatically different way, even when they perform equally well on multiple-choice question-answering tasks. 🧵1/6