🚀 We are hiring! 🚀 🔍 Join us as a Postdoctoral Researcher (fully-funded) at the Helmholtz Institute for Human-Centered AI in Munich.
@lucaschubu.bsky.social
Excited to see our Centaur project out in @nature.com. TL;DR: Centaur is a computational model that predicts and simulates human behavior for any experiment described in natural language.
Excited to say our paper got accepted to ICML! We added new findings including this: models fine-tuned on a visual counterfactual reasoning task do not generalize to the underlying factual physical reasoning task, even with test images matched to the fine-tuning data set.
In previous work we found that VLMs fall short of human visual cognition. To make them better, we fine-tuned them on visual cognition tasks. We find that while this improves performance on the fine-tuning task, it does not lead to models that generalize to other related tasks:
In previous work we found that VLMs fall short of human visual cognition. To make them better, we fine-tuned them on visual cognition tasks. We find that while this improves performance on the fine-tuning task, it does not lead to models that generalize to other related tasks: