Brandon Amos

@bdamos.bsky.social

🧙🏻‍♀️ scientist at Meta NYC | http://bamos.github.io

Looking for a principled evaluation method for ranking of *general* agents or models, i.e. that get evaluated across a myriad of different tasks? I’m delighted to tell you about our new paper, Soft Condorcet Optimization (SCO) for Ranking of General Agents, to be presented at AAMAS 2025! 🧵 1/N