Kale-ab Tessera

@kale-ab.bsky.social

ML PhD Student @ Uni. of Edinburgh, working on Multi-Agent Problems. | Organiser @deeplearningindaba.bsky.socialโ€ฌ @rl-agents-rg.bsky.socialโ€ฌ | ๐Ÿ‡ช๐Ÿ‡น๐Ÿ‡ฟ๐Ÿ‡ฆ kaleabtessera.com

I'm delighted to announce the release of OpenSpiel 2.0! โ™Ÿ๏ธ๐ŸŽฒโ™ฆ๏ธ๐ŸŽ‰ Structured types for states, observations, and actions, standard trajectories (based on JSON), 19 new games, AlphaZero ported to JAX, Windows PyPI support, language model fine-tuning examples and an MCP server (demo below ๐Ÿคฉ๐Ÿ‘‡)! ๐Ÿงต 1/N

OpenSpiel visual asset

Can LLM agents coordinate in long-horizon, open-ended worlds? We test 13 LLMs in Alem, a new benchmark. Most struggle, averaging ~6% normalised return. Yet on Hard, zero-shot Gemini 3.1 Pro matches the best MARL agent after 1B training steps. Our ablations show communication matters most. ๐Ÿงต

As we build more cooperative multi-agent envs, we should ask: are we really testing the properties that make Dec-POMDPs hard, partial observability and decentralised coordination, or can agents succeed via shortcuts? ๐Ÿค” Our #AAMAS2026 Oral probes this question ๐Ÿ”Ž๐Ÿงต

Hello world! This is the RL & Agents Reading Group We organise regular meetings to discuss recent papers in Reinforcement Learning (RL), Multi-Agent RL and related areas (open-ended learning, LLM agents, robotics, etc). Meetings take place online and are open to everyone ๐Ÿ˜Š

๐Ÿ“œ๐Ÿค– Can a shared multi-agent RL policy support both specialised & homogeneous team behaviours -- without changing the learning objective, requiring preset diversity levels or sequential updates? Our preprint โ€œ๐˜๐˜บ๐˜ฑ๐˜ฆ๐˜ณ๐˜”๐˜ˆ๐˜™๐˜“: ๐˜ˆ๐˜ฅ๐˜ข๐˜ฑ๐˜ต๐˜ช๐˜ท๐˜ฆ ๐˜๐˜บ๐˜ฑ๐˜ฆ๐˜ณ๐˜ฏ๐˜ฆ๐˜ต๐˜ธ๐˜ฐ๐˜ณ๐˜ฌ๐˜ด ๐˜ง๐˜ฐ๐˜ณ ๐˜”๐˜ถ๐˜ญ๐˜ต๐˜ช-๐˜ˆ๐˜จ๐˜ฆ๐˜ฏ๐˜ต ๐˜™๐˜“โ€ explores this!

๐Ÿช˜The 2024 Impact Report is here! Our last Indabaโ€™s theme, Xam Xamlรฉ (Wolof for "To Gather Knowledge and Share It"), beautifully reflected our mission: Empowering and educating through African AI. Read our report ๐Ÿ”—https://deeplearningindaba.com/blog/2025/04/xam-xamle-our-latest-indaba-impact-report/

Bild

We recently asked the authors to do a talk on this - m.youtube.com/watch?v=Bmdn... . A lot of awesome gems (goal focused rl research, don't waste your time, etc) ๐Ÿ”ฅ Thanks for bringing this to my attention @brandonrohrer.com !

Adam White - Empirical Design in Reinforcement Learning

YouTube video by UoE Agents Group

m.youtube.com

Brandon Rohrer@brandonrohrer.com ยท 2y ago

Adding my love letter to arxiv.org/pdf/2304.01315 Empirical Design in Reinforcement Learning by Andrew Patterson, Samuel Neumann, Martha White, Adam White JMLR 25 (2024) 1-63 #ReinforcementLearning These arenโ€™t the heroes we deserve, but they are the heroes we need.

The Deep Learning Indaba has made significant progress over the last few years. This piece in @mittechreviewbr.bsky.social paints a great picture of the opportunities and tensions we face ๐Ÿ™๐Ÿพ. And Lots more work for us still ๐Ÿš€๐Ÿš€ www.technologyreview.com/2024/11/11/1...

What Africa needs to do to become a major AI player

Inadequate funding, infrastructure issues, and fights over regulation mean the sector's future remains uncertain.

technologyreview.com

Posting a call for help: does anyone know of a good way to simultaneously treat both POTS and Mรฉniรจreโ€™s disease? Please contact me if youโ€™re either a clinician with experience doing this or a patient who has found a good solution. Context in thread