Keshav Ramji

@keshavramji.bsky.social

working toward continually self-improving AI, reasoning and alignment @ IBM Research AI | Prev: Penn CS + Wharton

Does a LLM really need to think in English or Chinese? How about it thinks using a short sequence of reserved "abstract" tokens through reinforcement learning? They find out that it is as performant as verbalized CoT at a fraction of the cost, achieving major gains in inference-time efficiency.

Bild