Thanks for sharing our work! Please check it out -- we have a Twitter (X) thread as well, bsky thread coming soon!
Does a LLM really need to think in English or Chinese? How about it thinks using a short sequence of reserved "abstract" tokens through reinforcement learning? They find out that it is as performant as verbalized CoT at a fraction of the cost, achieving major gains in inference-time efficiency.