It’s Tokens All The Way Down: How RLMs are Different — Kevin Madura, AlixPartners
Kevin Madura from AlixPartners presents recursive language models (RLM) — a method to make AI models work more like programmers than language processors.
Instead of attempting to read through all tokens directly, an RLM can treat its content as variables in a code environment, write Python code to solve problems, and even delegate to itself or other models to break down difficult tasks. On a benchmark for long reasoning, accuracy improved from 2.6 percent to 45.4 percent, particularly for tasks that become straightforward when written as code. He demonstrates practical examples such as analyzing customer data, summarizing long invoices, and finding patterns in logs — all without needing to break up the text into pieces.
Vibekollen prepared this summary with AI from the original publication. The content belongs to AI Engineer.