Scalable Chain of Thoughts via Elastic Reasoning Artwork

Deep Papers

Deep Papers is a podcast series featuring deep dives on today’s most important AI papers and research. Hosted by Arize AI founders and engineers, each episode profiles the people and techniques behind cutting-edge breakthroughs in machine learning.

All Episodes

Deep Papers

Scalable Chain of Thoughts via Elastic Reasoning

May 16, 2025 • Arize AI

0:00 | 28:54

In this week's episode, we talk about Elastic Reasoning, a novel framework designed to enhance the efficiency and scalability of large reasoning models by explicitly separating the reasoning process into two distinct phases: thinking and solution.

This separation allows for independent allocation of computational budgets, addressing challenges related to uncontrolled output lengths in real-world deployments with strict resource constraints.

Our discussion explores how Elastic Reasoning contributes to more concise and efficient reasoning, even in unconstrained settings, and its implications for deploying LRMs in resource-limited environments.

Read the paper
Join us live
Read the blog

Learn more about AI observability and evaluation, join the Arize AI Slack community or get the latest on LinkedIn and X.

Dylan Couzon

Host

Parth Shisode

Host