Reasoning Risks and Production Controls
Overthinking, prompt injection and agent hijacking in reasoning models with tools. Design criteria for bounding risk in production.
Links containing ?t= open the video at a specific second.
Video summary
The ideas to retain
1. Overthinking: when more reasoning degrades the answer
The intuition that more reasoning always produces better answers is incorrect. There is a documented phenomenon in reasoning models called overthinking (Apple Research, 2025): the model…
2. Quality vs cost vs latency in a real product
The three-way tension that every generative-AI system has to manage becomes sharper in reasoning models.
The debt profile
In models without extended reasoning, quality debt is paid relatively uniformly: the model works well on the distribution of cases it was trained for and fails on cases outside that…


