
HV
Ha Van Dau, Thanh Tung Khuat, Nguyen Thanh Dung
· 1 min read
ResearcharXiv cs.AI
Beyond Depth Truncation: Controlled Evaluation of Depth Utilization in Recursive Language Models
arXiv:2609.19934v1 Announce Type: new
Abstract: Depth-recurrent language models iteratively apply a small layer stack, decoupling per-token compute from distinct parameter count. To determine whether such a model genuinely utilizes its depth, both recurrence and layer-pruning literatures rely on a shared evaluation: truncating depth at inference time, plotting quality against retained depth fraction, and reading off the slope. While cheap and training-free, this metric suffers from an unexamined flaw: it extracts a single scalar from an intervention that alters multiple model properties simultaneously. Depth truncation concurrently reduces the number of block applications, decreases the volume of distinct computation performed, and pushes the readout head onto an out-of-distribution residual stream. The observed slope conflates all three factors, yet is conventionally interpreted as reflecting solely the second.
We propose the Depth Control Protocol (DCP), a diagnostic suite that disentangles these three quantities. DCP comprises three positive controls that isolate each factor while varying the others, a negative control applying the identical interventions to dense transformers to ensure the effect is not an artifact of the measurement protocol, and a controlled training intervention to verify causality. The linchpin control, running the full budget of block applications while executing only a single distinct iteration, is strictly realizable only in depth-wise weight-sharing architectures, since in a dense network repeating a layer yields an entirely different model rather than the same model in an alternative configuration.
Original source
This story was published by arXiv cs.AI and written by Ha Van Dau, Thanh Tung Khuat, Nguyen Thanh Dung. SyncAI.news shows a preview; the complete article is on the publisher's site.
Read the full story on arxiv.org


