AI Trend Notifier
EN한
← wiki

$ cat wiki/papers/2026/2610.05076-feedback-ttt.md

Self-Generated Feedback Destabilizes Test-Time Training: A Causal Decomposition of Long-Horizon Adaptation

paperupdated 2026-10-07created 2026-10-07

TL;DR

Updating a model on its own generated text can worsen prediction on independent human text. The authors trace a feedback loop between the changing generator and the updates retained in its weights, and test an independent-evidence gate before committing updates. (source)

Authors & Org

Cheng Luo, Bing Li and Bernard Ghanem. Affiliations are not established by the abstract page. Submitted 2026-10-04; read scope is the abstract and bibliographic record. (source)

Method

Across 128K-token streams, the study tests TTT-E2E configurations labeled 125M, 760M and 3B, plus Adam updates to Qwen3-4B's existing weights. Fixed Generation freezes the model producing training chunks; Recorded Replay separates degraded-input effects from damage retained through updates. A paired one-update comparison tests source-text fit against new real text. (source)

Results

Fixed Generation removes over 98% of the damage at 125M and 760M. Updates can fit their own source better while predicting new real text worse; the same mechanisms can improve on real text, so updating itself is not the failure. (source)

Settlement checks a candidate state against independent real text before committing it. It leaves mean endpoint gaps of 0.07 and -0.02 nats at 125M and 760M, while retaining real-text adaptation. These results are not reported for every tested model in the abstract. (source)

Significance

This qualifies a neighboring mechanism to Test-Time Compute (Inference-Time Compute Scaling): test-time training changes weights, whereas the concept page's core definition keeps them fixed. Interpretation: spending more inference-time compute and retaining self-generated updates are different interventions and need different controls. (source)

Open Questions

The abstract does not establish the cost of Settlement or performance at frontier scale. It does not show that all synthetic-data training fails, and the reported 98% mitigation is limited to the two named configurations. (source)

Cite

Cheng Luo, Bing Li, Bernard Ghanem. “Self-Generated Feedback Destabilizes Test-Time Training: A Causal Decomposition of Long-Horizon Adaptation.” arXiv:2610.05076, 2026. arXiv.

Referenced by

Sources