Signals 4 · free daily AI digest

Every Ablation Is a Dose: Counterweights and the Semblance of Self-Repair

Paper recorded by Signals 4 on 2026-10-01 in cs.LG. Abstract reproduced from arXiv; link to the original below.

Published 2026-10-01 on arXiv · recorded by Signals 4 on 2026-10-02

Category: cs.LG · 机器学习 · first seen 2026-10-02

Abstract

Ablate a component of a language model, and other components often appear to adjust and compensate. This phenomenon, termed self-repair, has been observed repeatedly, but its mechanism remains unclear. The most systematic study to date concluded that self-repair is noisy and unlikely to have a single explanation. We argue that it has one: a gain already present before any ablation. Any interventio

Read on arXiv →

#10 most recent of 362 cs.LG papers we have recorded · ↑ newer: Effective Resistance and Graph Neural Network Reliability in Tissue-Sp · ↓ older: When Do Intrinsic Rewards Lead to Exploration?
Cite this page: Every Ablation Is a Dose: Counterweights and the Semblance of Self-Repair: the #10 most recent of 362 cs.LG papers we have recorded (as of 2026-10-01). Source: Signals 4 (Signals API) — https://data.jiangzhang.ca/signals4/t/papers/every-ablation-is-a-dose-counterweights-and-the-semblance-of-self-repair.html
Free to quote with attribution to “Signals 4 (Signals API)”. Machine-readable: papers.json
Related: More cs.LG papers · arXiv signals · All papers · Today in AI
Get 4 AI signals a day by email — free.
Subscribe free → See all plans →
Get 4 AI signals a day by email — free
All models · All repos · By company · Daily editions