Logit Refiner: Improving Visual Autoregressive Models via Intra-Scale Dependency Modeling
Paper recorded by Signals 4 on 2026-09-10 in cs.AI. Abstract reproduced from arXiv; link to the original below.
Published 2026-09-10 on arXiv · recorded by Signals 4 on 2026-09-11
Category: cs.AI · 人工智能 · first seen 2026-09-11
Abstract
Visual Autoregressive Models (VAR) generate images through next-scale prediction, producing all tokens within each scale in parallel. We show that this parallel decoding constitutes a mean-field-style approximation that discards spatial dependencies among same-scale tokens, causing locally incoherent samples regardless of backbone capacity -- a limitation of the decoding rule. Addressing this limi
Read on arXiv →
Cite this page: Logit Refiner: Improving Visual Autoregressive Models via Intra-Scale Dependency Modeling: the #116 most recent of 300 cs.AI papers we have recorded (as of 2026-09-10). Source: Signals 4 (Signals API) — https://data.jiangzhang.ca/signals4/t/papers/logit-refiner-improving-visual-autoregressive-models-via-intra-scale-dependency-.html
Free to quote with attribution to “Signals 4 (Signals API)”. Machine-readable:
papers.json
Get 4 AI signals a day by email — free.
Get 4 AI signals a day by email — free