Chang Xu, Hoang Lam, Nguyen, Daochang Liu, Dan Ky, Anh-Dung Dinh, Weidong Cai, Xiuying Wang
We have not lifted any functions out of this paper's repositories yet, so there is nothing we have run. If it links a repository, it is listed below.
Some links come from the archived Papers with Code dataset (CC BY-SA 4.0): attribution and licence.
Autoregressive (AR) models based on nextscale prediction have emerged as a powerful tool for image generation, but they face a critical weakness: information inconsistencies between patches across timesteps introduced by progressive resolution scaling. These inconsistencies scatter guidance signals, causing them to drift away from salient regions within the image and leaving behind ambiguous, unfaithful features during sampling. We tackle this challenge with Information-Grounding Guidance (IGG), a novel framework that anchors guidance to semantically important tokens via an attention-based dynamic weighting formulation, consequently ensuring that guidance and semantic contents remain tightly aligned. Across both class-conditioned and textto-image generation tasks, IGG delivers sharper, more coherent, and semantically grounded images, demonstrating its efficacy for correcting AR-based methods. Our code is available at https://github.com/dnngky/info ground-guidance.
The same record, over MCP at https://syntology.ai/mcp:
get_harvested_code_for_paper("2509.23876")
get_code_for_paper("2509.23876")
have("2509.23876")
Connect an agent — have() is free.