Papers
arxiv:2605.16817

Adaptive Fused Prior Transfer for Controllable Generative Image Compression

Published on Sep 28
ยท Submitted by
Yifei Pei
on Oct 6
Authors:
,
,

Abstract

Learned image compression achieves competitive rate-distortion performance, but very-low-bitrate reconstruction remains challenging because the transmitted representation cannot preserve fine textures and local structures. Perceptual and generative codecs synthesize missing details using reconstruction priors, while controllable codecs allow one model to cover different bitrate and reconstruction preferences. However, existing codebook-based controllable designs generally rely on single-codebook reconstruction priors. We propose Adaptive Fused Prior Transfer for Controllable Generative Image Compression (AFP-GIC), a controllable codec that transfers an adaptive fused prior from a frozen pretrained AdaCode model. Encoder-side fused-prior features guide latent formation, while the decoder predicts a compatible fused prior from the compressed representation and selected control variables, enabling prior-guided reconstruction without transmitting the fused prior itself. A motivating analysis shows that better decoder-side fused-prior alignment tightens a reconstruction-error upper bound and that the fused-prior family contains single-codebook choices as special cases. Under the unified benchmark, AFP-GIC achieves 18.1% lower decoder latency and uses 31.10 million (20.5%) fewer inference parameters than DC-VIC. Experiments on Kodak, CLIC2020, and DIV2K show competitive PSNR and SSIM, with the clearest perceptual gains in NIQE scores and very-low-bitrate visual comparisons.

Community

๐Ÿš€ AFP-GIC: Controllable Generative Image Compression | IEEE Access 2026

One pretrained model. Five bitrate operating points. No model switching.

167:1 compression in the example shown: a 768ร—512 image becomes a 3.57 KiB bitstream at 0.0744 bpp, without resizing. The ratio compares the source PNG with the bitstream, including headers, and varies by image and source format.

AFP-GIC uses content-adaptive guidance to reconstruct natural-looking details at very low bitrates, without transmitting the fused prior.

Compared with DC-VIC, a state-of-the-art controllable generative image compression model:

  • 18.1% lower decoder latency: 80.47 vs. 98.27 ms.
  • 20.5% fewer inference parameters: a reduction of 31.1M.

Latency measured on an NVIDIA RTX 4090 using 256ร—256 patches.

๐Ÿค— Try your own images: compress, decompress, compare, and download bitstreams and metrics. The public demo runs on CPU.

๐Ÿ’ป Code and pretrained model
๐Ÿ“ฅ 2,760 reconstructions and metrics across Kodak, CLIC2020, and DIV2K for research comparisons.
๐Ÿ“„ Published paper

Find it useful? โญ Star the GitHub repository and ๐Ÿค— like the Hugging Face Space. Your support helps others discover this work!

This is an automated message from the Librarian Bot. I found the following papers similar to this paper.

The following papers were recommended by the Semantic Scholar API

Please give a thumbs up to this comment if you found it helpful!

If you want recommendations for any Paper on Hugging Face checkout this Space

You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend

Sign up or log in to comment

Get this paper in your agent:

hf papers read 2605.16817
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper 1

Datasets citing this paper 0

No dataset linking this paper

Cite arxiv.org/abs/2605.16817 in a dataset README.md to link it from this page.

Spaces citing this paper 1

Collections including this paper 0

No Collection including this paper

Add this paper to a collection to link it from this page.