Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Mitigating Premature Discretization with Progressive Quantization for Robust Vector Tokenization

About

Vector Quantization (VQ) has become the cornerstone of tokenization for many multimodal Large Language Models and diffusion synthesis. However, existing VQ paradigms suffer from a fundamental conflict: they enforce discretization before the encoder has captured the underlying data manifold. We term this phenomenon Premature Discretization. To resolve this, we propose Progressive Quantization (ProVQ), which incorporates the dynamics of quantization hardness as a fundamental yet previously overlooked axis in VQ training. By treating quantization as a curriculum that smoothly anneals from a continuous latent space to a discrete one, ProVQ effectively guides the codebook toward the well-expanded manifolds. Extensive experimental results demonstrate the broad effectiveness of ProVQ across diverse modalities. We report improved reconstruction and generative performance on the ImageNet-1K and ImageNet-100 benchmarks, highlighting the ProVQ's boost for generative modeling. Furthermore, ProVQ proves highly effective for modeling complex biological sequences, establishing a new performance ceiling for protein structure tokenization on the StrutTokenBench leaderboard.

Wenhao Zhao, Qiran Zou, Zhouhan Lin, Dianbo Liu• 2026

Related benchmarks

TaskDatasetResultRank
Image GenerationImageNet-1K 256x256 (val)
Inception Score235.5
113
Functional Site PredictionStructTokenBench (SupFam)
BindInt AUROC91.55
7
Functional Site PredictionStructTokenBench All Functional Splits
Average AUROC72.62
7
Physiochemical Property PredictionStructTokenBench (Fold)
Spearman's rho (RMSF)44.94
7
Physiochemical Property PredictionStructTokenBench (SupFam)
FlexRMSF (Spearman's ρ)41.28
7
Physiochemical Property PredictionStructTokenBench All Physiochemical Splits
Average Spearman's Rho (%)38.88
7
Structure Property PredictionStructTokenBench (Fold)
Homo Macro F138.21
7
Structure Property PredictionStructTokenBench (SupFam)
Homo Macro F141.49
7
Structure Property PredictionStructTokenBench (Fam)
Fam Macro F187.39
7
Structure Property PredictionStructTokenBench (All Structure Splits)
Macro F155.7
7
Showing 10 of 12 rows

Other info

Follow for update