Limits of Confidence in Diffusion
AI Classified by Officially
AuthorsRuss Webb, Amitis Shidani, Alice Bizeul, Dan Busbridge
Discrete diffusion, including remasking and uniform-state samplers, generate a sequence by writing multiple token positions per step, drawing each from a per-position distribution and choosing which positions to write from those same distributions. For domains of general interest (pixels, phonemes, or words) there are inherent dependencies between tokens. We show that a step matches the training distribution only when the positions it writes are conditionally independent given the tokens already fixed, that no product of per-position distributions can match a dependent group, and that per-position distributions do not determine whether a group is dependent: two joint distributions can have identical per-position marginals while differing in which combinations of values occur. On ScanAndAdd, a synthetic task whose joint distribution is available in closed form, we verify that every group of two or more undetermined positions a confidence ranking writes is dependent, and measure the generated distribution to be 29× the sampling-noise floor total variation while per-sample metrics are 1.0.
Unmasking On-Policy Distillation: Where It Helps, Where It Hurts, and Why
On-policy distillation offers dense, per-token supervision for training reasoning models; however, it remains unclear under which conditions this signal is beneficial and under which it is detrimental. Which teacher model should be used, and in the case of self-distillation, which specific context should serve as the supervisory signal? Does the optimal choice vary from one token to the next? At present, addressing these questions typically…
Position Prediction as an Effective Pre-training Strategy
This is an extract. The publication continues at the source.
Read the original at the source: https://machinelearning.apple.com/research/limits-confidence-diffusion
Officially imported this from Apple Machine Learning Research’s own source and shows an extract. If you work there, claiming the profile and verifying the domain lets you choose to show the full text here.
Provenance
- Organization
- Apple Machine Learning Research — imported from official source
- Official source
- https://machinelearning.apple.com/rss.xml RSS
- Imported
- October 02, 2026 21:00
- Versions
- 1 recorded
- Identity
limits-confidence-diffusion