Skip to content

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Simple Diffusion Language Models Are More Effective Few-Step Generators Than Reported

Diffusion language models (DLMs) promise fast parallel generation, yet high-quality samples often require large number of refinement steps, which diminishes their advantage in practice. This has led to massive interest in and rapid development of new methods for effective few-step generation. We show that much of the s...

H. Amin, Ming Yin, Rajiv Khanna · 0 citations
#machine learning Preprint Sep 2026

When Does Scale-Invariant Optimization Become Unstable? An Exact Schedule Law with Weight Decay

Normalization renders large parts of neural networks effectively scale invariant, inducing a hidden feedback loop in which learning-rate schedules and weight decay interact through the parameter norm to control the effective step taken by the optimizer. We show that this interaction is governed by an exact discrete-tim...

H. Amin, Wei-Kai Chang, Rajiv Khanna · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.