The  LGTM
  • Home
  • Agentic Coding
  • Claude Code
  • Codex
Sign in Subscribe

model post-training

A collection of 1 post
d-OPSD Gives Diffusion LLMs a Post-Training Recipe That Does Not Pretend They Decode Left-to-Right
ai-models

d-OPSD Gives Diffusion LLMs a Post-Training Recipe That Does Not Pretend They Decode Left-to-Right

Diffusion language models keep getting discussed as if they are autoregressive models with a different decoding costume. That shortcut is convenient, but it breaks down the moment you try to post-train them seriously. d-OPSD, short for on-policy self-distillation for diffusion LLMs, is worth covering because it starts from the obvious-but-often-ignored
17 Jun 2026 3 min read
Page 1 of 1
The LGTM © 2026
  • Sign up
Powered by Ghost