Interactive Demo · Sudoku Hard
Visualizing Diffusion Language Model Reasoning
Comparing reasoning performance under our modified training objective (noised prompt) vs default training objective.
step 0 / 0
Sampled from a random set of 128 puzzles. Both models use the same
seeded noise for fair comparison. Puzzle turns green when correct, red
highlighted cells are incorrect cells.