ReadGlim
Beyond Entropy: Learning from Token-Level Distributional Deviations for LLM Reasoning — ReadGlim