ReadGlim
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards — ReadGlim