uncalibrated_reasoning. Code repository for "Uncalibrated Reasoning: GRPO Induces Overconfidence for Stochastic Outcomes"