Papers Discussed ETR: Entropy Trend Reward for Efficient Chain-of-Thought Reasoning Reason Only When Needed: Efficient Generative Reward Modeling via Model-Internal Uncertainty Purging the Gray Zone: Latent-Geometric Denoising for Precise Knowledge Boundary Awareness