The Missing "I Don't Know": Why Three Reasoning-Reliability Findings Converge on Calibrated Abstention

Three recent results describe what look like unrelated LLM reliability problems. Yin et al. (2026) show reasoning RL collapses tool-reliability…

Three recent results describe what look like unrelated LLM reliability problems. Yin et al. (2026) show reasoning RL collapses tool-reliability representations. Suleymanov et al. (2026) show that under safety-constrained generation, large models rewrite flagged spans while small…

Read the original source — arxiv.org

paper · Shared by tscosj

0 comments

No comments yet.