Three recent results describe what look like unrelated LLM reliability problems. Yin et al. (2026) show reasoning RL collapses tool-reliability representations. Suleymanov et al. (2026) show that under safety-constrained generation, large models rewrite flagged spans while small…
Read the original source — arxiv.org
paper · Shared by tscosj
0 comments
No comments yet.