Topic
large reasoning models
Artificial Intelligence #artificial intelligence#large reasoning models
Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models
A new research paper from arXiv shows that reinforcement learning with verifiable rewards (RLVR) can cause large reasoning models to forget foundational capabilities like perception and faithfulness. The authors propose RECAP, a replay strategy with dynamic objective reweighting that preserves general knowledge while maintaining reasoning gains.
Jun 21, 2026 1 source
Artificial Intelligence #artificial intelligence#ai safety
Adaptive and Explicit safe: Triggering Latent Safety Awareness in Large Reasoning Models
A new method called Safe Trigger leverages the latent safety awareness of Large Reasoning Models to improve safety alignment without external data. Using Supervised Fine-Tuning and Direct Preference Optimization, the approach reduces Attack Success Rate on harmful and jailbreak benchmarks while preserving general performance.
Jun 16, 2026 1 source