arXiv (OAI)open access
RL-PLUS: Countering Capability Boundary Collapse of LLMs in Reinforcement Learning with Hybrid-policy Optimization
artificial intelligence, computation and language, machine learning
This document is indexed with metadata only — full text is not available in the archive for this record.
Open the official source ↗
Related documents
Record · ID 144336
Retrieved via
Conceptio — every document is proof-bundled with source, license, and retrieval metadata.