Conceptio
›
Archive
›
arXiv (All)
arXiv (All)
open access
An evidence-guided reinforcement learning method to improve psychiatric reasoning in small language models
Lin, Xinxin et al.
arXiv (All) · Papers · License: Open Access
Open Source ↗
Direct PDF ↓
computation-and-language
i-2-7
computation and language, i.2.7
This document is indexed with metadata only — full text is not available in the archive for this record.
Open the official source ↗
Related documents
RPMem: Learning Long-Term Recurrent Parametric Memory Across Sessions for LLM Agents
#1052293
Are Human-Aligned Models Models of Humans? A Turing-Test Gap in Preference Alignment
#1052296
Record
· ID 798761
Retrieved via
Conceptio
— every document is proof-bundled with source, license, and retrieval metadata.