Conceptio
›
language-policy
Topic
language-policy
Knowledge-graph topic
· documents ABOUT language-policy across the archive
23
Documents about language-policy
Documents about language-policy
Policy-Grounded Safety Evaluation of 20 Large Language Models
#183283
arXiv (OAI)
Same Evidence, Different Answers: Canonical-Context On-Policy Distillation for Multi-Turn Language Models
#238655
arXiv CS
DanceOPD: On-Policy Generative Field Distillation
#614110
arXiv (OAI Expanded)
DanceOPD: On-Policy Generative Field Distillation
#615077
arXiv (All)
Personalized Group Relative Policy Optimization for Heterogenous Preference Alignment
#799968
arXiv (All)
Local Edits, Global Ripples: Replay-Informed Policy Adaptation for Workflow Synthesis
#1000094
arXiv (All)
EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents
#1011363
arXiv (All)
STOP: Structured On-Policy Pruning of Long-Form Reasoning in Low-Data Regimes
#1048653
arXiv (All)
Distill What You Trust: Reliability-Aware Multi-Teacher On-Policy Distillation
#1049433
arXiv (All)
Review: Privacy-Preservation in the Context of Natural Language Processing
#400642
OpenAlex
Law no. 3465/1989 on the functioning of languages in the Moldavian SSR: fundamentals, impact and applicability in higher education
#220337
Zenodo (CERN)
Global flood research patterns, disparities, and risk–attention mismatches revealed by large language models
#253449
Figshare
X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment
#394393
arXiv CS
Call Neighbours Yourself: Graph Walks with Destination-Conditioned On-Policy Self-Distillation
#796815
arXiv (OAI Expanded)
PA3: Policy-Aware Agent Alignment through Chain-of-Thought
#797385
arXiv (OAI Expanded)
Call Neighbours Yourself: Graph Walks with Destination-Conditioned On-Policy Self-Distillation
#798228
arXiv (All)
PA3: Policy-Aware Agent Alignment through Chain-of-Thought
#798799
arXiv (All)
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement
#809084
arXiv (OAI Expanded)
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement
#819076
arXiv (All)
HiDiffTIR: Hierarchical Difficulty-Aware Policy Optimization for Multi-Turn Tool-Integrated Reasoning
#820626
arXiv (All)
Rethinking RL for LLM Reasoning: It's Sparse Policy Selection, Not Capability Learning
#973163
arXiv (All)
AgentServeSim: Serving-System Simulation and Policy Search for LLM Agent Programs
#973212
arXiv (All)
SOCIOLINGUISTIC STUDY OF THE USE OF VARIETY OF TRADERS' LANGUAGE IN BUYING AND BUYING TRANSACTIONS
#1003502
KABILAH : Journal of Social Community
Patch the Distribution Mismatch: RL Rewriting Agent for Stable Off-Policy SFT
#1011042
arXiv (All)
EAVer: Long-Form Factuality Verification as an End-to-End Agentic Policy
#1036906
arXiv (All)
English as a Medium of Instruction in Algerian Higher Education: Students’ Attitudes towards Learning using English in Blida 2 University.
#197039
OSF
Cost Effectiveness of Language Services in Hospital Emergency Departments (EDs)
#1007716
ClinicalTrials.gov
Extending Layered Privacy Language to Support Privacy Icons for a Personal Privacy Policy User Interface
#404561
OpenAlex
Speaking the State: Linguistic Displacement and Political Support
#251692
OSF
DUET: Dual-Teacher On-Policy Distillation via Same-Weight Disagreement for Prohibition Compliance
#609081
arXiv (OAI Expanded)
Why Summaries Turn Neutral: Policy Attribution for Sentiment Drift in Reinforcement Learning from Human Feedback
#616189
arXiv (OAI Expanded)
Why Summaries Turn Neutral: Policy Attribution for Sentiment Drift in Reinforcement Learning from Human Feedback
#616689
arXiv (All)
Distilling Aggregate Mobility Statistics into a Language Model Policy for Post-Event Crowd Simulation
#656775
arXiv (OAI Expanded)
Distilling Aggregate Mobility Statistics into a Language Model Policy for Post-Event Crowd Simulation
#658911
arXiv (All)
OISD: On-Policy Internal Self-Distillation of Language Models
#789688
arXiv (OAI Expanded)
OISD: On-Policy Internal Self-Distillation of Language Models
#793692
arXiv (All)
Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation
#797056
arXiv (OAI Expanded)
Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation
#798469
arXiv (All)
AKTS: Sub-Microsecond Kernel Policy Switching for Language-Model Agents
#1000236
arXiv (All)
Summarize, Judge, Refine: Decoupled Content Understanding and Policy Learning for Multimodal Content Moderation
#1036778
arXiv (All)
1% of Tokens Can Be Enough: On Gradient Estimation in On-Policy Distillation
#1050609
arXiv (All)
Intercomprehension in Theory and Practice: Teaching and Learning Plurilingual Competence in a European Context
#191617
DigitalCommons@Kennesaw State University
TRACE: Temporal Rhetorical Analysis and Consistency Evaluation for Legislative Speech
#343917
DigitalCommons@CalPoly
Trust Is Not Enough: Influence Calibration for On-Policy Self-Distillation in Agentic RL
#609358
arXiv (OAI Expanded)
Policy Iteration with Human Feedback: Bringing Post-Training RL to In-context Learning
#623107
arXiv (OAI Expanded)
Policy Iteration with Human Feedback: Bringing Post-Training RL to In-context Learning
#625380
arXiv (All)
STAR-OPD: Structured Aspect-Cascade-Aware On-Policy Reward Distillation for ABSA Quadruple Extraction
#657239
arXiv (OAI Expanded)
STAR-OPD: Structured Aspect-Cascade-Aware On-Policy Reward Distillation for ABSA Quadruple Extraction
#659375
arXiv (All)
SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation
#661416
arXiv (OAI Expanded)
SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation
#662466
arXiv (All)
← Previous
Page 3 of 5
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.