Conceptio
›
reinforcement-learning
Topic
reinforcement-learning
Knowledge-graph topic
· documents ABOUT reinforcement-learning across the archive
165
Documents about reinforcement-learning
Documents about reinforcement-learning
Reinforcement Learning from Human Feedback in LLMs: Whose Culture, Whose Values, Whose Perspectives?
#707091
PhilArchive
When LLM Signals Hurt: A Coverage-Density Analysis of LLM-Augmented Reinforcement Learning for Stock Trading
#181083
OSF
Advancements and Challenges in Machine Learning: A Comprehensive Review of Models, Libraries, Applications, and Algorithms
#461964
OpenAlex
RL-PLUS: Countering Capability Boundary Collapse of LLMs in Reinforcement Learning with Hybrid-policy Optimization
#144336
arXiv (OAI)
The social transmission of linguistic alignment: the role of observational reinforcement learning
#316564
OSF
Deep Learning and Reinforcement Learning for Autonomous Unmanned Aerial Systems: Roadmap for Theory to Deployment.
#489661
DBLP
From Thermal Preference Prediction to Adaptive Thermal Intervention: A Reinforcement Learning Approach Using Physiological and Environmental Sensing
#656859
arXiv (OAI Expanded)
CDRL: Certification-Driven Reinforcement Learning for Neutrino Flavor Model Discovery
#657109
arXiv (OAI Expanded)
From Thermal Preference Prediction to Adaptive Thermal Intervention: A Reinforcement Learning Approach Using Physiological and Environmental Sensing
#658995
arXiv (All)
CDRL: Certification-Driven Reinforcement Learning for Neutrino Flavor Model Discovery
#659245
arXiv (All)
Agentic-Kube: A Graph-Enhanced Multi-Agent Reinforcement Learning Framework for Multi-Objective Kubernetes Scheduling
#682607
arXiv (OAI Expanded)
Agentic-Kube: A Graph-Enhanced Multi-Agent Reinforcement Learning Framework for Multi-Objective Kubernetes Scheduling
#684382
arXiv (All)
A novel multi−agent deep reinforcement learning framework for fast frequency response in inverter−based hybrid power plants
#299042
Research Online
Ontology Synthesis Using Semi-Automatic Semantic AI Using Deep Learning and Reinforcement Learning.
#489675
DBLP
Achieving Scale-Invariant Reinforcement Learning Performance with Reward Range Normalization
#660471
OSF
Sharing the Control Authority Between Deep Reinforcement Learning and Model Predictive Control: Application to Multi-Class Transportation Networks
#661002
arXiv (OAI Expanded)
Sharing the Control Authority Between Deep Reinforcement Learning and Model Predictive Control: Application to Multi-Class Transportation Networks
#662052
arXiv (All)
PromptEcho: Annotation-Free Reward from Vision-Language Models for Text-to-Image Reinforcement Learning
#13156
arXiv CS
The Coherence Engine, part deux; Collapsing Bayesian Inference, Free Energy, Gradient Descent, Markov Models, and Reinforcement Learning
#89364
PhilArchive
Neural decision dynamics underlying reinforcement learning and working memory
#125753
Europe PMC
Fast State Stabilization using Deep Reinforcement Learning for Measurement-based Quantum Feedback Control
#128306
arXiv (OAI)
Semi-Markov Reinforcement Learning for City-Scale EV Ride-Hailing with Feasibility-Guaranteed Actions
#141520
arXiv CS
ProRank: Prompt Warmup via Reinforcement Learning for Small Language Models Reranking
#147101
arXiv (OAI)
Mitigating False Positives in Static Memory Safety Analysis of Rust Programs via Reinforcement Learning
#155356
arXiv CS
Queue-Aware and Resilient Routing in LEO Satellite Networks Using Multi-Agent Reinforcement Learning
#168256
arXiv CS
AegisTS: A Hierarchical Agent System with Reinforcement Learning for Multivariate Time Series Data Cleaning
#168423
arXiv CS
Reward Is Enough: LLMs Are In-Context Reinforcement Learners
#173309
arXiv (OAI)
BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models
#175344
arXiv CS
Data privacy model using blockchain reinforcement federated learning approach for scalable internet of medical things
#177252
Semantic Scholar
MARLIN: Multi-Agent Game-Theoretic Reinforcement Learning for Sustainable LLM Inference in Cloud Datacenters
#180623
arXiv CS
R3DM: Enabling Role Discovery and Diversity Through Dynamics Models in Multi-agent Reinforcement Learning
#183249
arXiv (OAI)
Angel or Demon: Investigating the Plasticity Interventions' Impact on Backdoor Threats in Deep Reinforcement Learning
#187246
arXiv CS
Multi-objective application placement in fog computing using graph neural network-based reinforcement learning
#187283
arXiv CS
Heterogeneous Tasks Offloading in Vehicular Edge Computing: A Federated Meta Deep Reinforcement Learning Approach
#200447
arXiv CS
SAFE AND AGILE UAV NAVIGATION IN CONFINED AND TIGHT SPACES VIA DEEP REINFORCEMENT LEARNING
#211397
KFUPM ePrints
Spreadsheet-RL: Advancing Large Language Model Agents on Realistic Spreadsheet Tasks via Reinforcement Learning
#216875
arXiv CS
A Reinforcement Learning-driven Transformer GAN for Molecular Generation
#248665
Springer Nature OA
Shape Formation for the Cooperative Transportation of Arbitrary Objects Using Multi-Agent Reinforcement Learning
#267699
arXiv CS
Dynamic-layer transformer-based reinforcement learning for observation-constrained multi-agent roundup scenarios.
#285151
NCBI PubMed Central
Fractional-Order Complex Systems: Advanced Control, Intelligent Estimation and Reinforcement Learning Image-Processing Algorithms
#330208
HAL (France)
Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning
#343552
arXiv CS
An Explainable Multimodal AI Framework with Reinforcement Learning for Post-Surgical Clinical Decision Support
#353182
bioRxiv / medRxiv
Modeling Advanced Persistent Threats Using Stackelberg Game Theory and Reinforcement Learning Under Partial Observability
#403990
KFUPM ePrints
WarmTuner: Program-Specific Warm Starts for Compiler Autotuning via Offline-to-Online Reinforcement Learning
#411129
arXiv CS
Translation with Thought: Difficulty-Adaptive Reasoning via Reinforcement Learning for Multi-Domain Machine Translation
#422302
arXiv CS
LLM-Driven Automated Reward Design for Reinforcement Learning-Based Routing in LEO Satellite Networks
#423896
arXiv CS
Chess on Ice: Curling Tactical Decision-Making via Backward Induction and Deep Reinforcement Learning
#424020
arXiv CS
An Emerging Retail Portfolio Management Application: Personalized, Tax-Aware Reinforcement Learning with Natural Language Goals
#432477
arXiv CS
EvoRIC: Reinforcement Learning Fine-Tuned LLM-empowered RAN Intelligent Control Toward Autonomous O-RAN
#440325
arXiv CS
Message-Dropout: An Efficient Training Method for Multi-Agent Deep Reinforcement Learning
#445787
KOASAS: KAIST Open Access Self-Archiving System
← Previous
Page 15 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.