Conceptio
›
reinforcement-learning
Topic
reinforcement-learning
Knowledge-graph topic
· documents ABOUT reinforcement-learning across the archive
165
Documents about reinforcement-learning
Documents about reinforcement-learning
Parameter Sharing with Network Pruning for Scalable Multi-Agent Deep Reinforcement Learning
#445789
KOASAS: KAIST Open Access Self-Archiving System
MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer
#445792
KOASAS: KAIST Open Access Self-Archiving System
Sample-efficient and safe deep reinforcement learning via reset deep ensemble agents
#445794
KOASAS: KAIST Open Access Self-Archiving System
Adaptive and Explainable Deployment of Navigation Skills via Hierarchical Deep Reinforcement Learning
#472158
KOASAS: KAIST Open Access Self-Archiving System
Deep Reinforcement Learning for 6G AI-RAN: A Comprehensive Survey
#609296
arXiv (OAI Expanded)
DiSA-IQL: Offline Reinforcement Learning for Robust Soft Robot Control under Distribution Shifts
#609823
arXiv (All)
Designing Quantum Error Correcting Codes to fit decoders via Reinforcement Learning
#616394
arXiv (OAI Expanded)
Designing Quantum Error Correcting Codes to fit decoders via Reinforcement Learning
#616894
arXiv (All)
How working memory and reinforcement learning interact when avoiding punishment and pursuing reward concurrently.
#631069
NCBI PubMed Central
Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs
#656923
arXiv (OAI Expanded)
Toward Cost-efficient Adaptive Clinical Trials in Knee Osteoarthritis with Reinforcement Learning
#657813
Cornell eCommons
Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs
#659059
arXiv (All)
WM-R1: Training GUI Agents to Reason and leverage World Models with Reinforcement Learning
#779447
arXiv (OAI Expanded)
Learning and control in trustworthy and responsible artificial intelligence for cyber-physical systems
#206978
RUcore, Rutgers University Community Repository
Reward-specific learning parameters change across normative adolescent development and are blunted in youth with high risk for depression
#183588
Europe PMC
Decentralized Q-Learning Supervisory Control for Coordinated Multi-Loop Tuning in Pump Stations
#173982
Digital Commons @ Michigan Tech
Adapting Autonomous Agents for Automotive Driving Games
#388273
IRIS
Closer to Human: Hybrid Training of VR Agents Beyond Robotic Motion
#411882
HAL (France)
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks
#633605
arXiv (OAI Expanded)
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks
#635954
arXiv (All)
Reinforcement learning for medical image analysis: a systematic review of algorithms, engineering challenges, and clinical deployment.
#9744
PubMed
Evolutionary reinforcement learning framework for energy-efficient fault resilience and topological stability in WSNs.
#1036
NCBI PubMed Central
An Adaptive Blockchain Framework for Federated IoMT with Reinforcement Learning-Based Consensus and Resource Forecasting.
#1218
NCBI PubMed Central
Migrating SofaGym plugin to Gymnasium
#68418
DataCite
WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforcement Learning
#124118
arXiv CS
Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning
#128403
arXiv (OAI)
Bridging Continuous-time LQR and Reinforcement Learning via Gradient Flow of the Bellman Error
#147105
arXiv (OAI)
QoS Assurance Mechanism for 5G Network Slicing Based on the Deep Reinforcement Learning PPO Algorithm
#155185
arXiv CS
Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning
#158544
arXiv CS
Analysis of Randomization Effects on Sim2Real Transfer in Reinforcement Learning for Robotic Manipulation Tasks
#176454
arXiv Biology
Analysis of Randomization Effects on Sim2Real Transfer in Reinforcement Learning for Robotic Manipulation Tasks
#176938
arXiv (OAI)
Analysis of Randomization Effects on Sim2Real Transfer in Reinforcement Learning for Robotic Manipulation Tasks
#177192
arXiv (OAI Expanded)
Analysis of Randomization Effects on Sim2Real Transfer in Reinforcement Learning for Robotic Manipulation Tasks
#177672
arXiv
The Reward Positivity Tracks Positive Reward Prediction Errors From Feedback to Cues During Reinforcement Learning.
#182289
NCBI PubMed Central
Fair-Aurora: Comparing Fairness Strategies for Reinforcement Learning-Based Congestion Control in Multi-Flow Environments
#204740
arXiv CS
Reinforcement Learning and Savings Behavior
#223977
BYU ScholarsArchive
BORA: Bridging Offline Reinforcement Learning and Online Residual Adaptation for Real-World Dexterous VLA Models
#238660
arXiv CS
Optimized glycemic control of type 2 diabetes with reinforcement learning: a proof-of-concept trial
#240367
Ora Oxford University Research
When in Doubt, Plan It Out: Committed Small Language Model Deliberation for Reactive Reinforcement Learning
#280193
arXiv CS
LOLLA: Deep Reinforcement Learning for Closed-Loop Link Adaptation Towards a GPU-Accelerated AI-RAN
#299860
arXiv CS
Neural signatures of model-based and model-free reinforcement learning across prefrontal cortex and striatum.
#305002
NCBI PubMed Central
Effect of Al₂O₃ nanoparticle reinforcement and annealing on PMMA composite performance for prosthetic feet.
#326921
NCBI PubMed Central
Robust multi-agent reinforcement learning framework for intelligent PV-integrated smart energy systems under uncertainty.
#354793
NCBI PubMed Central
A shared functional architecture for error-based and reinforcement-based motor learning in the human brain
#379822
bioRxiv / medRxiv
Equity-Preserving Public Health Resource Allocation Using Multi-Objective Safe Reinforcement Learning: Evidence from Thailand.
#409333
NCBI PubMed Central
SVR: Self-Verifying Refinement via Joint Verdict-Confidence Reinforcement Learning for Adaptive Test-Time Compute
#414147
arXiv CS
Reinforcement Learning for Chronic Care Pathway Optimization: A Unified Framework across Three Clinical Goal Types
#424132
bioRxiv / medRxiv
Efficient Replay Memory Architectures in Multi-Agent Reinforcement Learning for Traffic Congestion Control
#430510
KOASAS: KAIST Open Access Self-Archiving System
TLCR: Token-Level Continuous Reward for Fine-grained Reinforcement Learning from Human Feedback
#436429
KOASAS: KAIST Open Access Self-Archiving System
A variational approach to mutual information-based coordination for multi-agent reinforcement learning
#445790
KOASAS: KAIST Open Access Self-Archiving System
← Previous
Page 16 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.