Conceptio
›
reinforcement-learning
Topic
reinforcement-learning
Knowledge-graph topic
· documents ABOUT reinforcement-learning across the archive
165
Documents about reinforcement-learning
Documents about reinforcement-learning
Optimization of broadband metamaterial absorber using twin delayed deep deterministic policy gradient reinforcement learning technique.
#67460
NCBI PubMed Central
Reinforcement Learning for Intensity Control: An Application to Choice-Based Network Revenue Management
#139074
arXiv (OAI)
Multi-objective inventory optimization using reinforcement learning: a comparative study on profitability and carbon emissions.
#145872
NCBI PubMed Central
Off by a beat: the effects of temporal misalignment in reinforcement learning for sepsis treatment.
#165575
NCBI PubMed Central
OSP-Next: Efficient High-Quality Video Generation with Sparse Sequence Parallelism, HiF8 Quantization, and Reinforcement Learning
#236514
Papers With Code
Online Trading Models with Deep Reinforcement Learning in the Forex Market Considering Transaction Costs
#478198
CiNii
Safe and adaptive control of non-stationary stochastic systems via Lyapunov-constrained distributional reinforcement learning.
#479279
NCBI PubMed Central
Designing Reinforcement Learning for Diffusion Models: A Unified Path-Space View
#609600
arXiv (All)
TRACE-CASH: Trial-History-Conditioned Reinforcement Learning for Adaptive Configuration Exploration in Time-Series CASH
#619124
arXiv (OAI Expanded)
TRACE-CASH: Trial-History-Conditioned Reinforcement Learning for Adaptive Configuration Exploration in Time-Series CASH
#620803
arXiv (All)
The Ethical Decision Head: Operationalizing Normative Ethics in Autonomous Vehicles via Reinforcement Learning from Human Feedback
#622998
arXiv (OAI Expanded)
The Ethical Decision Head: Operationalizing Normative Ethics in Autonomous Vehicles via Reinforcement Learning from Human Feedback
#625271
arXiv (All)
E2HiL: Entropy-Guided Sample Selection for Efficient Real-World Human-in-the-Loop Reinforcement Learning
#668408
arXiv (OAI Expanded)
E2HiL: Entropy-Guided Sample Selection for Efficient Real-World Human-in-the-Loop Reinforcement Learning
#670473
arXiv (All)
Optimal Control with Natural Images: Efficient Reinforcement Learning using Overcomplete Sparse Codes
#674149
arXiv (OAI Expanded)
Optimal Control with Natural Images: Efficient Reinforcement Learning using Overcomplete Sparse Codes
#677346
arXiv (All)
The Effectiveness Of Using Star Boards As A Learning Motivation Strategy For Grade 1 Students (Case Study At Min 1 Langsa)
#218405
Jurnal al Muta'aliyah: Pendidikan Guru Madrasah Ibtidaiyah
Towards collective intelligence in agriculture: Deep reinforcement learning and digital twins for efficient management of collective irrigation water distribution systems
#274202
HAL (France)
Energy-Efficient Cooperative Data Offloading in Cellular Networks Using Reinforcement Learning
#667416
ARPHA OAI-PMH Endpoint
Reward Is Not a Simple Learning Signal: A Post-Inferential Representation of Evaluation
#228536
OSF
Transdiagnostic factors differentially shape choices and reaction times in social evaluative learning
#486066
PsyArXiv
Supplementary file 1_Application-driven pedagogical knowledge optimization of open-source LLMs via reinforcement learning and supervised fine-tuning.pdf
#325675
Figshare
Prompt Optimization Through Reinforcement Learning for Generative Language Model Code Synthesis in Multi-Robot Systems
#383735
Digital Commons @ Michigan Tech
Bridging Learning and Planning for Construction in a 3D environment
#306767
WIReDSpace
Offline Reinforcement Learning for Hemodynamic Management of Sepsis in the ICU: a MIMIC-IV Study with Dual Off-Policy Evaluation
#459679
arXiv CS
TL-RL-FusionNet: An Adaptive and Efficient Reinforcement Learning-Driven Transfer Learning Framework for Detecting Evolving Ransomware Threats
#123961
arXiv CS
Not All Tokens Matter: Towards Efficient LLM Reasoning via Token Significance in Reinforcement Learning
#149293
arXiv (OAI)
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty
#153090
arXiv (OAI)
Quantum framework for Reinforcement Learning: Integrating Markov decision process, quantum arithmetic, and trajectory search
#158655
arXiv (OAI)
Joint Optimization of Training and Inference in Federated Edge Learning via Constrained Multi-Objective Deep Reinforcement Learning
#224442
arXiv CS
Design of dynamic adaptive control framework for exoskeleton robot driven by neuromuscular signal based on deep reinforcement learning.
#378404
NCBI PubMed Central
A vibration control method for chatter mitigation in milling process based on sliding mode control and reinforcement learning.
#380773
NCBI PubMed Central
Energy-efficient wireless network control via spatio-temporal deep learning and multi-agent reinforcement learning.
#418600
NCBI PubMed Central
DreamWaQ: Learning Robust Quadrupedal Locomotion With Implicit Terrain Imagination via Deep Reinforcement Learning
#472146
KOASAS: KAIST Open Access Self-Archiving System
A reinforcement learning-driven adaptive hybrid PLC-RF communication architecture for IoT-based smart metering systems.
#483386
NCBI PubMed Central
SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation
#661416
arXiv (OAI Expanded)
SMOPD: Multi-Reward Reinforcement Learning via Specialize-and-Merge Online Policy Distillation
#662466
arXiv (All)
Highway Congestion Reduction through Reinforcement Learning Based Eulerian Headway Control
#668242
arXiv (OAI Expanded)
Highway Congestion Reduction through Reinforcement Learning Based Eulerian Headway Control
#670307
arXiv (All)
Residual Reward Models: Leveraging Prior Knowledge for Efficient Preference-based Reinforcement Learning in Robotics
#674192
arXiv (OAI Expanded)
Residual Reward Models: Leveraging Prior Knowledge for Efficient Preference-based Reinforcement Learning in Robotics
#677389
arXiv (All)
Now or later: A reinforcement learning model of behavioural delay
#31135
OSF
Optimizing AI Models for Biomedical Signal Processing Using Reinforcement Learning in Edge Computing
#71805
PhilArchive
OGER: A Robust Offline-Guided Exploration Reward for Hybrid Reinforcement Learning
#120545
arXiv CS
Reinforcement Learning for Antibiotic Stewardship: Optimizing Prescribing Policies Under Antimicrobial Resistance Dynamics
#130995
bioRxiv / medRxiv
Beam Scheduling for Cross-Layer ISAC: A Deep Reinforcement Learning Approach
#138872
arXiv CS
StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning
#178942
arXiv CS
Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation
#180710
arXiv CS
Scale: Deep Reinforcement Learning for Container Scheduling in Serverless Edge Computing
#194288
arXiv CS
Less Effort, Shorter Proofs: Reinforcement Learning for Security Protocol Analysis in Tamarin
#222541
arXiv CS
← Previous
Page 11 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.