Conceptio
›
reinforcement-learning
Topic
reinforcement-learning
Knowledge-graph topic
· documents ABOUT reinforcement-learning across the archive
165
Documents about reinforcement-learning
Documents about reinforcement-learning
Human-Level Text-to-SQL via Reinforcement Learning on Verified Data, Without Pipeline Engineering
#661363
arXiv (OAI Expanded)
Human-Level Text-to-SQL via Reinforcement Learning on Verified Data, Without Pipeline Engineering
#662413
arXiv (All)
RecoverFly: A Failure-Aware Reinforcement Learning Post-Training Framework for Aerial Vision-Language Navigation
#682802
arXiv (OAI Expanded)
RecoverFly: A Failure-Aware Reinforcement Learning Post-Training Framework for Aerial Vision-Language Navigation
#684578
arXiv (All)
Real-time control of urban drainage system for flood and combined sewer overflow mitigation with a novel recurrent deep reinforcement learning framework.
#136119
PubMed
Predicting Liquidity-Aware Bond Yields using Causal GANs and Deep Reinforcement Learning with LLM Evaluation
#163095
arXiv (OAI)
Bis-peri-dinaphtho-rylenes: Facile Synthesis via Radical-Mediated Coupling Reactions and their Distinctive Electronic Structures
#294976
ScholarBank@NUS
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight
#131282
arXiv (OAI)
An Intelligent Fair and Decentralized Consensus Mechanism for Blockchain-Based Supply Chain
#239031
HAL (France)
Exploring disassembly challenges: jamming, compliance effects, and reinforcement learning in complex peg-holes processes
#328218
UBIRA ETheses
Longitudinal Deep Truck: Deep learning and deep reinforcement learning for modeling and control of longitudinal dynamics of heavy duty trucks.
#489662
DBLP
Efficient surrogate-based optimization framework integrating physics-informed neural networks, deep active learning and deep reinforcement learning: Multilayer thin films case study.
#489677
DBLP
<p dir="ltr">Optimization of the Laser Cladding Process for Q355 Iron-Based Alloys Using SHAP-Based Explainable Machine Learning </p>
#614851
Figshare
Audiovisual cues must be predictable and win-paired to drive risky choice
#67598
Europe PMC
Explainable AI for reinforcement learning based dynamic scheduling solutions in semiconductor manufacturing
#130586
Springer Nature OA
Behavioral basis of intrinsic motivation in primates
#291710
HAL (France)
Otto—Design and Control of an 8-DoF SEA-Driven Quadrupedal Robot
#415079
IRIS
Investigating the relationship between affective valence and reinforcement learning
#439839
OSF
Adaptive Laser Welding Control: A Reinforcement Learning Approach
#448950
Infoscience
Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving
#664177
arXiv (OAI Expanded)
Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving
#665677
arXiv (All)
Exploring cooperation mechanisms via deep reinforcement learning in network common-pool resource games
#668555
arXiv (OAI Expanded)
Exploring cooperation mechanisms via deep reinforcement learning in network common-pool resource games
#670620
arXiv (All)
A Bibliometric Analysis of Research on the Convergence of Artificial Intelligence and Blockchain in Smart Cities
#151874
HAL (France)
A novel deep self-learning method for flexible job-shop scheduling problems with multiplicity: Deep reinforcement learning assisted the fluid master-apprentice evolutionary algorithm.
#489723
DBLP
Bootstrapping of parameterized skills through hybrid optimization in task and policy spaces
#657522
Bielefeld Pub
AURA: An AI-Powered Multimodal Prototype for Adaptive Apraxia of Speech Therapy and Communication Support
#228064
ODU Digital Commons
Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models
#131299
arXiv (OAI)
SAR-ERL: an evolutionary reinforcement learning optimization method based on state–action co-representation embedding
#145128
Springer Nature OA
Poster Guided Reinforcement Learning technique on sugar intake and oral hygiene knowledge, attitude and practices among adolescents: A cluster randomized controlled trial.
#146044
NCBI PubMed Central
Experimental insights on CIRLEM: Enhancing energy efficiency and flexibility in buildings
#194412
HAL (France)
A novel hybrid task scheduling method in fog computing using reinforcement learning and metaheuristic algorithms
#252097
Springer Nature OA
Current practices and Limitations in E Health Informatics
#285494
OpenAlex
Sashimi-Bot: autonomous tri-manual advanced manipulation and cutting of deformable objects
#294660
HAL (France)
Differentiating Ischemic From Nonischemic T-Wave Inversion Using a Multimodal Vision-Language Model With Reinforcement Learning (ECG-R1): Development and Validation Study.
#297303
NCBI PubMed Central
Developmental differences in probabilistic reinforcement learning reflect strategy engagement rather than parameter changes
#416320
OSF
ProCAVE: A Self-Adaptive, Full-Lifecycle Edge Caching Framework for Video Streaming via Predictive Bandwidth Estimation and Preference-Aware Deep Reinforcement Learning
#429312
arXiv CS
Degradation-constrained multi-agent reinforcement learning with centralized training and decentralized execution for vehicle-to-grid optimization in renewable-dominated distribution networks.
#455456
NCBI PubMed Central
Towards a Digital Twin of a Solar Power Plant
#482537
HAL (France)
Mitigating Post-Training Effects on Generative Diversity in Language Models
#634924
eScholarship
Advantage-level Aggregation Reinforcement Learning for X-point Target Magnetic Configuration Control in an EXL-50U Experiment-Calibrated Simulation Environment
#660981
arXiv (OAI Expanded)
Reinforcement Learning to Harness Approximation Errors for Long-Time Quantum Simulation
#661443
arXiv (OAI Expanded)
Advantage-level Aggregation Reinforcement Learning for X-point Target Magnetic Configuration Control in an EXL-50U Experiment-Calibrated Simulation Environment
#662031
arXiv (All)
Reinforcement Learning to Harness Approximation Errors for Long-Time Quantum Simulation
#662493
arXiv (All)
Reward-induced endogenous pain inhibition scales with action-outcome certainty in humans
#676460
Europe PMC
Digital-Twin-Driven Predictive Maintenance and Fault-Tolerant Control for Electrified Agricultural Machinery: A Multiphysics and Deep Reinforcement Learning Framework for PMSM In-Wheel Drives
#707434
PhilArchive
A Unified Sequence Modeling Framework for Imperfect-Information Trick-Taking Card Games
#439648
OpenMETU
Reinforcement learning in maintenance decision-making and optimization: a literature review
#137664
Springer Nature OA
Autonomous System For Identifying and Capturing Floating Waste
#294561
HAL (France)
Federated deep reinforcement learning with transformer-based anomaly detection for cybersecurity in next-generation agricultural IoT networks
#313477
Springer Nature OA
← Previous
Page 19 of 20
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.