HAL (France)open access
Apprentissage dans les jeux stochastiques
Jeux stochastiques, Q-Apprentissage optimiste, Apprentissage par renforcement, 0, Nash equilibrium, Counterfactual regret minimization, Équilibre de Nash, Optimistic Q-Learning
This document is indexed with metadata only — full text is not available in the archive for this record.
Open the official source ↗
Record · ID 262041
Retrieved via
Conceptio — every document is proof-bundled with source, license, and retrieval metadata.