OpenAI Gym
From MaRDI portal
Cited in
(only showing first 100 items - show all)- quadrature-ML
- Verisig
- GoTube
- nnenum
- DeepLoco
- TF-Agents
- Sherlock
- LaplacianSmoothing-GradientDescent
- DeepSynth
- MOGPTK
- Safety Gym
- CURL
- QT-Opt
- VIREL
- VIME
- SUNRISE
- LGSVL
- StereoSet
- MASS
- pymdp
- SeaPearl
- Isaac Gym
- ElegantRL
- CleanRL
- PyTorchRL
- IMPALA
- SafePILCO
- Deep active inference
- PIKAIA
- CompEcon
- ToolboxLS
- M-TRAN
- HyFlex
- FODD-Planner
- Importance sampling in reinforcement learning with an estimated behavior policy
- Accelerating reinforcement learning with a directional-Gaussian-smoothing evolution strategy
- Convex optimization with an interpolation-based projection and its application to deep learning
- Air learning: a deep reinforcement learning gym for autonomous aerial robot visual navigation
- Permutation flow shop scheduling with multiple lines and demand plans using reinforcement learning
- How does momentum benefit deep neural networks architecture design? A few case studies
- Neural network repair with reachability analysis
- Recruitment-imitation mechanism for evolutionary reinforcement learning
- SAMBA: safe model-based \& active reinforcement learning
- Reinforcement learning for robotic manipulation using simulated locomotion demonstrations
- Deep reinforcement learning for the control of conjugate heat transfer
- Quantum-enhanced reinforcement learning for control: a preliminary study
- Dynamic metasurface control using deep reinforcement learning
- Towards finding longer proofs
- End-to-end learning for off-road terrain navigation using the chrono open-source simulation platform
- A theoretical and empirical comparison of gradient approximations in derivative-free optimization
- Lipschitzness is all you need to tame off-policy generative adversarial imitation learning
- Laplacian smoothing gradient descent
- Data science applications to string theory
- Deep active inference as variational policy gradients
- Active deep Q-learning with demonstration
- Counterfactual state explanations for reinforcement learning agents via generative deep learning
- NMRDPP
- A review on deep reinforcement learning for fluid mechanics
- The Hanabi challenge: a new frontier for AI research
- DSSAT
- Branes with brains: exploring string vacua with deep reinforcement learning
- TD-regularized actor-critic methods
- Orbifolder
- SUMO
- COCO
- TEXPLORE
- RLPy
- Approxrl
- Reinforcement learning control of constrained dynamic systems with uniformly ultimate boundedness stability guarantee
- Preparation of three-atom GHZ states based on deep reinforcement learning
- Cimlib
- SPM
- You only lie twice: a multi-round cyber deception game of questionable veracity
- CUBIC
- POMDPs.jl
- Pandapower
- Pypsa
- GAZEBO classic
- ABC-LMPC: Safe Sample-Based Learning MPC for Stochastic Nonlinear Dynamical Systems with Adjustable Boundary Conditions
- Unity3D
- FAUST2
- APES
- MazeBase
- SUMMARIST
- ParlAI
- XNMT
- Nematus
- ELF
- SeqGAN
- Chainer
- MuJoCo
- Ray
- Libratus
- PEORL
- BindsNET
- Nengo
- ANNarchy
- Torchmeta
- ONNX
- ckn_kernel
This page was built for software: OpenAI Gym