markov decision process (mdp)
4 episodes mention this concept
lexfridmanMar 12, 2019Leslie Kaelbling on Reinforcement Learning, Planning, Robotics, and the Philosophy of AI
TechnologyReinforcement LearningPlanning (AI)Robot NavigationArtificial Intelligence (AI)
lexfridmanJan 25, 2018Deep Reinforcement Learning: From Raw Data to Reasoning and Real-World Action
TechnologyDeep Reinforcement Learning (DRL)Artificial Intelligence (AI) stackEnd-to-end learningSupervised Learning
lexfridmanSep 27, 2016Deep Reinforcement Learning: Core Methods, Applications, and Policy Gradients
TechnologyDeep Reinforcement Learning (DRL)Reinforcement Learning (RL)Neural NetworksFunction Approximators
lexfridmanDeep Reinforcement Learning Fundamentals: From Perceptrons to Q-Learning for Motion Planning
TechnologyDeep Reinforcement Learning (DRL)Supervised LearningUnsupervised LearningSemi-supervised Learning
Knowledge Graph
Related concepts — line thickness indicates connection strength. Click any node to explore.
Related concepts
Concepts that appear alongside markov decision process (mdp) across episodes.
- supervised learning
- neural networks
- deep reinforcement learning (drl)
- q-learning
- policy
- reinforcement learning (rl)
- value function
- unsupervised learning
- monte carlo tree search (mcts)
- bellman equation
- reward clipping
- sparse reward data
- exploration vs. exploitation
- recurrent neural network
- temporal dynamics
- agent-environment interaction
- robot navigation
- perceptron
- artificial intelligence (ai) stack
- symbolic systems