Back to Home
AI on Medium··Industry Media

Q-Learning (Reinforcement learning): Bellman Equation, Markov Decision Processes, Q-Values, and…

中文摘要

本文探讨Q学习及其核心概念(贝尔曼方程、MDP和Q值),旨在解决非预测性的现实问题。

English Summary

This article explains Q-Learning and its fundamentals, including the Bellman Equation, MDPs, and Q-Values, to solve non-predictive real-world problems.

Original Excerpt

Machine learning models are excellent at making predictions. However, many real-world problems are not prediction problems. Continue reading on Medium »