 ##  [Markov Decision Process](/markov-decision-process) 

  ##  [Markov Decision Process](https://natural.quantumdictionary.io/markov-decision-process-0) 

  

 [![Natural & Formal Sciences Dictionary](/sites/default/files/styles/large/public/2026-01/Natural%20%26%20Formal%20Sciences.png.webp?itok=2kCDRVQv)](/topic-specific-dictionaries/natural-formal-sciences)



**Natural &amp; Formal Sciences Dictionary**

 







 

 

 

 



 

 

 

 

Definition

A mathematical model for sequential decision-making under uncertainty defined by a tuple (S, A, P, R, γ): state space S, action set A, transition kernel P(s'|s,a), reward function R(s,a), and discount factor γ; decisions (policies) map states to actions or distributions over actions.

 

 

 

 

 





 

 



 ##  [Markov Decision Process](https://engineering.quantumdictionary.io/markov-decision-process-1) 

  

 [![Engineering & Applied Technologies Dictionary](/sites/default/files/styles/large/public/2026-01/Engineering%20%26%20Applied%20Technologies.png.webp?itok=_IKO7_-n)](/topic-specific-dictionaries/engineering-applied-technologies)



**Engineering &amp; Applied Technologies Dictionary**

 







 

 

 

 



 

 

 

 

Definition

A formal model for sequential decision problems in which an agent observes a state from a state space, chooses an action from an action set, and the system transitions to a next state according to transition probabilities that depend only on the current state and action (the Markov property); a reward function assigns immediate payoffs and the objective is to maximize expected cumulative reward ov