State-Action Value Function : the total amount of reward an agent can expect to accumulate in the future from a given state when taking a given action.