Skip to content

other

Q-learning

Value-based reinforcement learning algorithm that learns action values from experience; in its tabular form it stores one value per state-action pair.

Known aliases

  • tabular Q-learning

Relationships

No evidence-backed relationships are recorded.

Current clusters