Atlas · GenAI 2026
Multi-armed Bandits
Sequential decision-making under uncertainty balancing exploration and exploitation
conceptPeak: 2017Sequential Decision-MakingAI consensus: 0/3
Prerequisites
No prerequisites.
Recommended reference
Slivkins, A. (2019) 'Introduction to Multi-Armed Bandits' — arXiv:1904.07272; comprehensive survey
Notes from AI deep research
Anthropic Opus
Recommender systems, ad serving, dynamic prompt/model selection. Elegant framework
Related skills
- → is subcategory of: Reinforcement Learning(3/3)