[MINI] Multi-armed Bandit Problems

02/10/2015 12 min

Listen "[MINI] Multi-armed Bandit Problems"

Descargar episodio Ver en sitio original

Episode Synopsis

The multi-armed bandit problem is named with reference to slot machines (one armed bandits). Given the chance to play from a pool of slot machines, all with unknown payout frequencies, how can you maximize your reward? If you knew in advance which machine was best, you would play exclusively that machine. Any strategy less than this will, on average, earn less payout, and the difference can be called the "regret". You can try each slot machine to learn about it, which we refer to as exploration. When you've spent enough time to be convinced you've identified the best machine, you can then double down and exploit that knowledge. But how do you best balance exploration and exploitation to minimize the regret of your play? This mini-episode explores a few examples including restaurant selection and A/B testing to discuss the nature of this problem. In the end we touch briefly on Thompson sampling as a solution.

More episodes of the podcast Data Skeptic

Video Recommendations in Industry 26/12/2025

Eye Tracking in Recommender Systems 18/12/2025

Cracking the Cold Start Problem 08/12/2025

Designing Recommender Systems for Digital Humanities 23/11/2025

DataRec Library for Reproducible in Recommend Systems 13/11/2025

Shilling Attacks on Recommender Systems 05/11/2025

Music Playlist Recommendations 29/10/2025

Bypassing the Popularity Bias 15/10/2025

Sustainable Recommender Systems for Tourism 09/10/2025

Interpretable Real Estate Recommendations 22/09/2025

Ver todos los episodios

ZARZA We are Zarza, the prestigious firm behind major projects in information technology.

[MINI] Multi-armed Bandit Problems

Listen "[MINI] Multi-armed Bandit Problems"

Episode Synopsis

More episodes of the podcast Data Skeptic

Orthographic errors in Web pages

Telecommuting for employees of trust

Bandwidth: Broadband or Narrowband?

Personnel recruitment via Web

Deep web or Invisible Internet

Subdomains, a glance with the experts!

Free Internet, a prediction in Nostradamus style

Educational Technology: From traditional to digital

Localhost, there’s no place like 127.0.0.1

Googling with breathtaking tricks you ignore

Gray Hat Hacking, those with ambiguous ethics…

Internet Predators on the prowl

Dot COM: The Internet’s dominant TLD