Stanford
University
  • Stanford Home
  • Maps & Directions
  • Search Stanford
  • Emergency Info
  • Terms of Use
  • Privacy
  • Copyright
  • Trademarks
  • Non-Discrimination
  • Accessibility
© Stanford University.  Stanford, California 94305.
HAI Weekly Seminar with Mohsen Bayati | Stanford HAI
Skip to content
  • About

    • About
    • People
    • Get Involved with HAI
    • Support HAI
    • Subscribe to Email
  • Research

    • Research
    • Fellowship Programs
    • Grants
    • Student Affinity Groups
    • Centers & Labs
    • Research Publications
    • Research Partners
  • Education

    • Education
    • Executive and Professional Education
    • Government and Policymakers
    • K-12
    • Stanford Students
  • Policy

    • Policy
    • Policy Publications
    • Policymaker Education
    • Student Opportunities
  • AI Index

    • AI Index
    • AI Index Report
    • Global Vibrancy Tool
    • People
  • News
  • Events
  • Industry
  • Centers & Labs
Navigate
  • About
  • Events
  • AI Glossary
  • Careers
  • Search
Participate
  • Get Involved
  • Support HAI
  • Contact Us

Stay Up To Date

Get the latest news, advances in research, policy work, and education program updates from HAI in your inbox weekly.

Sign Up For Latest News

Your browser does not support the video tag.
eventSeminar

HAI Weekly Seminar with Mohsen Bayati

Status
Past
Date
Wednesday, February 16, 2022 10:00 AM - 11:00 AM PST/PDT
Location
Virtual
Share
Link copied to clipboard!
Event Contact
Kaci Peel
kpeel@stanford.edu

Related Events

Alexandr Lenk & Arvind Karunakaran | Industry Conversation with Instacart
SeminarSep 23, 202612:00 PM - 1:15 PM
September
23
2026

This seminar pairs a case study of AI diffusion with an organizational approach to technological changes in the workplace.

Seminar

Alexandr Lenk & Arvind Karunakaran | Industry Conversation with Instacart

Sep 23, 202612:00 PM - 1:15 PM

This seminar pairs a case study of AI diffusion with an organizational approach to technological changes in the workplace.

Marlowe | AI + Data for Science with Stephen Baccus
SeminarSep 23, 2026
September
23
2026

Sessions run Wednesdays from 4:30–5:30 PM in CoDa E160. Each session features a different Stanford speaker; talk titles are announced by the organizers.

Event

Marlowe | AI + Data for Science with Stephen Baccus

Sep 23, 2026

Sessions run Wednesdays from 4:30–5:30 PM in CoDa E160. Each session features a different Stanford speaker; talk titles are announced by the organizers.

World Development Report 2026: The Promise of Artificial Intelligence
Sep 24, 20269:00 AM - 4:15 PM
September
24
2026
Event

World Development Report 2026: The Promise of Artificial Intelligence

Sep 24, 20269:00 AM - 4:15 PM

The Unreasonable Effectiveness of Greedy Algorithms in Multi-Armed Bandits

The stochastic multi-armed bandit (MAB) is a benchmark model for decision-making under uncertainty. In the classical MAB setting, a decision maker sequentially chooses between a set of alternatives ("arms"), and earns a reward upon each choice. The decision maker's goal is to ensure these rewards are as high as possible over their decision horizon. MABs are used in a wide range of applications, from Internet advertising to healthcare.

It is well known that high performing MAB algorithms must balance "exploration", i.e., learning about relatively unknown arms, against "exploitation", i.e., leveraging arms that have already been seen to perform reasonably well. Unfortunately, due to practical constraints, fairness requirements, and ethical considerations, actively exploring may not be possible in some domains. For example, in health care, "exploration" may involve using an untested treatment on a prospective patient, but ethical considerations may preclude such use without appropriate safeguards.

Surprisingly, a body of recent research has suggested that in many practical regimes of interest, algorithms for MAB problems that focus solely on exploitation (i.e., choosing the empirical best arm) -- known as "greedy" algorithms -- in fact can perform quite well, due to exploration that happens for "free" during the run of the algorithm.  In this talk we describe this phenomenon; highlight its specific emergence in particular in MAB problems with large numbers of arms, as well as in a range of other settings; and suggest directions for future investigation.

Joint work with Nima Hamidi, Ramesh Johari, and Khashayar Khosravi.

Read "When 'Greedy' is Good" here

Mohsen Bayati

Associate Professor of Operations, Information and Technology at The Graduate School of Business and, by courtesy, of Electrical Engineering

No tweets available.