11e Journée Statistique et Informatique pour la Science des Données à Paris-Saclay

Collection 11e Journée Statistique et Informatique pour la Science des Données à Paris-Saclay

Organisateur(s) Avetik Karagulyan, Erwan Le Pennec

Date(s) 03/04/2026 - 03/04/2026

URL associée https://indico.math.cnrs.fr/event/16080/

00:00:00 / 00:00:00

1 6

Linear Bandits on Ellipsoids: Minimax Optimal Algorithms

De Richard Combes

We consider linear stochastic bandits where the set of actions is an ellipsoid. We provide the first known minimax optimal algorithm for this problem. We first derive a novel information-theoretic lower bound on the regret of any algorithm, which must be at least $\Omega(\min(d \sigma \sqrt{T} + d |\theta|_{A}, |\theta|_{A} T))$ where $d$ is the dimension, $T$ the time horizon, $\sigma^2$ the noise variance, $A$ a matrix defining the set of actions, and $\theta$ the vector of unknown parameters. We then provide an algorithm whose regret matches this bound to a multiplicative universal constant. The algorithm is non-classical in the sense that it is not optimistic, and it is not a sampling algorithm. The main idea is to combine a novel sequential procedure to estimate $|\theta|$, followed by an explore-and-commit strategy informed by this estimate. The algorithm is highly computationally efficient, and a run requires only time $\mathcal{O}(dT + d^2 \log(T/d) + d^3)$ and memory $\mathcal{O}(d^2)$, in contrast with known optimistic algorithms, which are not implementable in polynomial time. We go beyond minimax optimality and show that our algorithm is locally asymptotically minimax optimal, a much stronger notion of optimality. We further provide numerical experiments to illustrate our theoretical findings.

Informations sur la vidéo

Date de captation 03/04/2026
Date de publication 13/04/2026
Institut IHES
Langue Anglais
Audience Chercheurs
Format MP4

Domaine(s)

Networking and Internet Architecture

Dernières questions liées sur MathOverflow

Pour poser une question, votre compte Carmin.tv doit être connecté à mathoverflow

Poser une question sur MathOverflow

Toutes les vidéos de la collection

53:43

publiée le 13 avril 2026

Linear Bandits on Ellipsoids: Minimax Optimal Algorithms

De Richard Combes

54:40

publiée le 13 avril 2026

Statistical Analysis of Multiple Networks

De Tabea Rebafka

47:06

publiée le 13 avril 2026

A Computable Measure of Suboptimality for Entropy-Regularised Variational Objectives

De Anna Korba

46:12

publiée le 13 avril 2026

Asymptotic Theory of Iterated Empirical Risk Minimization, with Applications to Active Learning

De Hugo Cui

48:16

publiée le 13 avril 2026

Convergence and Linear Speed-Up in Stochastic Federated Learning

De Paul Mangold

49:24

publiée le 13 avril 2026

Beyond Kemeny Medians: Consensus Ranking Distributions Definition, Properties and Statistical Learning

De Ekhine Irurozki

Copyright Carmin.tv 2026

Donner son avis