pipette
ENEnglish

Efficient Linear Bandits via Cluster-Aware Sketching

Hantao Yang, Hong Xie, Defu Lian

PreprintAfirmaciones fuertes, leer con cuidado

En palabras de los autores

We study the problem of computational efficiency for linear bandits in high-dimensional settings with a finite arm set. In linear bandits, the increase in the dimension of the feature vectors leads to growing computational costs of at each round of update. Traditional sketching-based methods such as SOFUL reduce computation via fixed-size matrix sketching, yet run the risk of incurring vacuous linear regret when the spectral tail of the data is heavy and the sketch size is inadequately selected. To guarantee regret convergence and effectively reduce computational costs, we introduce a clustering mechanism and propose the Cluster Sketch Linear Bandit (CS-LB) algorithm. Our method preserves the full covariance information in each cluster to guarantee robust sublinear regret without spectral-tail vulnerabilities, performs cluster switching by assigning a sentinel for each cluster, and reduces per-round update computation to via a tunable sketch size . Experiments on synthetic datasets demonstrate that our method consistently maintains a favorable trade-off between efficiency and regret.

Resultado principalLimitación que admiten los autores

Apareció: jueves, 24 de septiembre. arXiv. Preprint, todavía sin revisión por pares.