Kernel-based methods for bandit convex optimization

We consider the adversarial convex bandit problem and we build the first poly(T)-time algorithm with poly(n) √T-regret for this problem. To do so we introduce three new ideas in the derivative-free optimization

Paper

Similar papers

© 2026 NYSGPT2525 LLC