Co-occurring keywords
Papers
Thompson Sampling with a Mixture Prior
AISTATS 2022
Near-optimal Policy Optimization Algorithms for Learning Adversarial Linear Mixture MDPs
AISTATS 2022
Gap-Dependent Bounds for Two-Player Markov Games
AISTATS 2022
Can Q-learning be Improved with Advice?
COLT 2022
Versatile Dueling Bandits: Best-of-both World Analyses for Learning from Relative Preferences
ICML 2022
Efficient Kernelized UCB for Contextual Bandits
AISTATS 2022
Learning Two-Player Markov Games: Neural Function Approximation and Correlated Equilibrium
NIPS 2022