Publications


Research in Learning Theory, Bandits, and Reinforcement Learning
Selected Research Output

Publications

Research on bandits, online learning, reinforcement learning, high-dimensional statistics, and PAC-Bayes theory.

Official Publications

AISTATS2026

Sparse Linear Bandits with Fixed Sparsity Support: Adversarial and Stochastic Regimes

Kyoungseok Jang, Nam Tran, Nicolò Cesa-Bianchi

The Conference on Artificial Intelligence and Statistics

AISTATS2026

GL-LowPopArt: A Nearly Instance-Wise Minimax Estimator for (Adaptive) Generalized Linear Low-Rank Trace Regression

Junghyun Lee, Kyoungseok Jang, Kwang-Sung Jun, Milan Vojnovic, Se-Young Yun

The Conference on Artificial Intelligence and Statistics

View paper
AISTATS2026

Rate-optimal Design for Anytime Best Arm Identification

Junpei Komiyama, Kyoungseok Jang, Honda Junya

The Conference on Artificial Intelligence and Statistics

View paper
AAAI2026

Online Linear Regression with Paid Stochastic Features

Nadav Merlis, Kyoungseok Jang, Nicolò Cesa-Bianchi

The Association for the Advancement of Artificial Intelligence

View paper
NeurIPS2024

Fixed Confidence Best Arm Identification in the Bayesian Setting

Kyoungseok Jang, Junpei Komiyama, Kazutoshi Yamazaki

Conference on Neural Information Processing Systems

View paper
NeurIPS2024

Sparsity-Agnostic Linear Bandits with Adaptive Adversaries

Tianyuan Jin, Kyoungseok Jang, Nicolò Cesa-Bianchi

Conference on Neural Information Processing Systems

View paper
ICML2024

Efficient Low-Rank Matrix Estimation, Experimental Design, and Arm-Set-Dependent Low-Rank Bandits

Kyoungseok Jang, Kwang-Sung Jun, Chicheng Zhang

International Conference on Machine Learning

View paper
COLT2024

Better-than-KL PAC-Bayes Bounds

Ilja Kuzborskij, Kwang-Sung Jun, Yulian Wu, Kyoungseok Jang, Francesco Orabona

Conference on Learning Theory

View paper
COLT2023

Tighter PAC-Bayes Bound through Coin-Betting

Kyoungseok Jang, Kwang-Sung Jun, Ilja Kuzborskij, Francesco Orabona

Conference on Learning Theory

View paper
NeurIPS2022

PopArt: Efficient Sparse Regression and Experimental Design for Optimal Sparse Linear Bandits

Kyoungseok Jang, Chicheng Zhang, Kwang-Sung Jun

Conference on Neural Information Processing Systems

View paper
ICML2021

Improved Regret Bounds of Bilinear Bandits using Action Space Analysis

Kyoungseok Jang, Kwang-Sung Jun, Se Young Yun, Wanmo Kang

International Conference on Machine Learning

View paper

Preprints

Preprint2025

Stability and Generalization for Bellman Residuals

Enoch H. Kang, Kyoungseok Jang

View paper
Preprint

Globally Convergent Offline Reinforcement Learning with Bellman Residual Minimization

Byungjun Park, Minhyeok Park, Enoch H. Kang, Kyoungseok Jang

Preprint2026

Efficient Algorithms for Contextual Apple Tasting with Log-Loss

Byeongwoo An, Kapilan Balagopalan, Sehwa Jeong, Jiun Jeong, Kyoungseok Jang, Hyowon Wi, Noseong Park, Gi-Soo Kim, Kwang-Sung Jun

Workshop Papers

  • 2026
    Stability and Generalization for Bellman Residuals

    Enoch H. Kang, Kyoungseok Jang

    ICML Workshop on Decision-Making from Offline Datasets to Online Adaptation · Poster
  • 2026
    Efficient Algorithms for Contextual Apple Tasting with Log-Loss

    Byeongwoo An, Kapilan Balagopalan, Hyowon Wi, Sehwa Jeong, Jiun Jeong, Kyoungseok Jang, Noseong Park, Gi-Soo Kim, Kwang-Sung Jun

    ICML Workshop on Decision-Making from Offline Datasets to Online Adaptation
  • 2026
    Globally Convergent Offline Reinforcement Learning with Smoothed Bellman Residual Minimization

    Byungjun Park, Minhyeok Park, Enoch Hyunwook Kang, Kyoungseok Jang

    ICML Workshop on Decision-Making from Offline Datasets to Online Adaptation · Poster
  • 2026
    Statistical Complexity of Soft Bellman Residual Minimization

    Enoch Hyunwook Kang, Kyoungseok Jang

    ICML Workshop on Decision-Making from Offline Datasets to Online Adaptation · Poster
  • 2024
    Fixed Confidence Best-Arm Identification in the Bayesian Setting

    Kyoungseok Jang, Junpei Komiyama, Kazutoshi Yamazaki

    Second RL Theory Workshop, co-located with COLT 2024 · Poster
  • 2023
    Improved Time-Uniform PAC-Bayes Bounds using Coin Betting

    Kyoungseok Jang, Kwang-Sung Jun, Ilja Kuzborskij, Francesco Orabona

    ICML Workshop: PAC-Bayes Meets Interactive Learning · Contributed Talk and Poster