Publications


Research in Learning Theory, Bandits, and Reinforcement Learning
Selected Research Output

Publications

Research on bandits, online learning, reinforcement learning, high-dimensional statistics, and PAC-Bayes theory.

Official Publications

NeurIPS2026

Statistical Complexity of Soft Bellman Residual Minimization

Enoch Hyunwook Kang*, Kyoungseok Jang*
* Equal contribution

Conference on Neural Information Processing Systems · Main Track · Accepted as Poster

View paper →
AISTATS2026

Sparse Linear Bandits with Fixed Sparsity Support: Adversarial and Stochastic Regimes

Kyoungseok Jang, Nam Tran, Nicolò Cesa-Bianchi

The Conference on Artificial Intelligence and Statistics

AISTATS2026

GL-LowPopArt: A Nearly Instance-Wise Minimax Estimator for (Adaptive) Generalized Linear Low-Rank Trace Regression

Junghyun Lee, Kyoungseok Jang, Kwang-Sung Jun, Milan Vojnovic, Se-Young Yun

The Conference on Artificial Intelligence and Statistics

View paper →
AISTATS2026

Rate-optimal Design for Anytime Best Arm Identification

Junpei Komiyama, Kyoungseok Jang, Honda Junya

The Conference on Artificial Intelligence and Statistics

View paper →
AAAI2026

Online Linear Regression with Paid Stochastic Features

Nadav Merlis, Kyoungseok Jang, Nicolò Cesa-Bianchi

The Association for the Advancement of Artificial Intelligence

View paper →
NeurIPS2024

Fixed Confidence Best Arm Identification in the Bayesian Setting

Kyoungseok Jang, Junpei Komiyama, Kazutoshi Yamazaki

Conference on Neural Information Processing Systems

View paper →
NeurIPS2024

Sparsity-Agnostic Linear Bandits with Adaptive Adversaries

Tianyuan Jin, Kyoungseok Jang, Nicolò Cesa-Bianchi

Conference on Neural Information Processing Systems

View paper →
ICML2024

Efficient Low-Rank Matrix Estimation, Experimental Design, and Arm-Set-Dependent Low-Rank Bandits

Kyoungseok Jang, Kwang-Sung Jun, Chicheng Zhang

International Conference on Machine Learning

View paper →
COLT2024

Better-than-KL PAC-Bayes Bounds

Ilja Kuzborskij, Kwang-Sung Jun, Yulian Wu, Kyoungseok Jang, Francesco Orabona

Conference on Learning Theory

View paper →
COLT2023

Tighter PAC-Bayes Bound through Coin-Betting

Kyoungseok Jang, Kwang-Sung Jun, Ilja Kuzborskij, Francesco Orabona

Conference on Learning Theory

View paper →
NeurIPS2022

PopArt: Efficient Sparse Regression and Experimental Design for Optimal Sparse Linear Bandits

Kyoungseok Jang, Chicheng Zhang, Kwang-Sung Jun

Conference on Neural Information Processing Systems

View paper →
ICML2021

Improved Regret Bounds of Bilinear Bandits using Action Space Analysis

Kyoungseok Jang, Kwang-Sung Jun, Se Young Yun, Wanmo Kang

International Conference on Machine Learning

View paper →

Preprints

Preprint

Globally Convergent Offline Reinforcement Learning with Bellman Residual Minimization

Byungjun Park, Minhyeok Park, Enoch H. Kang, Kyoungseok Jang

Preprint2026

Efficient Algorithms for Contextual Apple Tasting with Log-Loss

Byeongwoo An, Kapilan Balagopalan, Sehwa Jeong, Jiun Jeong, Kyoungseok Jang, Hyowon Wi, Noseong Park, Gi-Soo Kim, Kwang-Sung Jun

Workshop Papers

  • 2026
    Stability and Generalization for Bellman Residuals

    Enoch H. Kang, Kyoungseok Jang

    ICML Workshop on Decision-Making from Offline Datasets to Online Adaptation · Poster
  • 2026
    Efficient Algorithms for Contextual Apple Tasting with Log-Loss

    Byeongwoo An, Kapilan Balagopalan, Hyowon Wi, Sehwa Jeong, Jiun Jeong, Kyoungseok Jang, Noseong Park, Gi-Soo Kim, Kwang-Sung Jun

    ICML Workshop on Decision-Making from Offline Datasets to Online Adaptation
  • 2026
    Globally Convergent Offline Reinforcement Learning with Smoothed Bellman Residual Minimization

    Byungjun Park, Minhyeok Park, Enoch Hyunwook Kang, Kyoungseok Jang

    ICML Workshop on Decision-Making from Offline Datasets to Online Adaptation · Poster
  • 2026
    Statistical Complexity of Soft Bellman Residual Minimization

    Enoch Hyunwook Kang*, Kyoungseok Jang* (* Equal contribution)

    ICML Workshop on Decision-Making from Offline Datasets to Online Adaptation · Poster
  • 2024
    Fixed Confidence Best-Arm Identification in the Bayesian Setting

    Kyoungseok Jang, Junpei Komiyama, Kazutoshi Yamazaki

    Second RL Theory Workshop, co-located with COLT 2024 · Poster
  • 2023
    Improved Time-Uniform PAC-Bayes Bounds using Coin Betting

    Kyoungseok Jang, Kwang-Sung Jun, Ilja Kuzborskij, Francesco Orabona

    ICML Workshop: PAC-Bayes Meets Interactive Learning · Contributed Talk and Poster