Kwang-Sung Jun
Publications 5
Provably Efficient Regularized Online RLHF with Generalized Bilinear Preferences New
NeurIPS 2026
ICML 2026 Pluralistic Alignment Workshop
CKAIA 2026
· Distinguished Paper Award
GL-LowPopArt: A Nearly Instance-Wise Minimax-Optimal Estimator for Generalized Low-Rank Trace Regression
AISTATS 2026
Pointwise or Pairwise: When Do Pairwise Losses Help Reward Learning, Provably?
arXiv:2609.37209
A Unified Confidence Sequence for Generalized Linear Models, with Applications to Bandits
NeurIPS 2024
ICML 2024 ARLET Workshop
· Oral
CKAIA 2024
· Best Paper Award
Improved Regret Bounds of (Multinomial) Logistic Bandits via Regret-to-Confidence-Set Conversion
AISTATS 2024