Sanghwa Kim
Publications 4
Instance-Optimal Estimation with Multiple LLM Judges on a Budget New
NeurIPS 2026
ICML 2026 CTB Workshop
· Best Paper Award (Long Paper track)
Pointwise or Pairwise: When Do Pairwise Losses Help Reward Learning, Provably?
arXiv:2609.37209
A Jointly Efficient and Optimal Algorithm for Heteroskedastic Generalized Linear Bandits with Adversarial Corruptions
arXiv:2602.10971
Preliminary Empirical Study of Low-Rank, Hierarchical Gaussian Linear Bandits
KSC 2025