影响力指数
97.55/100
前 0.1%
全站排名 #71
发表论文45
平均评分5.9
年均产出15.0 篇/年

Min-hwan Oh

Associate Professor@Seoul National University·韩国·OpenReview
研究方向

Bandit Algorithms · Reinforcement Learning

7.8
27

Exploration via Feature Perturbation in Contextual Bandits

NeurIPS 2025Spotlight
二作
7.8
23

True Impact of Cascade Length in Contextual Cascading Bandits

NeurIPS 2025Poster
三作
7.3
23

Infrequent Exploration in Linear Bandits

NeurIPS 2025Poster
二作
7.3
22

Tractable Multinomial Logit Contextual Bandits with Non-Linear Utilities

NeurIPS 2025Poster
三作
7.3
21

Revisiting Follow-the-Perturbed-Leader with Unbounded Perturbations in Bandit Problems

NeurIPS 2025Poster
通讯
7.2
11

Improved Online Confidence Bounds for Multinomial Logistic Bandits

ICML 2025Poster
二作
7.1
25

Thompson Sampling for Multi-Objective Linear Contextual Bandit

NeurIPS 2025Poster
三作
7.0
27

Minimax Optimal Reinforcement Learning with Quasi-Optimism

ICLR 2025Poster
二作
7.0
17

Adversarial Policy Optimization for Offline Preference-based Reinforcement Learning

ICLR 2025Poster
二作
7.0
15

Oracle-Efficient Combinatorial Semi-Bandits

NeurIPS 2025Poster
三作
6.8
29

Preference-based Reinforcement Learning beyond Pairwise Comparisons: Benefits of Multiple Options

NeurIPS 2025Poster
三作
6.7
16

ADAM Optimization with Adaptive Batch Selection

ICLR 2025Poster
二作
6.5
16

Dynamic Assortment Selection and Pricing with Censored Preference Feedback

ICLR 2025Poster
二作
6.4
29

EUGens: Efficient, Unified and General Dense Layers

NeurIPS 2025Poster
6.3
17

Lasso Bandit with Compatibility Condition on Optimal Arm

ICLR 2025Poster
三作
6.1
11

Combinatorial Reinforcement Learning with Preference Feedback

ICML 2025Poster
二作
6.1
13

Symmetry-Aware GFlowNets

ICML 2025Poster
三作
5.6
30

GFlowNets Need Automorphism Correction for Unbiased Graph Generation

ICLR 2025Rejected
三作
5.5
32

Combinatorial Reinforcement Learning with Preference Feedback

ICLR 2025Rejected
二作
5.5
10

Optimal and Practical Batched Linear Bandit Algorithm

ICML 2025Poster
二作
5.3
21

Magnituder Layers for Implicit Neural Representations in 3D

ICLR 2025Rejected
通讯
5.0
27

Linear Bandits with Partially Observable Features

ICLR 2025Rejected
通讯
4.8
23

Stochastic Matching Bandits under Preference Feedback

ICLR 2025Withdrawn
二作
4.8
9

Linear Bandits with Partially Observable Features

ICML 2025Poster
通讯
4.3
9

Neural Dynamic Pricing: Provable and Practical Efficiency

ICLR 2025Withdrawn
通讯
4.0
14

Mostly Exploration-free Algorithms for Multi-Objective Linear Bandits

ICLR 2025Withdrawn
二作
-1

Coordinated Exploration in Distributed Reinforcement Learning

ICLR 2025Withdrawn
二作