影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
81.56/100
前 1.2%
全站排名 #795
发表论文21 篇
平均评分
年均产出7.0 篇/年
Hongning Wang
研究方向
bandit learning · learning to rank · information retrieval · text mining · machine learning
23
Trust-Region Adaptive Policy Optimization
ICLR 2026Poster
通讯19
BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs
ICLR 2026Poster
21
AgentRL: Scaling Agentic Reinforcement Learning with a Multi-Turn, Multi-Task Framework
ICLR 2026Rejected
24
Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen!
ICLR 2026Poster
23
Think Socially via Cognitive Reasoning
ICLR 2026Rejected
31
From Theft to Bomb-Making: The Ripple Effect of Unlearning in Defending Against Jailbreak Attacks
ICLR 2026Rejected
21
RLAR: An Agentic Reward System for Multi-task Reinforcement Learning on Large Language Models
ICLR 2026Withdrawn
5
Beyond Nash Equilibrium: Bounded Rationality of LLMs and humans in Strategic Decision-making
ICLR 2026Withdrawn
通讯25
Data Selection via Optimal Control for Language Models
ICLR 2025Oral
三作28
MAPS: Advancing Multi-Modal Reasoning in Expert-Level Physical Science
ICLR 2025Poster
通讯14
SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
ICLR 2025Poster
18
RecFlow: An Industrial Full Flow Recommendation Dataset
ICLR 2025Poster
21
Does RLHF Scale? Exploring the Effects of Data, Model, and Method
ICLR 2025Rejected
27
RedHat: Towards Reducing Hallucination in Essay Critiques with Large Language Models
ICLR 2025Rejected
29
CodePlan: Unlocking Reasoning Potential in Large Language Models by Scaling Code-form Planning
ICLR 2025Poster
三作5
Cost-Efficient Multi-Fidelity Alignment for LLMs
ICLR 2025Rejected