影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
87.97/100
前 0.7%
全站排名 #462
发表论文32 篇
平均评分
年均产出10.7 篇/年
Minlie Huang
研究方向
natural language generation · dialogue systems · large language models · safety and alignment · social intelligence
23
Trust-Region Adaptive Policy Optimization
ICLR 2026Poster
19
BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs
ICLR 2026Poster
通讯29
SafeSearch: Automated Red-Teaming for the Safety of LLM-Based Search Agents
ICLR 2026Rejected
24
Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen!
ICLR 2026Poster
通讯23
Think Socially via Cognitive Reasoning
ICLR 2026Rejected
通讯18
Survive at All Costs: Exploring LLM's Risky Behavior under Survival Pressure
ICLR 2026Rejected
通讯36
On the Self-awareness of Large Reasoning Models' Capability Boundaries
ICLR 2026Rejected
31
From Theft to Bomb-Making: The Ripple Effect of Unlearning in Defending Against Jailbreak Attacks
ICLR 2026Rejected
通讯21
RLAR: An Agentic Reward System for Multi-task Reinforcement Learning on Large Language Models
ICLR 2026Withdrawn
通讯5
Understanding the Dilemma of Unlearning for Large Language Models
ICLR 2026Withdrawn
5
PROBE: Benchmarking Reasoning Paradigm Overfitting in Large Language Models
ICLR 2026Withdrawn
5
Beyond Nash Equilibrium: Bounded Rationality of LLMs and humans in Strategic Decision-making
ICLR 2026Withdrawn
25
Data Selection via Optimal Control for Language Models
ICLR 2025Oral
通讯28
MAPS: Advancing Multi-Modal Reasoning in Expert-Level Physical Science
ICLR 2025Poster
14
SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
ICLR 2025Poster
通讯23
MiniPLM: Knowledge Distillation for Pre-training Language Models
ICLR 2025Poster
通讯20
Language Models Learn to Mislead Humans via RLHF
ICLR 2025Poster
21
Does RLHF Scale? Exploring the Effects of Data, Model, and Method
ICLR 2025Rejected
27
RedHat: Towards Reducing Hallucination in Essay Critiques with Large Language Models
ICLR 2025Rejected
通讯29
CodePlan: Unlocking Reasoning Potential in Large Language Models by Scaling Code-form Planning
ICLR 2025Poster
通讯10
Diffusion Attacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak
ICLR 2025Withdrawn
14
Model Extrapolation Expedites Alignment
ICLR 2025Withdrawn
23
Seeker: Enhancing Exception Handling in Code with a LLM-based Multi-Agent Approach
ICLR 2025Withdrawn
通讯5
BlackDAN: A Black-Box Multi-Objective Approach to Effective and Contextual Jailbreaking of Language Models
ICLR 2025Rejected
通讯