影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
95.8/100
前 0.2%
全站排名 #147
发表论文46 篇
平均评分
年均产出15.3 篇/年
Yue Zhang
研究方向
LLM · OOD Generalization · Reasoning · Summarization · Semantic parsing · Information Extraction · Generation · Machine Translation · Sentiment · Syntactic Parsing
16
TrustJudge: Inconsistencies of LLM-as-a-Judge and How to Alleviate Them
ICLR 2026Poster
24
Beyond English-Centric Training: How Reinforcement Learning Improves Cross-Lingual Reasoning in LLMs
ICLR 2026Poster
通讯13
Video Summarization Pretraining with Self-Discovery of Informative Frames
ICLR 2026Withdrawn
三作21
Video Language Models are Human-Aligned Evaluators for Text to Motion Generation
ICLR 2026Withdrawn
三作13
FiRE: Fine-Grained Ranking Evaluation for Machine Translation
ICLR 2026Rejected
通讯16
SafeReview: Building a Robust Deep Review Assistant Against Prompt Injection
ICLR 2026Rejected
22
Deep Literature Survey Automation with an Iterative Workflow
ICLR 2026Rejected
通讯31
DeepScientist: Advancing Frontier-Pushing Scientific Findings Progressively
ICLR 2026Poster
通讯24
RewardAnything: Generalizable Principle-Following Reward Models
ICLR 2026Rejected
47
AutoFigure: Generating and Refining Publication-Ready Scientific Illustrations
ICLR 2026Poster
通讯5
Safetylock: Guarding LLM againt FuneTuning Risks with Efficient Inference-time Addon
ICLR 2026Withdrawn
通讯5
Res-Bench: Reasoning Skill-Aware Reasoning Diagnostic evaluation benchmark for math reasoning
ICLR 2026Withdrawn
5
League: Leaderboard Generation on Demand
ICLR 2026Withdrawn
通讯5
Evaluating the Logical Reasoning Abilities of Large Reasoning Models
ICLR 2026Rejected
通讯14
MMQA: Evaluating LLMs with Multi-Table Multi-Hop Complex Questions
ICLR 2025Oral
通讯14
Learning to Reason under Off-Policy Guidance
NeurIPS 2025Poster
通讯5
Task Calibration: Calibrating Large Language Models on Inference Tasks
ICLR 2025Rejected
通讯28
ELICIT: LLM Augmentation Via External In-context Capability
ICLR 2025Poster
三作29
CycleResearcher: Improving Automated Research via Automated Review
ICLR 2025Poster
27
NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens
ICLR 2025Poster
通讯30
Glimpse: Enabling White-Box Methods to Use Proprietary Models for Zero-Shot LLM-Generated Text Detection
ICLR 2025Poster
通讯28
Personality Alignment of Large Language Models
ICLR 2025Poster
通讯63
Towards Homogeneous Lexical Tone Decoding from Heterogeneous Intracranial Recordings
ICLR 2025Poster
17
CofCA: A STEP-WISE Counterfactual Multi-hop QA benchmark
ICLR 2025Poster
通讯22
Learning to Rank for In-Context Example Retrieval
NeurIPS 2025Poster
通讯22
An Empirical Analysis of Uncertainty in Large Language Model Evaluations
ICLR 2025Poster
29
Direct Preference Optimization Using Sparse Feature-level Constraints
ICLR 2025Rejected
41
Human Simulacra: Benchmarking the Personification of Large Language Models
ICLR 2025Poster
通讯12
Constrain Alignment with Sparse Autoencoders
ICML 2025Poster
15
BALCONI: BALancing CONtext and Internal Knowledge For Training Flexible LLMs
ICLR 2025Rejected
通讯24
Locking Down the Finetuned LLMs Safety
ICLR 2025Rejected
通讯5
Keys to Robust Edits: From Theoretical Insights to Practical Advances
ICLR 2025Withdrawn
通讯5
Logic Agent: Enhancing Validity with Logic Rule Invocation
ICLR 2025Withdrawn
三作12
An Efficient Quantum Classifier Based on Hamiltonian Representations
ICLR 2025Withdrawn
三作