影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
68.17/100
前 2.9%
全站排名 #1,893
发表论文22 篇
平均评分
年均产出7.3 篇/年
Qingwei Lin
研究方向
Reinforcement Learning · LLM · GPT · Deep Learning · AIOps · Cloud Intelligence · Machine learning · Data mining · Cloud systems · Log Analysis · Anomaly Detection · Diagnosis · Root cause analysis
21
Pretrain Value, Not Reward: Decoupled Value Policy Optimization
ICLR 2026Poster
23
DoVer: Intervention-Driven Auto Debugging for LLM Multi-Agent Systems
ICLR 2026Poster
19
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
ICLR 2026Poster
30
Text2Grad: Reinforcement Learning from Natural Language Feedback
ICLR 2026Poster
16
WarriorMath: Empowering Mathematical Reasoning for Large Language Models via Expert Battles
ICLR 2026Rejected
14
Memory-Augmented Personalized Retrieval for Long-Context Egocentric Video
ICLR 2026Withdrawn
22
GUI-360° : A Comprehensive Dataset and Benchmark for Computer-Using Agents
ICLR 2026Rejected
46
G-KV: Decoding-Time KV Cache Eviction with Global Attention
ICLR 2026Withdrawn
14
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
ICLR 2026Withdrawn
5
Duet: Joint Exploration of User–Item Profiles
ICLR 2026Withdrawn
6
VEM: Environment-Free Exploration for Training GUI Agent with Value Environment Model
ICLR 2026Withdrawn
14
Distill Not Only Data but Also Rewards: Can Smaller Language Models Surpass Larger Ones?
ICLR 2026Rejected
30
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
ICLR 2025Oral
25
GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents
NeurIPS 2025Poster
40
OpenRCA: Can Large Language Models Locate the Root Cause of Software Failures?
ICLR 2025Poster
20
RuAG: Learned-rule-augmented Generation for Large Language Models
ICLR 2025Poster
37
SELF-EVOLVED REWARD LEARNING FOR LLMS
ICLR 2025Poster
29
Thread: A Logic-Based Data Organization Paradigm for How-To Question Answering with Retrieval Augmented Generation
ICLR 2025Rejected
6
Zoomer: Enhancing MLLM Performance with Adaptive Image Focus Optimization
ICLR 2025Withdrawn