影响力指数
95.79/100
前 0.2%
全站排名 #150
发表论文68
平均评分4.9
年均产出22.7 篇/年

Di Wang

Assistant Professor@KAUST·沙特阿拉伯·OpenReview
研究方向

interpretability · fairness · learning theory · Differential Privacy

5.5
40

Controlling Repetition in Protein Language Models

ICLR 2026Poster
三作
5.5
30

Dual-Kernel Adapter: Expanding Spatial Horizons for Data-Constrained Medical Image Analysis

ICLR 2026Poster
5.3
23

Predicting LLM Output Length via Entropy-Guided Representations

ICLR 2026Poster
通讯
5.0
24

Understanding and Improving Continuous LLM Adversarial Training via In-context Learning Theory

ICLR 2026Poster
二作
5.0
21

The Price of Amortized inference in Sparse Autoencoders

ICLR 2026Poster
二作
5.0
26

When Modalities Conflict: How Unimodal Reasoning Uncertainty Governs Preference Dynamics in MLLMs

ICLR 2026Rejected
5.0
16

Dissecting Representation Misalignment in Contrastive Learning via Influence Function

ICLR 2026Poster
通讯
4.7
15

Evaluating Data Influence in Meta Learning

ICLR 2026Poster
通讯
4.5
16

Mechanistic Analysis of Demonstration Conflicts in In-Context Learning

ICLR 2026Rejected
二作
4.5
21

Understanding Private Learning From Feature Perspective

ICLR 2026Rejected
4.5
25

Benign Overfitting in Adversarial Training for Vision Transformers

ICLR 2026Rejected
通讯
4.5
11

Untargeted Jailbreak Attack

ICLR 2026Withdrawn
4.5
36

MSRS: Adaptive Multi-Subspace Representation Steering for Attribute Alignment in Large Language Models

ICLR 2026Rejected
4.5
18

MONICA: Real-Time Monitoring and Calibration of Chain-of-Thought Sycophancy in Large Reasoning Models

ICLR 2026Rejected
通讯
4.0
13

Robust Learning of Diffusion Models with Extremely Noisy Conditions

ICLR 2026Rejected
4.0
5

D-LEAF: Localizing and Correcting Hallucinations in Multimodal LLMs via Layer-to-head Attention Diagnostics

ICLR 2026Withdrawn
4.0
5

Efficient and Stable Grouped RL Training for Large Language Models

ICLR 2026Withdrawn
通讯
4.0
16

Dynamic Target Attack

ICLR 2026Rejected
4.0
17

Curriculum-RLAIF: Curriculum Alignment with Reinforcement Learning from AI Feedback

ICLR 2026Withdrawn
通讯
4.0
16

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs

ICLR 2026Withdrawn
通讯
3.6
6

CACE-Net: Cascade Coupling Effect for Link Prediction in Multi-layer Networks

ICLR 2026Withdrawn
3.6
6

Investigating CoT Monitorability in Large Reasoning Models

ICLR 2026Withdrawn
通讯
3.6
6

Attributing Data for Sharpness-Aware Minimization

ICLR 2026Withdrawn
通讯
3.5
5

Pre-trained Encoder Inference: Revealing Upstream Encoders In Downstream Machine Learning Services

ICLR 2026Withdrawn
通讯
3.5
5

TAD-Net: Reinforced Anomaly Generation and Wavelet-enhanced Prediction for Temporal Anomaly Detection

ICLR 2026Withdrawn
通讯
3.5
5

Concept-Based Dictionary Learning for Inference-Time Safety in Vision–Language–Action Models

ICLR 2026Withdrawn
通讯
3.0
12

Flexible Feature Distillation for Large Language Models

ICLR 2026Rejected
二作
3.0
5

CoLa: A Choice Leakage Attack Framework To Expose Privacy Risks In Subset Training

ICLR 2026Withdrawn
通讯
3.0
14

Goal-oriented Backdoor Attack against Vision-Language-Action Models via Physical Objects

ICLR 2026Rejected
3.0
15

Model-agnostic Adversarial Attack and Defense for Vision-Language-Action Models

ICLR 2026Rejected
3.0
5

Backdooring CLIP through Concept Confusion

ICLR 2026Withdrawn
通讯
2.5
5

Compositional Architecture of Regret in Large Language Models

ICLR 2026Withdrawn
通讯
7.8
12

Locate-then-edit for Multi-hop Factual Recall under Knowledge Editing

ICML 2025Poster
通讯
7.3
28

Second-Order Convergence in Private Stochastic Non-Convex Optimization

NeurIPS 2025Poster
通讯
7.2
16

Editable Concept Bottleneck Models

ICML 2025Poster
通讯
6.8
23

Short-length Adversarial Training Helps LLMs Defend Long-length Jailbreak Attacks: Theoretical and Empirical Evidence

NeurIPS 2025Poster
通讯
6.8
26

EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification

NeurIPS 2025Poster
通讯
6.8
35

Private Training Large-scale Models with Efficient DP-SGD

NeurIPS 2025Poster
通讯
6.8
21

Private Stochastic Optimization for Achieving Second-Order Stationary Points

ICLR 2025Rejected
通讯
6.3
11

Scalable Zeroth-Order Fine-Tuning for Extremely Large Language Models with Limited GPU Memory

COLM 2025Poster
通讯
6.3
27

Locate-then-edit for Multi-hop Factual Recall under Knowledge Editing

ICLR 2025Rejected
通讯
5.7
29

Towards User-level Private Reinforcement Learning with Human Feedback

COLM 2025Poster
通讯
5.6
63

Editable Concept Bottleneck Models

ICLR 2025Rejected
通讯
5.3
31

Dissecting Misalignment of Multimodal Large Language Models via Influence Function

ICLR 2025Rejected
通讯
5.0
18

FlashDP: Memory-Efficient and High-Throughput DP-SGD Training for Large Language Models

ICLR 2025Withdrawn
通讯
4.8
9

Private Stochastic Convex Optimization with Tysbakov Noise Condition and Large Lipschitz Constant

ICLR 2025Withdrawn
通讯
4.8
11

Low-cost Enhancer for Text Attributed Graph Learning via Graph Alignment

ICLR 2025Withdrawn
通讯
4.5
15

Representation Confusion: Towards Representation Backdoor on CLIP via Concept Activation

ICLR 2025Rejected
通讯
4.0
5

XTraffic: A Dataset Where Traffic Meets Incidents with Explainability and More

ICLR 2025Withdrawn
3.8
20

ZO-Offloading: Fine-Tuning LLMs with 100 Billion Parameters on a Single GPU

ICLR 2025Withdrawn
通讯
3.5
5

Understanding Reasoning in Chain-of-Thought from the Hopfieldian View

ICLR 2025Withdrawn
通讯
3.0
5

What Makes Your Model a Low-empathy or Warmth Person: Exploring the Origins of Personality in LLMs

ICLR 2025Withdrawn
通讯