影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
95.79/100
前 0.2%
全站排名 #150
发表论文68 篇
平均评分
年均产出22.7 篇/年
Di Wang
研究方向
interpretability · fairness · learning theory · Differential Privacy
40
Controlling Repetition in Protein Language Models
ICLR 2026Poster
三作30
Dual-Kernel Adapter: Expanding Spatial Horizons for Data-Constrained Medical Image Analysis
ICLR 2026Poster
23
Predicting LLM Output Length via Entropy-Guided Representations
ICLR 2026Poster
通讯24
Understanding and Improving Continuous LLM Adversarial Training via In-context Learning Theory
ICLR 2026Poster
二作21
The Price of Amortized inference in Sparse Autoencoders
ICLR 2026Poster
二作26
When Modalities Conflict: How Unimodal Reasoning Uncertainty Governs Preference Dynamics in MLLMs
ICLR 2026Rejected
16
Dissecting Representation Misalignment in Contrastive Learning via Influence Function
ICLR 2026Poster
通讯15
Evaluating Data Influence in Meta Learning
ICLR 2026Poster
通讯16
Mechanistic Analysis of Demonstration Conflicts in In-Context Learning
ICLR 2026Rejected
二作21
Understanding Private Learning From Feature Perspective
ICLR 2026Rejected
25
Benign Overfitting in Adversarial Training for Vision Transformers
ICLR 2026Rejected
通讯11
Untargeted Jailbreak Attack
ICLR 2026Withdrawn
36
MSRS: Adaptive Multi-Subspace Representation Steering for Attribute Alignment in Large Language Models
ICLR 2026Rejected
18
MONICA: Real-Time Monitoring and Calibration of Chain-of-Thought Sycophancy in Large Reasoning Models
ICLR 2026Rejected
通讯13
Robust Learning of Diffusion Models with Extremely Noisy Conditions
ICLR 2026Rejected
5
D-LEAF: Localizing and Correcting Hallucinations in Multimodal LLMs via Layer-to-head Attention Diagnostics
ICLR 2026Withdrawn
5
Efficient and Stable Grouped RL Training for Large Language Models
ICLR 2026Withdrawn
通讯16
Dynamic Target Attack
ICLR 2026Rejected
17
Curriculum-RLAIF: Curriculum Alignment with Reinforcement Learning from AI Feedback
ICLR 2026Withdrawn
通讯16
Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs
ICLR 2026Withdrawn
通讯6
CACE-Net: Cascade Coupling Effect for Link Prediction in Multi-layer Networks
ICLR 2026Withdrawn
6
Investigating CoT Monitorability in Large Reasoning Models
ICLR 2026Withdrawn
通讯6
Attributing Data for Sharpness-Aware Minimization
ICLR 2026Withdrawn
通讯5
Pre-trained Encoder Inference: Revealing Upstream Encoders In Downstream Machine Learning Services
ICLR 2026Withdrawn
通讯5
TAD-Net: Reinforced Anomaly Generation and Wavelet-enhanced Prediction for Temporal Anomaly Detection
ICLR 2026Withdrawn
通讯5
Concept-Based Dictionary Learning for Inference-Time Safety in Vision–Language–Action Models
ICLR 2026Withdrawn
通讯12
Flexible Feature Distillation for Large Language Models
ICLR 2026Rejected
二作5
CoLa: A Choice Leakage Attack Framework To Expose Privacy Risks In Subset Training
ICLR 2026Withdrawn
通讯14
Goal-oriented Backdoor Attack against Vision-Language-Action Models via Physical Objects
ICLR 2026Rejected
15
Model-agnostic Adversarial Attack and Defense for Vision-Language-Action Models
ICLR 2026Rejected
5
Backdooring CLIP through Concept Confusion
ICLR 2026Withdrawn
通讯5
Compositional Architecture of Regret in Large Language Models
ICLR 2026Withdrawn
通讯12
Locate-then-edit for Multi-hop Factual Recall under Knowledge Editing
ICML 2025Poster
通讯28
Second-Order Convergence in Private Stochastic Non-Convex Optimization
NeurIPS 2025Poster
通讯16
Editable Concept Bottleneck Models
ICML 2025Poster
通讯23
Short-length Adversarial Training Helps LLMs Defend Long-length Jailbreak Attacks: Theoretical and Empirical Evidence
NeurIPS 2025Poster
通讯26
EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification
NeurIPS 2025Poster
通讯35
Private Training Large-scale Models with Efficient DP-SGD
NeurIPS 2025Poster
通讯21
Private Stochastic Optimization for Achieving Second-Order Stationary Points
ICLR 2025Rejected
通讯11
Scalable Zeroth-Order Fine-Tuning for Extremely Large Language Models with Limited GPU Memory
COLM 2025Poster
通讯27
Locate-then-edit for Multi-hop Factual Recall under Knowledge Editing
ICLR 2025Rejected
通讯29
Towards User-level Private Reinforcement Learning with Human Feedback
COLM 2025Poster
通讯63
Editable Concept Bottleneck Models
ICLR 2025Rejected
通讯31
Dissecting Misalignment of Multimodal Large Language Models via Influence Function
ICLR 2025Rejected
通讯18
FlashDP: Memory-Efficient and High-Throughput DP-SGD Training for Large Language Models
ICLR 2025Withdrawn
通讯9
Private Stochastic Convex Optimization with Tysbakov Noise Condition and Large Lipschitz Constant
ICLR 2025Withdrawn
通讯11
Low-cost Enhancer for Text Attributed Graph Learning via Graph Alignment
ICLR 2025Withdrawn
通讯15
Representation Confusion: Towards Representation Backdoor on CLIP via Concept Activation
ICLR 2025Rejected
通讯5
XTraffic: A Dataset Where Traffic Meets Incidents with Explainability and More
ICLR 2025Withdrawn
20
ZO-Offloading: Fine-Tuning LLMs with 100 Billion Parameters on a Single GPU
ICLR 2025Withdrawn
通讯5
Understanding Reasoning in Chain-of-Thought from the Hopfieldian View
ICLR 2025Withdrawn
通讯5
What Makes Your Model a Low-empathy or Warmth Person: Exploring the Origins of Personality in LLMs
ICLR 2025Withdrawn
通讯