影响力指数
94.87/100
前 0.3%
全站排名 #174
发表论文35
平均评分5.6
年均产出11.7 篇/年

Ruoxi Jia

Assistant Professor@Virginia Tech·美国·OpenReview
研究方向

Machine learning · Security · Privacy

8.0
27

Capturing the Temporal Dependence of Training Data Influence

ICLR 2025Oral
通讯
7.5
30

Data Shapley in One Training Run

ICLR 2025Oral
通讯
7.5
15

AIR-BENCH 2024: A Safety Benchmark based on Regulation and Policies Specified Risk Categories

ICLR 2025Spotlight
7.3
23

LLM Can be a Dangerous Persuader: Empirical Study of Persuasion Safety in Large Language Models

COLM 2025Poster
6.8
24

SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal

ICLR 2025Poster
6.5
21

Data-Centric Human Preference with Rationales for Direct Preference Alignment

COLM 2025Poster
通讯
6.5
42

Mind Control through Causal Inference: Predicting Clean Images from Poisoned Data

ICLR 2025Poster
6.4
31

LLMs Can Plan Only If We Tell Them

ICLR 2025Poster
二作
6.4
21

Probing Hidden Knowledge Holes in Unlearned LLMs

NeurIPS 2025Poster
通讯
6.1
12

LLMs Can Reason Faster Only If We Let Them

ICML 2025Poster
5.5
13

Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning

ICML 2025Poster
通讯
5.5
19

AutoScale: Automatic Prediction of Compute-optimal Data Compositions for Training LLMs

ICLR 2025Rejected
通讯
5.3
26

AutoScale: Scale-Aware Data Mixing for Pre-Training LLMs

COLM 2025Poster
通讯
5.3
22

LLM Spark: Critical Thinking Evaluation of Large Language Models

ICLR 2025Rejected
5.3
37

Data-Centric Human Preference Optimization with Rationales

ICLR 2025Rejected
通讯
5.0
15

SCOPE: Scalable and Adaptive Evaluation of Misguided Safety Refusal in LLMs

ICLR 2025Rejected
通讯
4.8
18

Fast and Noise-Robust Diffusion Solvers for Inverse Problems: A Frequentist Approach

ICLR 2025Rejected
4.5
9

CONCORD: Concept-informed Diffusion for Dataset Distillation

ICLR 2025Withdrawn
三作