影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
83.73/100
前 1.1%
全站排名 #676
发表论文28 篇
平均评分
年均产出9.3 篇/年
Gagandeep Singh
研究方向
Adversarial Robustness · Neural Network Verification · Abstract Interpretation · Formal Methods
15
How Catastrophic is Your LLM? Certifying Risks in Conversation
ICLR 2026Poster
通讯16
Compression Aware Certified Training
ICLR 2026Rejected
二作21
Certifying Robustness of Agent Tool-Selection Under Adversarial Attacks
ICLR 2026Rejected
三作11
Data Shifts Hurt CoT: A Theoretical Study
ICLR 2026Rejected
三作5
Towards Generalized Certified Robustness with Multi-Norm Training
ICLR 2026Withdrawn
三作13
SuperCoder: Assembly Program Superoptimization with Large Language Models
ICLR 2026Rejected
9
Pessimistic Reward Modeling in RLHF against Reward Hacking
ICLR 2026Rejected
通讯5
Deep Safety Alignment Requires Thinking Beyond the Top Token
ICLR 2026Withdrawn
二作27
DINGO: Constrained Inference for Diffusion LLMs
NeurIPS 2025Poster
通讯24
TRAP: Targeted Redirecting of Agentic Preferences
NeurIPS 2025Poster
三作11
IterGen: Iterative Semantic-aware Structured LLM Generation with Backtracking
ICLR 2025Poster
8
CRANE: Reasoning with constrained LLM generation
ICML 2025Poster
通讯21
Certifying Counterfactual Bias in LLMs
ICLR 2025Poster
通讯32
UTF-8 Plumbing: Byte-level Tokenizers Unavoidably Enable LLMs to Generate Ill-formed UTF-8
COLM 2025Poster
三作5
Support is All You Need for Certified VAE Training
ICLR 2025Poster
通讯25
Stochastic Monkeys at Play: Random Augmentations Cheaply Break LLM Safety Alignment
ICLR 2025Withdrawn
通讯15
ARQ: A Mixed-Precision Quantization Framework for Accurate and Certifiably Robust DNNs
ICLR 2025Rejected
22
Decoding Intelligence: A Framework for Certifying Knowledge Comprehension in LLMs
ICLR 2025Rejected
三作16
Robust Thompson Sampling Algorithms Against Reward Poisoning Attacks
ICLR 2025Rejected
三作7
Universal Black-Box Reward Poisoning Attack against Offline Reinforcement Learning
ICLR 2025Withdrawn
三作5
Towards Universal Certified Robustness with Multi-Norm Training
ICLR 2025Withdrawn
二作5
Binary Reward Labeling: Bridging Offline Preference and Reward-Based Reinforcement Learning
ICLR 2025Withdrawn
通讯5
Two-Step Offline Preference-Based Reinforcement Learning with Constrained Actions
ICLR 2025Withdrawn
通讯