影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
94.21/100
前 0.3%
全站排名 #200
发表论文34 篇
平均评分
年均产出11.3 篇/年
Paul Pu Liang
研究方向
multimodal machine learning · multisensory AI · deep learning · human-AI interaction
22
World-In-World: World Models in a Closed-Loop World
ICLR 2026Oral
17
Building Massively Multimodal Foundation Models with Interaction-aware Mixture-of-Experts
ICLR 2026Poster
19
MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents
ICLR 2026Poster
通讯16
PAGE-4D: Disentangled Pose and Geometry Estimation for VGGT-4D Perception
ICLR 2026Poster
20
RAVENEA: A Benchmark for Multimodal Retrieval-Augmented Visual Culture Understanding
ICLR 2026Poster
21
PuzzleWorld: A Benchmark for Multimodal, Open-Ended Reasoning in Puzzlehunts
ICLR 2026Poster
通讯41
Human Behavior Atlas: Benchmarking Unified Psychological And Social Behavior Understanding
ICLR 2026Poster
通讯17
FairGRPO: Fair Reinforcement Learning for Equitable Clinical Reasoning
ICLR 2026Rejected
通讯30
SmellNet: A Dataset for Sensor-Based Smell Recognition and Mixture Prediction
ICLR 2026Poster
通讯18
When Reasoning Meets Its Laws
ICLR 2026Rejected
14
MENTALMs: Teaching Language Models to Reason Efficiently in the Language of Thought
ICLR 2026Desk Rejected
三作20
Learn Globally, Speak Locally: Bridging the Gaps in Multilingual Reasoning
ICLR 2026Rejected
通讯32
HugAgent: Evaluating LLMs in Simulating Human-Like Individual Reasoning on Open-Ended Tasks
ICLR 2026Rejected
22
CTM-AI: A Blueprint for General AI Inspired by Consciousness
ICLR 2026Rejected
通讯5
ChartRef: Benchmarking Fine-Grained Visual Element Localization in Charts
ICLR 2026Withdrawn
二作6
EmoSign: A Multimodal Dataset for Understanding Emotions in American Sign Language
ICLR 2026Rejected
6
MULTI-FACETED MULTIMODAL MONOSEMANTICITY
ICLR 2026Rejected
9
Sotopia-RL: Reward Design for Social Intelligence
ICLR 2026Withdrawn
17
QoQ-Med: Building Multimodal Clinical Foundation Models with Domain-Aware GRPO Training
NeurIPS 2025Oral
通讯18
Balancing Multimodal Training Through Game-Theoretic Regularization
NeurIPS 2025Spotlight
29
Progressive Compositionality in Text-to-Image Generative Models
ICLR 2025Spotlight
通讯17
OS-ATLAS: Foundation Action Model for Generalist GUI Agents
ICLR 2025Spotlight
18
Can A Society of Generative Agents Simulate Human Behavior and Inform Public Health Policy? A Case Study on Vaccine Hesitancy
COLM 2025Poster
23
Partial Information Decomposition via Normalizing Flows in Latent Gaussian Distributions
NeurIPS 2025Poster
通讯30
What One Cannot, Two Can: Two-Layer Transformers Provably Represent Induction Heads on Any-Order Markov Chains
NeurIPS 2025Spotlight
通讯18
REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing
NeurIPS 2025Poster
26
VideoWebArena: Evaluating Long Context Multimodal Agents with Video Understanding Web Tasks
ICLR 2025Poster
17
Understanding the Emergence of Multimodal Representation Alignment
ICML 2025Poster
通讯9
CLIMB: Data Foundations for Large Scale Multimodal Clinical Foundation Models
ICML 2025Poster
通讯8
Video Active Perception: Efficient Inference-Time Long-Form Video Understanding with Vision-Language Models
ICLR 2025Rejected
22
TeaserGen: Generating Teasers for Long Documentaries
ICLR 2025Poster
二作