影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
88.23/100
前 0.7%
全站排名 #450
发表论文33 篇
平均评分
年均产出11.0 篇/年
15
SpatialHand: Generative Object Manipulation from 3D Prespective
ICLR 2026Poster
一作19
Depth Anything with Any Prior
ICLR 2026Poster
一作16
Vox-Infinity: Benchmarking the Limits of Long-Context Spoken Language Models
ICLR 2026Rejected
20
FocusDiff: Advancing Fine-Grained Text-Image Alignment for Autoregressive Visual Generation through RL
ICLR 2026Rejected
17
AlignSep: Temporally-Aligned Video-Queried Sound Separation with Flow Matching
ICLR 2026Poster
5
DSI-Bench: A Benchmark for Dynamic Spatial Intelligence
ICLR 2026Withdrawn
二作6
OmniChat: Enhancing Spoken Dialogue Systems with Scalable Synthetic Data for Diverse Scenarios
ICLR 2026Rejected
21
Orient Anything V2: Unifying Orientation and Rotation Understanding
NeurIPS 2025Spotlight
一作22
VoxDialogue: Can Spoken Dialogue Systems Understand Information Beyond Words?
ICLR 2025Poster
29
WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
ICLR 2025Poster
22
OmniBind: Large-scale Omni Multimodal Representation via Binding Spaces
ICLR 2025Poster
一作25
OmniSep: Unified Omni-Modality Sound Separation with Query-Mixup
ICLR 2025Poster
三作32
Diff-Prompt: Diffusion-Driven Prompt Generator with Mask Supervision
ICLR 2025Poster
27
Improving Long-Text Alignment for Text-to-Image Diffusion Models
ICLR 2025Poster
13
Orient Anything: Learning Robust Object Orientation Estimation from Rendering 3D Models
ICML 2025Poster
一作6
ControlSpeech: Towards Simultaneous Zero-shot Speaker Cloning and Zero-shot Language Style Control
ICLR 2025Withdrawn
23
T2A-Feedback: Improving Basic Capabilities of Text-to-Audio Generation via Fine-grained AI Feedback
ICLR 2025Withdrawn
一作14
OmniChat: Enhancing Spoken Dialogue Systems with Scalable Synthetic Data for Diverse Scenarios
ICLR 2025Withdrawn
5
AVSET-10M: An Open Large-Scale Audio-Visual Dataset with High Correspondence
ICLR 2025Withdrawn
6
Noise-Robust Audio-Visual Speech-Driven Body Language Synthesis
ICLR 2025Withdrawn
三作7
Advancing Multimodal Unified Discrete Representations
ICLR 2025Withdrawn
5
Dynamic Switching Teacher: How to Generalize Temporal Action Detection Models
ICLR 2025Withdrawn
4
MindLoc: A Secure Brain-Based System for Object Localization
ICLR 2025Withdrawn