影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
45.5/100
前 10.7%
全站排名 #6,877
发表论文16 篇
平均评分
年均产出5.3 篇/年
22
VoxDialogue: Can Spoken Dialogue Systems Understand Information Beyond Words?
ICLR 2025Poster
29
WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
ICLR 2025Poster
一作22
OmniBind: Large-scale Omni Multimodal Representation via Binding Spaces
ICLR 2025Poster
25
OmniSep: Unified Omni-Modality Sound Separation with Query-Mixup
ICLR 2025Poster
46
Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
ICLR 2025Withdrawn
32
MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization
ICLR 2025Rejected
6
ControlSpeech: Towards Simultaneous Zero-shot Speaker Cloning and Zero-shot Language Style Control
ICLR 2025Withdrawn
一作23
T2A-Feedback: Improving Basic Capabilities of Text-to-Audio Generation via Fine-grained AI Feedback
ICLR 2025Withdrawn
14
OmniChat: Enhancing Spoken Dialogue Systems with Scalable Synthetic Data for Diverse Scenarios
ICLR 2025Withdrawn
17
IRBridge: Solving Image Restoration Bridge with Pre-trained Generative Diffusion Models
ICML 2025Poster
7
Advancing Multimodal Unified Discrete Representations
ICLR 2025Withdrawn
三作4
MindLoc: A Secure Brain-Based System for Object Localization
ICLR 2025Withdrawn