影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
77.5/100
前 1.6%
全站排名 #1,041
发表论文29 篇
平均评分
年均产出9.7 篇/年
Rongjie Huang
研究方向
Deep generative model · speech · audio · singing
20
PrismAudio: Decomposed Chain-of-Thought and Multi-dimensional Rewards for Video-to-Audio Generation
ICLR 2026Poster
17
AlignSep: Temporally-Aligned Video-Queried Sound Separation with Flow Matching
ICLR 2026Poster
21
ReasonAudio: Semantic Reasoning and Temporal Synchrony in Video–Text-to-Audio Generation
ICLR 2026Withdrawn
一作6
OmniChat: Enhancing Spoken Dialogue Systems with Scalable Synthetic Data for Diverse Scenarios
ICLR 2026Rejected
21
Lumina-T2X: Scalable Flow-based Large Diffusion Transformer for Flexible Resolution Generation
ICLR 2025Spotlight
17
OmniAudio: Generating Spatial Audio from 360-Degree Video
ICML 2025Poster
22
VoxDialogue: Can Spoken Dialogue Systems Understand Information Beyond Words?
ICLR 2025Poster
29
WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
ICLR 2025Poster
22
OmniBind: Large-scale Omni Multimodal Representation via Binding Spaces
ICLR 2025Poster
25
OmniSep: Unified Omni-Modality Sound Separation with Query-Mixup
ICLR 2025Poster
23
T2A-Feedback: Improving Basic Capabilities of Text-to-Audio Generation via Fine-grained AI Feedback
ICLR 2025Withdrawn
14
OmniChat: Enhancing Spoken Dialogue Systems with Scalable Synthetic Data for Diverse Scenarios
ICLR 2025Withdrawn
5
AVSET-10M: An Open Large-Scale Audio-Visual Dataset with High Correspondence
ICLR 2025Withdrawn
6
Noise-Robust Audio-Visual Speech-Driven Body Language Synthesis
ICLR 2025Withdrawn
5
MultiBand: Multi-Task Song Generation with Personalized Prompt-Based Control
ICLR 2025Withdrawn
5
MEDIC: Zero-shot Music Editing with Disentangled Inversion Control
ICLR 2025Withdrawn