影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
69.68/100
前 2.7%
全站排名 #1,743
发表论文25 篇
平均评分
年均产出8.3 篇/年
Xize Cheng
研究方向
Spoken dialogue systems · Omni-modal understanding
15
SpatialHand: Generative Object Manipulation from 3D Prespective
ICLR 2026Poster
16
Vox-Infinity: Benchmarking the Limits of Long-Context Spoken Language Models
ICLR 2026Rejected
一作15
MARS-Sep: Multimodal-Aligned Reinforced Sound Separation
ICLR 2026Poster
二作17
AlignSep: Temporally-Aligned Video-Queried Sound Separation with Flow Matching
ICLR 2026Poster
一作21
WavReward: Spoken Dialogue Models With Generalist Reward Evaluators
ICLR 2026Rejected
6
OmniChat: Enhancing Spoken Dialogue Systems with Scalable Synthetic Data for Diverse Scenarios
ICLR 2026Rejected
一作5
Character Beyond Speech: Leveraging Role-Playing Evaluation in Large Audio Language Models via Reinforcement Learning
ICLR 2026Withdrawn
二作22
VoxDialogue: Can Spoken Dialogue Systems Understand Information Beyond Words?
ICLR 2025Poster
一作29
WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
ICLR 2025Poster
22
OmniBind: Large-scale Omni Multimodal Representation via Binding Spaces
ICLR 2025Poster
25
OmniSep: Unified Omni-Modality Sound Separation with Query-Mixup
ICLR 2025Poster
一作32
MuVi: Video-to-Music Generation with Semantic Alignment and Rhythmic Synchronization
ICLR 2025Rejected
三作6
ControlSpeech: Towards Simultaneous Zero-shot Speaker Cloning and Zero-shot Language Style Control
ICLR 2025Withdrawn
14
OmniChat: Enhancing Spoken Dialogue Systems with Scalable Synthetic Data for Diverse Scenarios
ICLR 2025Withdrawn
一作23
T2A-Feedback: Improving Basic Capabilities of Text-to-Audio Generation via Fine-grained AI Feedback
ICLR 2025Withdrawn
5
AVSET-10M: An Open Large-Scale Audio-Visual Dataset with High Correspondence
ICLR 2025Withdrawn
一作6
Noise-Robust Audio-Visual Speech-Driven Body Language Synthesis
ICLR 2025Withdrawn
一作5
Dynamic Switching Teacher: How to Generalize Temporal Action Detection Models
ICLR 2025Withdrawn
4
MindLoc: A Secure Brain-Based System for Object Localization
ICLR 2025Withdrawn
二作