影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
67.08/100
前 3.2%
全站排名 #2,030
发表论文13 篇
平均评分
年均产出4.3 篇/年
17
UALM: Unified Audio Language Model for Understanding, Generation and Reasoning
ICLR 2026Oral
30
TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization
ICLR 2026Poster
15
OmniVinci: Enhancing Architecture and Data for Omni-Modal Understanding LLM
ICLR 2026Poster
20
Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models
NeurIPS 2025Spotlight
9
Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities
ICML 2025Poster
11
ETTA: Elucidating the Design Space of Text-to-Audio Models
ICML 2025Poster
32
Synthio: Augmenting Small-Scale Audio Classification Datasets with Synthetic Data
ICLR 2025Poster
21
Fugatto 1: Foundational Generative Audio Transformer Opus 1
ICLR 2025Poster
一作25
OMCAT: Omni Context Aware Transformer
ICLR 2025Rejected
22
Elucidating the Design Space of Text-to-Audio Models
ICLR 2025Rejected
17
UniWav: Towards Unified Pre-training for Speech Representation Learning and Generation
ICLR 2025Poster
20
A$^2$-Flow: Alignment-Aware Pre-training for Speech Synthesis with Flow Matching
ICLR 2025Rejected