影响力指数
97.8/100
前 0.1%
全站排名 #63
发表论文39
平均评分6.4
年均产出13.0 篇/年

Bryan Catanzaro

Vice President@NVIDIA·美国·OpenReview
研究方向

speech recognition · deep learning · machine learning systems

8.7
20

Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models

NeurIPS 2025Spotlight
通讯
8.2
14

Prismatic Synthesis: Gradient-based Data Diversification Boosts Generalization in LLM Reasoning

NeurIPS 2025Spotlight
8.2
15

Efficient Hybrid Language Model Compression through Group-Aware SSM Pruning

NeurIPS 2025Poster
7.8
9

Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities

ICML 2025Poster
通讯
7.5
14

NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models

ICLR 2025Spotlight
7.3
22

AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning

NeurIPS 2025Poster
7.2
33

Eagle: Exploring The Design Space for Multimodal LLMs with Mixture of Encoders

ICLR 2025Spotlight
7.2
11

ETTA: Elucidating the Design Space of Text-to-Audio Models

ICML 2025Poster
通讯
6.8
20

Eagle 2.5: Boosting Long-Context Post-Training for Frontier Vision-Language Models

NeurIPS 2025Poster
6.8
32

Synthio: Augmenting Small-Scale Audio Classification Datasets with Synthetic Data

ICLR 2025Poster
6.7
21

Fugatto 1: Foundational Generative Audio Transformer Opus 1

ICLR 2025Poster
通讯
6.6
12

FeatSharp: Your Vision Model Features, Sharper

ICML 2025Poster
6.5
22

MM-EMBED: UNIVERSAL MULTIMODAL RETRIEVAL WITH MULTIMODAL LLMS

ICLR 2025Poster
6.0
22

Elucidating the Design Space of Text-to-Audio Models

ICLR 2025Rejected
通讯
6.0
25

OMCAT: Omni Context Aware Transformer

ICLR 2025Rejected
通讯
6.0
17

UniWav: Towards Unified Pre-training for Speech Representation Learning and Generation

ICLR 2025Poster
通讯
6.0
6

MIND: Math Informed syNthetic Dialogues for Pretraining LLMs

ICLR 2025Poster
通讯
5.8
27

ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities

ICLR 2025Poster
通讯
5.5
20

A$^2$-Flow: Alignment-Aware Pre-training for Speech Synthesis with Flow Matching

ICLR 2025Rejected
通讯
5.3
25

PHI-S: Distribution Balancing for Agglomerative Models

ICLR 2025Rejected
5.0
11

LLM Pruning and Distillation in Practice

ICLR 2025Rejected
通讯
4.8
7

Nemotron-CORTEXA: Enhancing LLM Agents for Software Engineering Tasks via Improved Localization and Solution Diversity

ICML 2025Poster
通讯