影响力指数
96.93/100
前 0.2%
全站排名 #98
发表论文58
平均评分5.2
年均产出19.3 篇/年

Yi Yang

Full Professor@Zhejiang University·中国·OpenReview
研究方向

deep learning · computer vision · machine perception · 3d vision · multi-modal learning

7.0
18

Uncertainty-Aware 3D Reconstruction for Dynamic Underwater Scenes

ICLR 2026Poster
6.0
21

LogiStory: A Logic-Aware Framework for Multi-Image Story Visualization

ICLR 2026Poster
5.5
28

Endowing GPT-4 with a Humanoid Body: Building the Bridge Between Off-the-Shelf VLMs and the Physical World

ICLR 2026Poster
三作
5.5
24

Structured Reasoning for LLMs: A Unified Framework for Efficiency and Explainability

ICLR 2026Poster
通讯
5.5
19

Open-Set Semantic Gaussian Splatting SLAM with Expandable Representation

ICLR 2026Poster
通讯
5.5
14

Frequency-aware Dynamic Gaussian Splatting

ICLR 2026Poster
5.5
24

CogFlow: Bridging Perception and Reasoning through Knowledge Internalization for Visual Mathematical Problem Solving

ICLR 2026Poster
5.3
20

Uncertainty-Aware Gaussian Map for Vision-Language Navigation

ICLR 2026Poster
5.0
22

ContextGen: Contextual Layout Anchoring for Identity-Consistent Multi-Instance Generation

ICLR 2026Poster
通讯
5.0
17

MoDE: Weight Denoising Towards Better LLM Performance through a Mixture of Domain Experts

ICLR 2026Desk Rejected
通讯
5.0
20

Lumos-1: On Autoregressive Video Generation with Discrete Diffusion from a Unified Model Perspective

ICLR 2026Poster
通讯
4.7
17

SEED-GRPO: Semantic Entropy Enhanced GRPO for Uncertainty-Aware Policy Optimization

ICLR 2026Rejected
通讯
4.5
24

Translution: Unifying Self-attention and Convolution for Adaptive and Relative Modeling

ICLR 2026Rejected
二作
4.5
22

Restore3D: Breathing Life into Broken Objects with Shape and Texture Restoration

ICLR 2026Rejected
三作
4.5
25

OSCAgent: Closing the Loop in Organic Solar Cell Discovery with LLM Agents

ICLR 2026Withdrawn
通讯
4.5
31

Stroke3D: Lifting 2D strokes into rigged 3D model via latent diffusion models

ICLR 2026Poster
通讯
4.5
25

BideDPO: Conditional Image Generation with Simultaneous Text and Condition Alignment

ICLR 2026Poster
通讯
4.3
24

Moving Beyond Diffusion: Hierarchy-to-Hierarchy Autoregression for fMRI-to-Image Reconstruction

ICLR 2026Poster
通讯
4.0
15

Learning from Reference Answers: Versatile Language Model Alignment without Binary Human Preference Data

ICLR 2026Withdrawn
通讯
4.0
17

Dynamic Experts Search: Enhancing Reasoning in Mixture-of-Experts LLMs at Test Time

ICLR 2026Withdrawn
通讯
4.0
4

RegionDoc-R1: Reinforcing Semantic Layout-Aware Learning for Document Understanding

ICLR 2026Withdrawn
通讯
4.0
19

CoMo: Compositional Motion Customization for Text-to-Video Generation

ICLR 2026Rejected
3.5
5

Improving Protein Sequence Design through Designability Preference Optimization

ICLR 2026Withdrawn
3.5
18

PRICIN: Principle-Centered Inorganic Retrosynthesis

ICLR 2026Withdrawn
3.0
5

Diffusion-Model Layers May Exhibit Diffusive Behavior at Each Step for Noise Estimation

ICLR 2026Withdrawn
三作
2.7
4

Hierarchical Speculative Decoding through Training-Free Slim-Verifier

ICLR 2026Withdrawn
通讯
2.0
5

StrucBooth: Structural Gradient Supervised Tuning for Enhanced Portrait Animation

ICLR 2026Withdrawn
通讯
7.3
23

Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer

NeurIPS 2025Poster
通讯
7.3
25

SparseDiT: Token Sparsification for Efficient Diffusion Transformer

NeurIPS 2025Poster
通讯
6.8
22

3DID: Direct 3D Inverse Design for Aerodynamics with Physics-Aware Optimization

NeurIPS 2025Poster
三作
6.6
13

DreamDPO: Aligning Text-to-3D Generation with Human Preferences via Direct Preference Optimization

ICML 2025Poster
6.4
23

DeltaPhi: Physical States Residual Learning for Neural Operators in Data-Limited PDE Solving

NeurIPS 2025Poster
二作
6.4
26

Restore3D: Breathing Life into Broken Objects with Shape and Texture Restoration

NeurIPS 2025Rejected
三作
6.4
23

FlexSelect: Flexible Token Selection for Efficient Long Video Understanding

NeurIPS 2025Poster
6.3
19

Hydra-SGG: Hybrid Relation Assignment for One-stage Scene Graph Generation

ICLR 2025Poster
通讯
6.1
11

Holistic Physics Solver: Learning PDEs in a Unified Spectral-Physical Space

ICML 2025Poster
二作
6.1
15

Origin Identification for Text-Guided Image-to-Image Diffusion Models

ICML 2025Poster
通讯
6.0
84

Reaction Graph: Toward Modeling Chemical Reactions with 3D Molecular Structures

ICLR 2025Rejected
通讯
5.5
31

Point-Calibrated Spectral Neural Operators

ICLR 2025Rejected
三作
5.5
15

Reaction Graph: Towards Reaction-Level Modeling for Chemical Reactions with 3D Structures

ICML 2025Poster
通讯
5.2
6

Collaborative Hybrid Propagator for Temporal Misalignment in Audio-Visual Segmentation

ICLR 2025Withdrawn
5.0
16

HYBRID MODEL COLLABORATION FOR SIGN LANGUAGE TRANSLATION WITH VQ-VAE AND RAG ENHANCED LLMS

ICLR 2025Withdrawn
三作
5.0
5

Generalizable Origin Identification for Text-Guided Image-to-Image Diffusion Models

ICLR 2025Withdrawn
通讯
4.8
5

Open-Ended 3D Metric-Semantic Representation Learning via Semantic-Embedded Gaussian Splatting

ICLR 2025Withdrawn
通讯
4.5
5

Towards Human-like Virtual Beings: Simulating Human Behavior in 3D Scenes

ICLR 2025Withdrawn
通讯
4.4
6

Protecting Copyrighted Material with Unique Identifiers in Large Language Model Training

ICLR 2025Withdrawn
通讯
3.4
6

ProteinAdapter: Adapting Pre-trained Large Protein Models for Efficient Protein Representation Learning

ICLR 2025Withdrawn
通讯