影响力指数
论文质量、代表作、近期表现、广度与样本量置信度综合计算
79.87/100
前 1.4%
全站排名 #904
发表论文31 篇
平均评分
年均产出10.3 篇/年
Yun Fu
研究方向
Machine Learning · Computer Vision · LLM · VLM
20
GmNet: Revisiting Gating Mechanisms From A Frequency View
ICLR 2026Poster
通讯22
SHIELD: Suppressing Hallucinations In LVLM Encoders via Bias and Vulnerability Defense
ICLR 2026Poster
通讯19
Seeing Through Words: Controlling Visual Retrieval Quality with Language Models
ICLR 2026Poster
通讯20
Ref-Adv: Exploring MLLM Visual Reasoning in Referring Expression Tasks
ICLR 2026Poster
通讯5
ARGen-Dexion: Autoregressive Image Generation Made Stronger by Vision Decoder
ICLR 2026Withdrawn
通讯23
On Computation and Generalization of Group Relative Policy Optimization
ICLR 2026Rejected
二作24
CompSRT: Quantization and Pruning for Image Super Resolution Transformers
ICLR 2026Rejected
通讯21
Boosting Large Language Models with Mask Fine-Tuning
ICLR 2026Rejected
通讯4
CoT Referring: Improving Localization with Grounded Reasoning in Referring Expression Tasks
ICLR 2026Desk Rejected
通讯5
Token-Shuffle: Towards High-Resolution Image Generation with Autoregressive Models
ICLR 2026Rejected
通讯5
SSTP: Efficient Sample Selection for Trajectory Prediction
ICLR 2026Withdrawn
三作22
Multimodal Generative Composition Recommendation
ICLR 2026Rejected
通讯5
Dense Video Understanding with Gated Residual Tokenization
ICLR 2026Withdrawn
通讯26
DiffuPhyGS: Text-to-Video Generation with 3D Gaussians and Learnable Physical Properties via Diffusion Priors
ICLR 2026Withdrawn
二作6
I$^2$BQ: Quantizing LLMs via Intra- and Inter-Block Optimization
ICLR 2026Withdrawn
通讯29
Accessing Vision Foundation Models via ImageNet-1K
ICLR 2025Poster
通讯28
Sports-Traj: A Unified Trajectory Generation Model for Multi-Agent Movement in Sports
ICLR 2025Poster
二作41
VQToken: Neural Discrete Token Representation Learning for Extreme Token Reduction in Video Large Language Models
NeurIPS 2025Poster
二作23
The Indra Representation Hypothesis
NeurIPS 2025Poster
通讯26
Scale-Free Graph-Language Models
ICLR 2025Poster
通讯6
VaQuitA: Enhancing Alignment in LLM-Assisted Zero-Shot Video Understanding
ICLR 2025Withdrawn
5
Understanding, Abstracting and Checking: Evoking Complicated Multimodal Reasoning in LMMs
ICLR 2025Withdrawn
二作4
Seeing is Knowing: Advancing Semantic Understanding with MLLMs in Grounding Tasks
ICLR 2025Withdrawn
通讯6
Improve Code Generation with Feedback
ICLR 2025Rejected
二作8
improve weakly supervised visual grounding by learning where to focus on
ICLR 2025Rejected
二作