← Search

Zihao Yu

8 accepted papers

2026

4DWorldBench: A Comprehensive Evaluation Framework for 3D/4D World Generation Models

CVPR 2026

World Generation Models are emerging as a cornerstone of next-generation multimodal intelligence systems. Unlike traditional 2D visual generation, World Models aim to construct realistic, dynamic, and physically consistent 3D/4D worlds from images, videos, or text. These models not only need to prod

Cited by 0SourcecodeScholar
2025

Animation Anycolor: Enhancing Line Drawing Colorization with Keypoint Matching

ICASSP 2025accepted

Colorization is a crucial but labor-intensive and time-consuming process of animation production. The automation of animation line-drawing colorization has become a prominent research topic. Recently, methods based on pre-trained text-to-image models have been explored for the task of line-drawing c…

Cited by 0SourceScholar
2025

COSMIC: Generalized Refusal Direction Identification in LLM Activations

ACL 2025finding

Large Language Models encode behaviors like refusal within their activation space, but identifying these behaviors remains challenging. Existing methods depend on predefined refusal templates detectable in output tokens or manual review. We introduce **COSMIC** (Cosine Similarity Metrics for Inversi…

2025

QMamba: On First Exploration of Vision Mamba for Image Quality Assessment

ICML 2025poster

In this work, we take the first exploration of the recently popular foundation model, *i.e.,* State Space Model/Mamba, in image quality assessment (IQA), aiming at observing and excavating the perception potential in vision Mamba. A series of works on Mamba has shown its significant potential in va…

2025

Unifews: You Need Fewer Operations for Efficient Graph Neural Networks

ICML 2025poster

Graph Neural Networks (GNNs) have shown promising performance, but at the cost of resource-intensive operations on graph-scale matrices. To reduce computational overhead, previous studies attempt to sparsify the graph or network parameters, but with limited flexibility and precision boundaries. In t…

Cited by 0SourcePDFScholar
2024

Accelerating Text-to-Image Editing via Cache-Enabled Sparse Diffusion Inference

AAAI 2024technical

Due to the recent success of diffusion models, text-to-image generation is becoming increasingly popular and achieves a wide range of applications. Among them, text-to-image editing, or continuous text-to-image generation, attracts lots of attention and can potentially improve the quality of generat…

2023

AdaSfM: From Coarse Global to Fine Incremental Adaptive Structure from Motion

ICRA 2023poster

Despite the impressive results achieved by many existing Structure from Motion (SfM) approaches, there is still a need to improve the robustness, accuracy, and efficiency on large-scale scenes with many outlier matches and sparse view graphs. In this paper, we propose AdaSfM: a coarse-to-fine adapti…

Cited by 10SourceScholar