← Search

Chao Zhou

46 accepted papers

2026

Bridging Domains through Subspace-Aware Model Merging

CVPR 2026

Model merging integrates multiple task-specific models into a single consolidated one. Recent research has made progress in improving merging performance for in-distribution or multi-task scenarios, but domain generalization in model merging remains underexplored. We investigate how merging models f

Cited by 0SourcecodeScholar
2026

Coloring the Noise: Adversarial Sobolev Alignment for Faithful Image Super Resolution

ICML 2026poster

Generative priors in Image Super-Resolution (SR) often compromise faithful restoration, we attribute this limitation to a fundamental spectral misalignment between isotropic objectives and the intrinsic natural image manifold. While Direct Preference Optimization offers a path to alignment, its reli…

Cited by 0SourceScholar
2026

Noise Suppression Method of Passive Electric Sense System for Robotic Shark

RA-L 2026

Passive electric sense, a unique underwater perception ability in aquatic species, enables fish to detect external electrical signals and perform essential life activities in dark or turbid environments. Inspired by the passive electrosensing capability of natural shark, this letter proposes a low-c

Cited by 0SourceScholar
2026

Score2Instruct: Scaling Up Video Quality-Centric Instructions via Automated Dimension Scoring

CVPR 2026

Classical video quality assessment (VQA) methods generate a numerical score to judge a video's perceived visual fidelity and clarity. Yet, a score fails to describe the video's complex quality dimensions (e.g., noise), restricting its applicability. Benefiting from the human-friendly linguistic outp

Cited by 0SourcecodeScholar
2026

ShiftLUT: Spatial Shift Enhanced Look-Up Tables for Efficient Image Restoration

CVPR 2026

Look-Up Table based methods have emerged as a promising direction for efficient image restoration tasks. Recent LUT-based methods focus on improving their performance by expanding the receptive field. However, they inevitably introduce extra computational and storage overhead, which hinders their de

Cited by 0SourcecodeScholar
2025

Accelerating Diffusion-based Super-Resolution with Dynamic Time-Spatial Sampling

IJCAI 2025

Diffusion models have gained attention for their success in modeling complex distributions, achieving impressive perceptual quality in SR tasks. However, existing diffusion-based SR methods often suffer from high computational costs, requiring numerous iterative steps for training and inference. Exi

Cited by 0SourcePDFScholar
2025

KVQ: Boosting Video Quality Assessment via Saliency-guided Local Perception

CVPR 2025poster

Video Quality Assessment (VQA), which intends to predict the perceptual quality of videos, has attracted increasing attention. Due to factors like motion blur or specific distortions, the quality of different regions in a video varies. Recognizing the region-wise local quality within a video is bene…

2025

Neural B-frame Video Compression with Bi-directional Reference Harmonization

NeurIPS 2025poster

Neural video compression (NVC) has made significant progress in recent years, while neural B-frame video compression (NBVC) remains underexplored compared to P-frame compression. NBVC can adopt bi-directional reference frames for better compression performance. However, NBVC's hierarchical coding ma…

Cited by 0SourcecodeScholar
2025

Plug-and-Play Tri-Branch Invertible Block for Image Rescaling

AAAI 2025technical

High-resolution (HR) images are commonly downscaled to low-resolution (LR) to reduce bandwidth, followed by upscaling to restore their original details. Recent advancements in image rescaling algorithms have employed invertible neural networks (INNs) to create a unified framework for downscaling and…

2025

Quantitative Analysis of Fish-Like Propulsion Mechanisms Based on the Discrete Vortex Method

RA-L 2025

Fish actively control their bodies to interact with the surrounding vortex flow field and generate thrust. Current research on the propulsion mechanisms of robotic fish primarily involves qualitative comparisons of vortex structures from simulations and experiments, lacking quantitative analysis. Th

Cited by 0SourceScholar
2025

Rethinking Personality Assessment from Human-Agent Dialogues: Fewer Rounds May Be Better Than More

EMNLP 2025

Personality assessment is essential for developing user-centered systems, playing a critical role across domains including hiring, education, and personalized system design. With the integration of conversational AI systems into daily life, automatically assessing human personality through natural l

2025

Scale Your Instructions: Enhance the Instruction-Following Fidelity of Unified Image Generation Model by Self-Adaptive Attention Scaling

ICCV 2025poster

Recent advancements in unified image generation models, such as OmniGen, have enabled the handling of diverse image generation and editing tasks within a single framework, accepting multimodal, interleaved texts and images in free form. This unified architecture eliminates the need for text encoders…

2025

Ultra Lowrate Image Compression with Semantic Residual Coding and Compression-aware Diffusion

ICML 2025poster

Existing multimodal large model-based image compression frameworks often rely on a fragmented integration of semantic retrieval, latent compression, and generative models, resulting in suboptimal performance in both reconstruction fidelity and coding efficiency. To address these challenges, we propo…

Cited by 0SourcePDFScholar
2025

Visual Autoregressive Modeling for Image Super-Resolution

ICML 2025poster

Image Super-Resolution (ISR) has seen significant progress with the introduction of remarkable generative models. However, challenges such as the trade-off issues between fidelity and realism, as well as computational complexity, have also posed limitations on their application. Building upon the tr…

2024

A New Dataset and Framework for Real-World Blurred Images Super-Resolution

ECCV 2024poster

"Recent Blind Image Super-Resolution (BSR) methods have shown proficiency in general images. However, we find that the efficacy of recent methods obviously diminishes when employed on image data with blur, while image data with intentional blur constitute a substantial proportion of general data. To…

2024

BAE-Net: a Low Complexity and High Fidelity Bandwidth-Adaptive Neural Network for Speech Super-Resolution

ICASSP 2024accepted

Speech bandwidth extension (BWE) has demonstrated promising performance in enhancing the perceptual speech quality in real communication systems. Most existing BWE researches primarily focus on fixed upsampling ratios, disregarding the fact that the effective bandwidth of captured audio may fluctuat…

Cited by 0SourceScholar
2024

CPGA: Coding Priors-Guided Aggregation Network for Compressed Video Quality Enhancement

CVPR 2024poster

Recently numerous approaches have achieved notable success in compressed video quality enhancement (VQE). However these methods usually ignore the utilization of valuable coding priors inherently embedded in compressed videos such as motion vectors and residual frames which carry abundant temporal a…

2024

KVQ: Kwai Video Quality Assessment for Short-form Videos

CVPR 2024poster

Short-form UGC video platforms like Kwai and TikTok have been an emerging and irreplaceable mainstream media form thriving on user-friendly engagement and kaleidoscope creation etc. However the advancing content generation modes e.g. special effects and sophisticated processing workflows e.g. de-art…

2024

OAPT: Offset-Aware Partition Transformer for Double JPEG Artifacts Removal

ECCV 2024poster

"Deep learning-based methods have shown remarkable performance in single JPEG artifacts removal task. However, existing methods tend to degrade on double JPEG images, which are prevalent in real-world scenarios. To address this issue, we propose Offset-Aware Partition Transformer for double JPEG art…

2024

PTM-VQA: Efficient Video Quality Assessment Leveraging Diverse PreTrained Models from the Wild

CVPR 2024poster

Video quality assessment (VQA) is a challenging problem due to the numerous factors that can affect the perceptual quality of a video e.g. content attractiveness distortion type motion pattern and level. However annotating the Mean opinion score (MOS) for videos is expensive and time-consuming which…

Cited by 5SourcePDFScholar
2024

XPSR: Cross-modal Priors for Diffusion-based Image Super-Resolution

ECCV 2024poster

"Diffusion-based methods, endowed with a formidable generative prior, have received increasing attention in Image Super-Resolution (ISR) recently. However, as low-resolution (LR) images often undergo severe degradation, it is challenging for ISR models to perceive the semantic and degradation inform…

2023

A Flow-Guided Non-Local Alignment Network for Video Compressive Sensing Reconstruction

ICASSP 2023accepted

Video compressive sensing (VCS) presents a promising encoder paradigm for efficient video signals acquisition at resource-limited applications. In order to recover complete and accurate signals at the decoder, powerful reconstruction algorithms are desired to exploit rich temporal redundancies withi…

Cited by 0SourceScholar
2023

A Performance Optimization Strategy Based on Improved NSGA-II for a Flexible Robotic Fish

ICRA 2023poster

The high speed and low energy cost are two conflicting objectives in the motion optimization of bio-inspired underwater robots, but playing a very important role. To this end, this paper proposes an optimization strategy for swimming speed and power cost using an improved NSGA-II for a flexible robo…

Cited by 2SourceScholar
2023

Design and Modeling of a Sperm-Inspired Helical Propulsion Robot

RA-L 2023

The development of biomimetics and the demand for higher propulsion efficiency lead to more research in helical propulsion robots. This letter presents a novel sperm-inspired robot that utilizes flexible tail as propulsion and analyzes the motion performance. Firstly, the robot's propulsion system i

Cited by 1SourceScholar
2022

An Indeterministic Vision-Based State Observer for Growing Magnetic Microrobot Motion Status Estimation

ICRA 2022poster

To date, untethered micro/nanorobots have attracted considerable attention in various aspects due to their unique potential for in-vivo applications such as the targeted therapy. One of the most promising types of micro/nanorobots is the class of ferromagnetic microrobots which can be efficiently ac…

Cited by 5SourceScholar
2022

Development and Stiffness Optimization for a Flexible-Tail Robotic Fish

RA-L 2022

The integral flexible tail has the potential advantage of lifelike undulating motion. However, due to the complex manufacturing process and difficult modification of structural parameters, its application in robotic fish encounters many challenges. Combining rigid structure and flexible material, th

Cited by 17SourceScholar
2022

Dynamic Modeling and Performance Analysis for a Wire-Driven Elastic Robotic Fish

RA-L 2022

The complex and continuous undulation of fishtail facilitates extraordinary underwater motion performance for natural fish. For the widely used Multi-Joint robotic fish, a lot of joints are used to simulate continuum fishtail, resulting in some challenges, e.g., the mechanism complexity, friction lo

Cited by 12SourceScholar
2022

Modeling and Characterization of Artificial Bacteria Flagella with Micro-structured Soft-magnetic Teeth

IROS 2022poster

Sub-structures such as micro-structured magnetic teeth fabricated with an artificial bacteria flagellum (ABF) are designed for achieving more motion modes, higher precision, and better controllability. To achieve these, a more precise model considering the non-circular cross-sectional features is se…

Cited by 2SourceScholar
2021

3D Periodic Magnetic Servoing System for Microrobot Actuation Using Decoupled Asynchronous Repetitive Control Approach

ICRA 2021poster

To date, untethered microrobots have been receiving tremendous attention for playing implacable roles of maneuverable tools in fields such as microfabrication and biomanipulation. Typical actuation of such untethered tiny robots is the magnetic field-based approaches, including gradient and rotation…

Cited by 3SourceScholar
2021

REST: Robust lEarned Shrinkage-Thresholding Network Taming Inverse Problems with Model Mismatch

ICASSP 2021accepted

We consider compressive sensing problems with model mismatch where one wishes to recover a sparse high-dimensional vector from low-dimensional observations subject to uncertainty in the measurement operator. In particular, we design a new robust deep neural network architecture by applying algorithm…

Cited by 0SourceScholar
2021

Scene Coordinate Regression Network With Global Context-Guided Spatial Feature Transformation for Visual Relocalization

RA-L 2021

Among visual relocalization from a single RGB image, the scene coordinate regression (SCoRe) based on convolutional neural network (CNN) becomes prevailing, however, it is insufficient to extract invariant features under different viewpoints due to fixed geometric structures of CNN. In this letter,

Cited by 19SourceScholar
2019

Learning Shape-Aware Embedding for Scene Text Detection

CVPR 2019poster

We address the problem of detecting scene text in arbitrary shapes, which is a challenging task due to the high variety and complexity of the scene. Specifically, we treat text detection as instance segmentation and propose a segmentation-based framework, which extracts each text instance as an inde…

Cited by 259PDFScholar
2017

Development of a power line inspection robot with hybrid operation modes

IROS 2017poster

In this paper, we design and build a power line inspection robot capable of hybrid operation modes. Specifically, the developed robot is able to land on the overhead ground wire (OGW) and to move as the climbing robot. When to negotiate obstacles, it can vertically take off the wire and fly over the…

Cited by 89SourceScholar
2017

High-Quality Correspondence and Segmentation Estimation for Dual-Lens Smart-Phone Portraits

ICCV 2017poster

Estimating correspondence between two images and extracting the foreground object are two challenges in computer vision. With dual-lens smart phones, such as iPhone 7Plus and Huawei P9, coming into the market, two images of slightly different views provide us new information to unify the two topics.…

Cited by 15PDFScholar
2017

Robotic Pick-And-Place of Multiple Embryos for Vitrification

RA-L 2017

Embryo vitrification is an essential cryopreservation technique in IVF (in vitro fertilization) clinics. Vitrification involves pick-and-place of an embryo in multiple types of cryoprotectant solutions for processing before placing the embryo on a vitrification straw for cryopreservation in liquid n

Cited by 37SourceScholar