← Search

Jiahao Wu

22 accepted papers

2026

ClipGStream: Clip-Stream Gaussian Splatting for Any Length and Any Motion Multi-View Dynamic Scene Reconstruction

CVPR 2026

Dynamic 3D scene reconstruction is essential for immersive media such as VR, MR, and XR, yet remains challenging for long multi-view sequences with large-scale motion. Existing dynamic Gaussian approaches are either Frame-Stream, offering scalability but poor temporal stability, or Clip, achieving l

Cited by 0SourceScholar
2026

Intrinsic Geometry-Appearance Consistency Optimization for Sparse-View Gaussian Splatting

CVPR 2026

3D Gaussian Splatting (3DGS) represents scenes through primitives with coupled intrinsic properties: geometric attributes (position, covariance, opacity) and appearance attributes (view-dependent color). Faithful reconstruction requires intrinsic geometry-appearance consistency, where geometry accur

Cited by 0SourceScholar
2026

Neural QAOA$^2$: Differentiable Joint Graph Partitioning and Parameter Initialization for Quantum Combinatorial Optimization

ICML 2026poster

The quantum approximate optimization algorithm (QAOA) holds promise for combinatorial optimization but is constrained by limited qubits. While divide-and-conquer frameworks like QAOA$^2$ address scalability by partitioning graphs into subgraphs, existing methods suffer from two fundamental limitatio…

Cited by 0SourceScholar
2026

Pano-GS: Perception-Aware Gaussian Optimization with Gradient Consistency and Multi-Criteria Densification for High-Quality Rendering

AAAI 2026technical

Reconstructing 3D scenes from multi-view image sequences remains a significant challenge in practical applications. While recent advances in 3D Gaussian Splatting have enabled high-quality rendering, existing methods rely heavily on pixel-level L1 loss, which misaligns with human perception, leading

Cited by 0SourcePDFScholar
2026

Principled Synthetic Data Enables the First Scaling Laws for LLMs in Recommendation

ICML 2026poster

Large Language Models (LLMs) represent a promising frontier for recommender systems, yet their development has been impeded by the absence of predictable scaling laws, which are crucial for guiding research and optimizing resource allocation. We hypothesize that this may be attributed to the inheren…

Cited by 0SourceScholar
2025

ADC-GS: Pose-Free 3D Gaussian Splatting with Adaptive Depth Consistency

ICASSP 2025accepted

Recently proposed 3D Gaussian Splatting (3DGS) has achieved state-of-the-art results in the fields of novel view synthesis, but it heavily relies on pre-computed camera poses. Although recent methods mitigate by leveraging explicit representations achieve novel view synthesis without requiring camer…

Cited by 0SourceScholar
2025

DSE: A Denoising State Estimator for RL-Based Bipedal Robot Locomotion

RA-L 2025

Recent advancements in legged robot locomotion and reinforcement learning have demonstrated significant potential for the development of bipedal robot. But the state estimation accuracy and bipedal robot locomotion robustness of Reinforcement Learning based (RL-based) controller is significantly inf

Cited by 0SourceScholar
2025

HDA-GS: Hierarchical Density-Controlled for Anisotropic 3D Gaussian Splatting

ICASSP 2025accepted

Recently, 3D Gaussian Splatting (3D-GS) has demonstrated impressive results in novel view synthesis, achieving outstanding rendering quality and speed. However, 3D-GS heavily relies on the quality of the initial point cloud, and its original Adaptive Density Control (ADC) module has difficulty handl…

Cited by 0SourceScholar
2025

Instant Gaussian Stream: Fast and Generalizable Streaming of Dynamic Scene Reconstruction via Gaussian Splatting

CVPR 2025highlight

Building Free-Viewpoint Videos in a streaming manner offers the advantage of rapid responsiveness compared to offline training methods, greatly enhancing user experience. However, current streaming approaches face challenges of high per-frame reconstruction time (10s+) and error accumulation, limiti…

2025

IntelliCockpitBench: A Comprehensive Benchmark to Evaluate VLMs for Intelligent Cockpit

ACL 2025finding

The integration of sophisticated Vision-Language Models (VLMs) in vehicular systems is revolutionizing vehicle interaction and safety, performing tasks such as Visual Question Answering (VQA). However, a critical gap persists due to the lack of a comprehensive benchmark for multimodal VQA models in…

2025

LocalDyGS: Multi-view Global Dynamic Scene Modeling via Adaptive Local Implicit Feature Decoupling

ICCV 2025poster

Due to the complex and highly dynamic motions in the real world, synthesizing dynamic videos from multi-view inputs for arbitrary viewpoints is challenging. Previous works based on neural radiance field or 3D Gaussian splatting are limited to modeling fine-scale motion, greatly restricting their app…

Cited by 0SourcePDFScholar
2025

Multi-View Image Enhancement Inconsistency Decoupling Guided 3D Gaussian Splatting

ICASSP 2025accepted

3D Gaussian Splatting (3DGS) has recently made breakthrough progress in radiance field reconstruction but struggles with multi-view inconsistency. Modern cameras often apply tailored enhancements to each view when capturing multi-view images. While this improves individual image quality, it inevitab…

Cited by 0SourceScholar
2025

OneForecast: A Universal Framework for Global and Regional Weather Forecasting

ICML 2025poster

Accurate weather forecasts are important for disaster prevention, agricultural planning, etc. Traditional numerical weather prediction (NWP) methods offer physically interpretable high-accuracy predictions but are computationally expensive and fail to fully leverage rapidly growing historical data.…

2025

Rising from Ashes: Generalized Federated Learning via Dynamic Parameter Reset

NeurIPS 2025poster

Although Federated Learning (FL) is promising in privacy-preserving collaborative model training, it faces low inference performance due to heterogeneous data among clients. Due to heterogeneous data in each client, FL training easily learns the specific overfitting features. Existing FL methods ad…

Cited by 0SourceScholar
2025

SAP: Exact Sorting in Splatting via Screen-Aligned Primitives

NeurIPS 2025poster

Recently, 3D Gaussian Splatting (3DGS) has achieved state-of-the-art rendering results. However, its efficiency relies on simplifications that disregard the thickness of Gaussian primitives and their overlapping interactions. These simplifications can lead to popping artifacts due to inaccurate sort…

Cited by 0SourceScholar
2025

Safe Delta: Consistently Preserving Safety when Fine-Tuning LLMs on Diverse Datasets

ICML 2025poster

Large language models (LLMs) have shown great potential as general-purpose AI assistants across various domains. To fully leverage this potential in specific applications, many companies provide fine-tuning API services, enabling users to upload their own data for LLM customization. However, fine-tu…

2025

Swift4D: Adaptive divide-and-conquer Gaussian Splatting for compact and efficient reconstruction of dynamic scene

ICLR 2025poster

Novel view synthesis has long been a practical but challenging task, although the introduction of numerous methods to solve this problem, even combining advanced representations like 3D Gaussian Splatting, they still struggle to recover high-quality results and often consume too much storage memory…

Cited by 1SourcePDFScholar
2024

Adapt2Reward: Adapting Video-Language Models to Generalizable Robotic Rewards via Failure Prompts

ECCV 2024poster

"For a general-purpose robot to operate in reality, executing a broad range of instructions across various environments is imperative. Central to the reinforcement learning and planning for such robotic agents is a generalizable reward function. Recent advances in vision-language models, such as CLI…

Cited by 2SourcePDFScholar
2021

Spotlight-Based 3D Instrument Guidance for Autonomous Task in Robot-Assisted Retinal Surgery

RA-L 2021

Retinal surgery is known to be a complicated and challenging task for an ophthalmologist even for retina specialists. Image guided robot-assisted intervention is among the novel and promising solutions that may enhance human capabilities during microsurgery. In this paper, a novel method is proposed

Cited by 21SourceScholar
2020

An Optimized Tilt Mechanism for a New Steady-Hand Eye Robot

IROS 2020poster

Robot-assisted vitreoretinal surgery can filter surgeons' hand tremors and provide safe, accurate tool manipulation. In this paper, we report the design, optimization, and evaluation of a novel tilt mechanism for a new Steady-Hand Eye Robot (SHER). The new tilt mechanism features a four-bar linkage…

Cited by 12SourceScholar
2020

Decidable Variable-Rate Dataflow for Heterogeneous Signal Processing Systems

ICASSP 2020accepted

Dynamic dataflow models of computation have become widely used through their adoption to popular programming frameworks such as TensorFlow and GNU Radio. Although dynamic dataflow models offer more programming freedom, they lack analyzability compared to their static counterparts (such as synchronou…

Cited by 0SourceScholar