← Search

Zhong Wang

14 accepted papers

2026

Genome-Factory: A Library for Tuning, Deploying, and Interpreting Genomic Foundation Models

ICML 2026poster

We introduce Genome-Factory, the first integrated Python library for tuning, deploying, and interpreting genomic foundation models. Our core contribution is to simplify and unify the workflow for genomic model development: data collection, model tuning, inference, benchmarking, and interpretability.…

Cited by 0SourceScholar
2026

Global-Local Confidence Fusion for Hallucination Detection in Mathematical Reasoning Task

AAAI 2026technical

Large Reasoning Models (LRMs) achieve promising results on complex reasoning tasks but remain susceptible to hallucinations. Existing hallucination detection methods based on Large Language Models (LLMs) often focus solely on final answers, overlooking inconsistencies between the answer and reasonin

Cited by 0SourcePDFScholar
2026

SceneTransporter: Optimal Transport-Guided Compositional Latent Diffusion for Single-Image Structured 3D Scene Generation

ICLR 2026poster

We introduce SceneTransporter, an end-to-end framework for structured 3D scene generation from a single image. While existing methods generate part-level 3D objects, they often fail to organize these parts into distinct instances in open-world scenes. Through a debiased clustering probe, we reveal a…

Cited by 0SourcecodeScholar
2026

SmartSplat: Feature-Smart Gaussians for Scalable Compression of Ultra-High-Resolution Images

AAAI 2026technical

Recent advances in generative AI have accelerated the production of ultra-high-resolution visual content. However, traditional image formats face significant limitations in efficient compression and real-time decoding, which restricts their applicability on end-user devices. Inspired by 3D Gaussian

Cited by 0SourcePDFScholar
2025

Representing Sounds as Neural Amplitude Fields: A Benchmark of Coordinate-MLPs and a Fourier Kolmogorov-Arnold Framework

AAAI 2025technical

Although Coordinate-MLP-based implicit neural representations have excelled in representing radiance fields, 3D shapes, and images, their application to audio signals remains underexplored. To fill this gap, we investigate existing implicit neural representations, from which we extract 3 types of po…

2025

SafeConf: A Confidence-Calibrated Safety Self-Evaluation Method for Large Language Models

EMNLP 2025

Large language models (LLMs) have achieved groundbreaking progress in Natural Language Processing (NLP). Despite the numerous advantages of LLMs, they also pose significant safety risks. Self-evaluation mechanisms have gained increasing attention as a key safeguard to ensure safe and controllable co

Cited by 0SourcePDFScholar
2025

TermDiffuSum: A Term-guided Diffusion Model for Extractive Summarization of Legal Documents

COLING 2025main

Extractive summarization for legal documents aims to automatically extract key sentences from legal texts to form concise summaries. Recent studies have explored diffusion models for extractive summarization task, showcasing their remarkable capabilities. Despite these advancements, these models oft…

2025

Towards Autonomous Indoor Parking: A Globally Consistent Semantic SLAM System and A Semantic Localization Subsystem

IROS 2025

We propose a globally consistent semantic SLAM system (GCSLAM) and a semantic-fusion localization subsystem (SF-Loc), which achieves accurate semantic mapping and robust localization in complex parking lots. Visual cameras (front-view and surround-view), IMU, and wheel encoder form the input sensor

Cited by 3SourceScholar
2024

Beyond the Snowfall: Enhancing Snowy Day Object Detection Through Progressive Restoration and Multi-Feature Fusion

ICASSP 2024accepted

In the field of computer vision, object detection is a prominent and challenging task. Despite the favorable performance of deep learning-based object detection techniques on clear images, it fails in inclement weather conditions like snow because of image degradation. Recent efforts have explored u…

Cited by 0SourceScholar
2024

RVDNet: A Two-Stage Network for Real-World Video Desnowing with Domain Adaptation

ICASSP 2024accepted

Video snow removal is an important task in computer vision, as the snowflakes in videos reduce visibility and negatively affect the performance of outdoor visual systems. However, due to the complexity of real snowy scenarios, it is difficult to apply existing supervised learning-based methods to pr…

Cited by 0SourceScholar
2023

Dynamic Object Classification of Low-Resolution Point Clouds: An LSTM-Based Ensemble Learning Approach

RA-L 2023

In unmanned vehicle perception, dynamic object classification is applied to classify objects accurately and timely, providing decision-making for obstacle avoidance and planning. Low-resolution LiDAR is one of the most important sensors for this task. Unfortunately, the existing approaches perform u

Cited by 2SourceScholar
2021

Towards Robust Autonomous Coverage Navigation for Carlike Robots

RA-L 2021

Thanks to their high carrying capacity and strong maneuverability, carlike robots which move with non-holonomic constraints, are frequently utilized in numerous coverage operation fields. In such fields, the robots need to complete the coverage task via autonomous path planning and tracking, which i

Cited by 5SourcecodeScholar