← Search

Siwei Zhang

15 accepted papers

2026

TG-RAG: A Retrieval-Augmented Framework for Reasoning Guidance in Specialized Domains

ICML 2026oral

Enhancing Large Reasoning Models (LRMs) for specialized domains remains a critical challenge. While recent industrial frameworks attempt to encapsulate Standard Operating Procedures into modular "skills" for dynamic retrieval, utilizing them via context engineering often proves insufficient for comp…

Cited by 0SourceScholar
2025

LLM-GAN: Constructing Generative Adversarial Network Through Large Language Models for Explainable Fake News Detection

ICASSP 2025accepted

Explainable fake news detection predicts the authenticity of news items with annotated explanations. Today, Large Language Models (LLMs) are known for their powerful natural language understanding and explanation generation abilities. However, using LLMs for explainable fake news detection remains t…

Cited by 0SourceScholar
2025

Rethinking Time Encoding via Learnable Transformation Functions

ICML 2025poster

Effectively modeling time information and incorporating it into applications or models involving chronologically occurring events is crucial. Real-world scenarios often involve diverse and complex time patterns, which pose significant challenges for time encoding methods. While previous methods focu…

2025

Unifying Text Semantics and Graph Structures for Temporal Text-attributed Graphs with Large Language Models

NeurIPS 2025poster

Temporal graph neural networks (TGNNs) have shown remarkable performance in temporal graph modeling. However, real-world temporal graphs often possess rich textual information, giving rise to temporal text-attributed graphs (TTAGs). Such combination of dynamic text semantics and evolving graph struc…

Cited by 0SourceScholar
2025

VolumetricSMPL: A Neural Volumetric Body Model for Efficient Interactions, Contacts, and Collisions

ICCV 2025poster

Parametric human body models play a crucial role in computer graphics and vision, enabling applications ranging from human motion analysis to understanding human-environment interactions. Traditionally, these models use surface meshes, which pose challenges in efficiently handling interactions with…

2024

A Joint Look on Lunar Satellite and Cooperative Surface PNT

ICASSP 2024accepted

A large number of missions to the Moon is planned in the coming years by both, public and private sectors. A high demand for a Lunar position, navigation and timing (PNT) system exists to aid landing, to support autonomous robotic exploration etc. ESA has proposed a satellite based system for commun…

Cited by 0SourceScholar
2024

EgoGen: An Egocentric Synthetic Data Generator

CVPR 2024poster

Understanding the world in first-person view is fundamental in Augmented Reality (AR). This immersive perspective brings dramatic visual changes and unique challenges compared to third-person views. Synthetic data has empowered third-person-view vision models but its application to embodied egocentr…

Cited by 17SourcePDFScholar
2024

RoHM: Robust Human Motion Reconstruction via Diffusion

CVPR 2024poster

We propose RoHM an approach for robust 3D human motion reconstruction from monocular RGB(-D) videos in the presence of noise and occlusions. Most previous approaches either train neural networks to directly regress motion in 3D or learn data-driven motion priors and combine them with optimization at…

2023

Autonomous Navigation of a Robotic Swarm in Space Exploration Missions

ICASSP 2023accepted

In recent years, the paradigm of navigation has shifted from pinpointing the location of a single agent to continuously estimating the full kinematic state of networked autonomous agents. In this paper, we propose a kinematics-aware information seeking algorithm for swarm navigation. The algorithm t…

Cited by 0SourceScholar
2023

Probabilistic Human Mesh Recovery in 3D Scenes from Egocentric Views

ICCV 2023oral

Automatic perception of human behaviors during social interactions is crucial for AR/VR applications, and an essential component is estimation of plausible 3D human pose and shape of our social partners from the egocentric view. One of the biggest challenges of this task is severe body truncation du…

Cited by 30PDFcodeScholar
2022

EgoBody: Human Body Shape and Motion of Interacting People from Head-Mounted Devices

ECCV 2022poster

"Understanding social interactions from egocentric views is crucial for many applications, ranging from assistive robotics to AR/VR. Key to reasoning about interactions is to understand the body pose and motion of the interaction partner from the egocentric view. However, research in this area is se…

2022

SAGA: Stochastic Whole-Body Grasping with Contact

ECCV 2022poster

"The synthesis of human grasping has numerous applications including AR/VR, video games and robotics. While methods have been proposed to generate realistic hand-object interaction for object grasping and manipulation, these typically only consider interacting hand alone. Our goal is to synthesize w…

2021

Learning Motion Priors for 4D Human Body Capture in 3D Scenes

ICCV 2021poster

Recovering high-quality 3D human motion in complex scenes from monocular videos is important for many applications, ranging from AR/VR to robotics. However, capturing realistic human-scene interactions, while dealing with occlusions and partial views, is challenging; current approaches are still far…

Cited by 116PDFcodeScholar
2020

The ARCHES Space-Analogue Demonstration Mission: Towards Heterogeneous Teams of Autonomous Robots for Collaborative Scientific Sampling in Planetary Exploration

RA-L 2020

Teams of mobile robots will play a crucial role in future missions to explore the surfaces of extraterrestrial bodies. Setting up infrastructure and taking scientific samples are expensive tasks when operating in distant, challenging, and unknown environments. In contrast to current single-robot spa

Cited by 92SourceScholar