← Search

Shizhuo Zhang

5 accepted papers

2026

GSUC-VLM: Geometrically-Guided Spatial Understanding Chain of Vision Language Model for Autonomous Driving

ICRA 2026poster

Robust spatial understanding is crucial for Visual Question Answering (VQA) in autonomous driving that aims to enhance decision-making, reduce positional risks, and ensure road safety by providing answers based on the perception, prediction, and planning of driving scenarios. Despite remarkable succ…

Cited by 0Scholar
2025

Standardization Status of MPEG Video-based Dynamic Mesh Coding (V-DMC)

ICASSP 2025accepted

3D dynamic meshes are extensively utilized to represent immersive 3D content. The need to transmit a vast quantity of these captured dynamic meshes over networks has created a significant demand for efficient dynamic mesh compression techniques. In response to this critical requirement, the Moving P…

Cited by 0SourceScholar
2023

GraphPrompt: Graph-Based Prompt Templates for Biomedical Synonym Prediction

AAAI 2023technical

In the expansion of biomedical dataset, the same category may be labeled with different terms, thus being tedious and onerous to curate these terms. Therefore, automatically mapping synonymous terms onto the ontologies is desirable, which we name as biomedical synonym prediction task. Unlike biomedi…

2023

Making Language Models Better Reasoners with Step-Aware Verifier

ACL 2023long

Few-shot learning is a challenging task that requires language models to generalize from limited examples. Large language models like GPT-3 and PaLM have made impressive progress in this area, but they still face difficulties in reasoning tasks such as GSM8K, a benchmark for arithmetic problems. To…

Cited by 185SourcePDFScholar
2022

Near-Optimal Task Selection for Meta-Learning with Mutual Information and Online Variational Bayesian Unlearning

AISTATS 2022poster

This paper addresses the problem of active task selection which involves selecting the most informative tasks for meta-learning. We propose a novel active task selection criterion based on the mutual information between latent task vectors. Unfortunately, such a criterion scales poorly in the number…

Cited by 11SourcePDFScholar