← Search

Chuang Wang

18 accepted papers

2026

MANSION: Multi-floor lANguage-to-3D Scene generatIOn for loNg-horizon tasks

CVPR 2026

Real-world robotic tasks are long-horizon and often span multiple floors, demanding rich spatial reasoning. However, existing embodied benchmarks are largely confined to single-floor in-house environments, failing to reflect the complexity of real-world tasks. We introduce MANSION, the first languag

Cited by 0SourceScholar
2026

RxnCaption: Reformulating Reaction Diagram Parsing as Visual Prompt Guided Captioning

CVPR 2026

Large-scale chemical reaction datasets are crucial for AI research in chemistry. However, existing chemical reaction data often exist as images within papers, making them not machine-readable and unusable for training machine learning models. In response to this challenge, we propose the RxnCaption

Cited by 0SourcecodeScholar
2026

VAnim: Rendering-Aware Sparse State Modeling for Structure-Preserving Vector Animation

ICML 2026poster

Scalable Vector Graphics (SVG) animation generation is pivotal for professional design due to their structural editability and resolution independence. However, this task remains challenging as it requires bridging discrete code representations with continuous visual dynamics. Existing optimization-…

Cited by 0SourceScholar
2025

Achieving binary weight and activation for LLMs using Post-Training Quantization

ACL 2025finding

Quantizing large language models (LLMs) to 1-bit precision significantly reduces computational costs, but existing quantization techniques suffer from noticeable performance degradation when using weight and activation precisions below 4 bits (W4A4). In this paper, we propose a post-training quantiz…

2025

Multimodal Task Attention Residual Reinforcement Learning: Advancing Robotic Assembly in Unstructured Environment

RA-L 2025

Robotic assembly in dynamic and unstructured environments poses challenges for recent methods, due to background noise and wide-ranging errors. Directly learning from environments relies on complex models and extensive training iterations to adapt. Representation selection approaches, which depend o

Cited by 5SourceScholar
2025

TrackGo: A Flexible and Efficient Method for Controllable Video Generation

AAAI 2025technical

Recent years have seen substantial progress in diffusion-based controllable video generation. However, achieving precise control in complex scenarios, including fine-grained object parts, sophisticated motion trajectories, and coherent background movement, remains a challenge. In this paper, we in…

Cited by 11SourcePDFScholar
2025

ViewCraft3D: High-fidelity and View-Consistent 3D Vector Graphics Synthesis

NeurIPS 2025poster

3D vector graphics play a crucial role in various applications including 3D shape retrieval, conceptual design, and virtual reality interactions due to their ability to capture essential structural information with minimal representation. While recent approaches have shown promise in generating 3D v…

Cited by 0SourceScholar
2024

Controlling Soft Robotic Arms Using Hybrid Modelling and Reinforcement Learning

RA-L 2024

Soft robotic arms exhibit high deformability and degrees of freedom, which brings challenges in modeling accuracy and susceptibility to gravitational effects, resulting in imprecise control. This study proposes a reinforcement learning control strategy based on a hybrid model for precise control of

Cited by 8SourceScholar
2024

Joint Pre-Encoding Representation and Structure Embedding for Efficient and Low-Resource Knowledge Graph Completion

EMNLP 2024main

Knowledge graph completion (KGC) aims to infer missing or incomplete parts in knowledge graph. The existing models are generally divided into structure-based and description-based models, among description-based models often require longer training and inference times as well as increased memory usa…

2024

SVGDreamer: Text Guided SVG Generation with Diffusion Model

CVPR 2024poster

Recently text-guided scalable vector graphics (SVGs) synthesis has shown promise in domains such as iconography and sketch. However existing text-to-SVG generation methods lack editability and struggle with visual quality and result diversity. To address these limitations we propose a novel text-gui…

2023

DiffSketcher: Text Guided Vector Sketch Synthesis through Latent Diffusion Models

NeurIPS 2023poster

Even though trained mainly on images, we discover that pretrained diffusion models show impressive power in guiding sketch synthesis. In this paper, we present DiffSketcher, an innovative algorithm that creates \textit{vectorized} free-hand sketches using natural language input. DiffSketcher is deve…

2022

On the Use of Bert for Automated Essay Scoring: Joint Learning of Multi-Scale Essay Representation

NAACL 2022long

In recent years, pre-trained models have become dominant in most natural language processing (NLP) tasks. However, in the area of Automated Essay Scoring (AES), pre-trained models such as BERT have not been properly used to outperform other deep learning models such as LSTM. In this paper, we introd…

2021

Prototype Augmentation and Self-Supervision for Incremental Learning

CVPR 2021poster

Despite the impressive performance in many individual tasks, deep neural networks suffer from catastrophic forgetting when learning new tasks incrementally. Recently, various incremental learning methods have been proposed, and some approaches achieved acceptable performance relying on stored data o…

Cited by 486PDFScholar
2021

Proxy Graph Matching with Proximal Matching Networks

AAAI 2021technical

Estimating feature point correspondence is a common technique in computer vision. A line of recent data-driven approaches utilizing the graph neural networks improved the matching accuracy by a large margin. However, these learning-based methods require a lot of labeled training data, which are expe…

Cited by 8SourcePDFScholar
2017

Neural network modeling for steering control of an autonomous vehicle

IROS 2017poster

Model-based control of dynamical systems typically requires accurate domain-specific knowledge and specifications of possibly proprietary system components. In the context of autonomous driving, steering actuator dynamics can be difficult to model due to an integrated proprietary power steering cont…

Cited by 63SourceScholar