← Search

Nan Sun

8 accepted papers

2026

AMS-IO-Bench and AMS-IO-Agent: Benchmarking and Structured Reasoning for Analog and Mixed-Signal Integrated Circuit Input/Output Design

AAAI 2026technical

In this paper, we propose AMS-IO-Agent, a domain-specialized LLM-based agent for structure-aware input/output (I/O) subsystem generation in analog and mixed-signal (AMS) integrated circuits (ICs). The central contribution of this work is a framework that connects natural language design intent with

Cited by 0SourcePDFScholar
2026

CollabVLA: Self-Reflective Vision-Language-Action Model Dreaming Together with Human

ICRA 2026poster

In this work, we present CollabVLA, a self-reflective vision-language-action framework that transforms a standard visuomotor policy into a collaborative assistant. CollabVLA tackles key limitations of prior VLAs, including domain overfitting, non-interpretable reasoning, and the high latency of auxi…

2026

GlyphShield: Document Watermarking for the Physical World via Vector Typeface Synthesis

AAAI 2026technical

Document protection has become a critical issue for preventing unauthorized copying, distribution, and tampering. Document encryption is a proven solution, but it is not resistant to attacks from the physical world such as screenshots, printing and photographing. A common document protection techniq

Cited by 0SourcePDFScholar
2025

AssistantX: An LLM-Powered Proactive Assistant in Collaborative Human-Populated Environments

IROS 2025

Current service robots suffer from limited natural language communication abilities, heavy reliance on predefined commands, ongoing human intervention, and, most notably, a lack of proactive collaboration awareness in human-populated environments. This results in narrow applicability and low utility

Cited by 6SourcecodeScholar
2025

END^2: Robust Dual-Decoder Watermarking Framework Against Non-Differentiable Distortions

AAAI 2025technical

DNN-based watermarking methods have rapidly advanced, with the ``Encoder-Noise Layer-Decoder'' (END) framework being the most widely used. To ensure end-to-end training, the noise layer in the framework must be differentiable. However, real-world distortions are often non-differentiable, leading to…

Cited by 0SourcePDFScholar
2025

Ultra-high Resolution Watermarking Framework Resistant to Extreme Cropping and Scaling

NeurIPS 2025poster

Recent developments in DNN-based image watermarking techniques have achieved impressive results in protecting digital content. However, most existing methods are constrained to low-resolution images as they need to encode the entire image, leading to prohibitive memory and computational costs when a…

Cited by 0SourceScholar
2024

Enhancing Multimodal Knowledge Graph Representation Learning through Triple Contrastive Learning

IJCAI 2024poster

Multimodal knowledge graphs incorporate multimodal information rather than pure symbols, which significantly enhance the representation of knowledge graphs and their capacity to understand the world. Despite these advancements, existing multimodal fusion techniques still face significant challenges…

Cited by 2SourcePDFScholar