← Search

Cheng Tang

12 accepted papers

2026

CloserToMe: A Unified Framework for Accurate and Transferable Latency Prediction Across Heterogeneous Devices

AAAI 2026technical

Hardware accelerators such as GPUs, NPUs, and FPGAs are essential to meeting AI’s computational demands. With the proliferation of heterogeneous devices across cloud and edge, various model optimization techniques adapt to diverse hardware characteristics through operator transformations and structu

Cited by 0SourcePDFScholar
2026

Escaping Low-Rank Traps: Interpretable Visual Concept Learning via Implicit Vector Quantization

ICLR 2026poster

Concept Bottleneck Models (CBMs) achieve interpretability by interposing a human-understandable concept layer between perception and label prediction. The foundation of CBMs lies in the many-to-many mapping that translates high-dimensional visual features to a set of discrete concepts. However, we…

Cited by 0SourceScholar
2026

EvoGraph-R1: Self-Evolving Multimodal Knowledge Hypergraphs for Agentic Retrieval

CVPR 2026

Retrieval-augmented generation (RAG) has emerged as a critical paradigm for grounding Multimodal Large Language Models (MLLMs) in external knowledge. Recent GraphRAG methods introduce structured entity-relation graphs to improve retrieval and reasoning. However, they remain limited by treating knowl

Cited by 0SourceScholar
2026

KUMA: A Novel Framework with Koopman Separation and Efficient Multilevel Extraction in Time Series Forecasting

ICML 2026poster

Time series forecasting plays a crucial role in a wide range of real-world applications and has become increasingly complex with the growth of multivariate dimensions and extended historical observations, leading to the prosperity of deep forecasting models. Previous models are hindered by three maj…

Cited by 0SourceScholar
2026

UniMedVL: Unifying Medical Multimodal Understanding and Generation through Observation-Knowledge-Analysis

ICML 2026poster

Medical diagnosis demands models that can process multimodal medical inputs, such as medical images and patient histories, and generate diverse outputs including textual reports and visual content, such as annotations or segmentation masks. Despite this need, existing medical AI models disrupt this …

Cited by 0SourceScholar
2025

Attention-Seeker: Dynamic Self-Attention Scoring for Unsupervised Keyphrase Extraction

COLING 2025main

This paper proposes Attention-Seeker, an unsupervised keyphrase extraction method that leverages self-attention maps from a Large Language Model to estimate the importance of candidate phrases. Our approach identifies specific components – such as layers, heads, and attention vectors – where the mod…

2025

Robust Offline Reinforcement Learning with Linearly Structured $f$-Divergence Regularization

ICML 2025poster

The Robust Regularized Markov Decision Process (RRMDP) is proposed to learn policies robust to dynamics shifts by adding regularization to the transition dynamics in the value function. Existing methods mostly use unstructured regularization, potentially leading to conservative policies under unreal…

Cited by 0SourcePDFScholar
2022

Evolutionary Neural Architecture Design of Liquid State Machine for Image Classification

ICASSP 2022accepted

As a recurrent spiking neural network, liquid state machine (LSM) has attracted more and more attention in neuromorphic computing due to its biological plausibility, computation power, and hardware implementation. However, the neural architecture of LSM, such as hidden neuron number, synaptic densit…

Cited by 0SourceScholar
2016

On Lloyd’s Algorithm: New Theoretical Insights for Clustering in Practice

AISTATS 2016poster

We provide new analyses of Lloyd’s algorithm (1982), commonly known as the k-means clustering algorithm. Kumar and Kannan (2010) showed that running k-SVD followed by a constant approximation k-means algorithm, and then Lloyd’s algorithm, will correctly cluster nearly all of the dataset with respect…

Cited by 32SourcePDFScholar