← Search

Jundong Zhou

6 accepted papers

2026

ConceptMoE: Adaptive Token-to-Concept Compression for Implicit Compute Allocation

ICML 2026poster

Large language models allocate uniform computation across all tokens, ignoring that some sequences are trivially predictable while others require deep reasoning. We introduce ConceptMoE, which dynamically merges semantically similar tokens into concepts through learnable chunking at target compressi…

Cited by 0SourceScholar
2026

ImpQuant: Fine-Grained Importance-Aware Quantization for Large Vision-Language Models

ICML 2026poster

Large Vision–Language Models (LVLMs) have demonstrated remarkable capabilities across diverse multimodal tasks, yet their high inference costs necessitate low-bit deployment. Existing post-training quantization (PTQ) pipelines primarily adopt methodologies from text-only LLMs by treating multimodal …

Cited by 0SourceScholar
2026

SPUR: Scale-Partitioned Uncertainty Rectification for Robust UAV-on-UAV Interception

ICML 2026poster

Robust aerial target detection for autonomous UAV-on-UAV pursuit is severely hindered by continuous scale drift, long-tailed scale imbalance, and flight-induced visual noise, rendering standard empirical risk minimization strategies poorly aligned with real-world deployment. To address these challen…

Cited by 0SourceScholar
2024

PNAS-MOT: Multi-Modal Object Tracking With Pareto Neural Architecture Search

RA-L 2024

Multiple object tracking is a critical task in autonomous driving. Existing works primarily focus on the heuristic design of neural networks to obtain high accuracy. As tracking accuracy improves, however, neural networks become increasingly complex, posing challenges for their practical application

Cited by 21SourcecodeScholar
2023

Box-Level Active Detection

CVPR 2023highlight

Active learning selects informative samples for annotation within budget, which has proven efficient recently on object detection. However, the widely used active detection benchmarks conduct image-level evaluation, which is unrealistic in human workload estimation and biased towards crowded images.…

2022

ReMoNet: Recurrent Multi-Output Network for Efficient Video Denoising

AAAI 2022technical

While deep neural network-based video denoising methods have achieved promising results, it is still hard to deploy them on mobile devices due to their high computational cost and memory demands. This paper aims to develop a lightweight deep video denoising method that is friendly to resource-constr…

Cited by 13SourcePDFScholar