← Search

Jianquan Li

14 accepted papers

2026

PTQ4ARVG: Post-Training Quantization for AutoRegressive Visual Generation Models

ICLR 2026poster

AutoRegressive Visual Generation (ARVG) models retain an architecture compatible with language models, while achieving performance comparable to diffusion-based models. Quantization is commonly employed in neural networks to reduce model size and computational latency. However, applying quantization…

Cited by 0SourcecodeScholar
2025

Huatuo-26M, a Large-scale Chinese Medical QA Dataset

NAACL 2025findings

Large Language Models infuse newfound vigor into the advancement of the medical domain, yet the scarcity of data poses a significant bottleneck hindering community progress. In this paper, we release the largest ever medical Question Answering (QA) dataset with 26 Million QA pairs named Huatuo-26M.…

2025

K-Sort Arena: Efficient and Reliable Benchmarking for Generative Models via K-wise Human Preferences

CVPR 2025poster

The rapid advancement of visual generative models necessitates efficient and reliable evaluation methods. Arena platform, which gathers user votes on model comparisons, can rank models with human preferences. However, traditional Arena methods, while established, require an excessive number of compa…

Cited by 4SourcePDFScholar
2025

MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria

NAACL 2025long

Multimodal large language models (MLLMs) have broadened the scope of AI applications. Existing automatic evaluation methodologies for MLLMs are mainly limited in evaluating objective queries without considering real-world user experiences, inadequately addressing the nuances of creative and associat…

2024

AceGPT, Localizing Large Language Models in Arabic

NAACL 2024long

This paper is devoted to the development of a localized Large Language Model (LLM) specifically for Arabic, a language imbued with unique cultural characteristics inadequately addressed by current mainstream models. Significant concerns emerge when addressing cultural sensitivity and local values. T…

2024

CMB: A Comprehensive Medical Benchmark in Chinese

NAACL 2024long

Large Language Models (LLMs) provide a possibility to make a great breakthrough in medicine. The establishment of a standardized medical benchmark becomes a fundamental cornerstone to measure progression. However, medical environments in different regions have their local characteristics, e.g., the…

2024

Incorporating Lexical and Syntactic Knowledge for Unsupervised Cross-Lingual Transfer

COLING 2024main

Unsupervised cross-lingual transfer involves transferring knowledge between languages without explicit supervision. Although numerous studies have been conducted to improve performance in such tasks by focusing on cross-lingual knowledge, particularly lexical and syntactic knowledge, current approac…

2024

Locomotion Control on Human-Centaur System With Spherical Joint Interaction

RA-L 2024

This paper presents a locomotion controller for a novel human-augmented legged robot, the Centaur robot, which is primarily developed to extend the human's ability to carry load. For such a human-robot walking system, there are requirements for the robot to maintain a balanced posture, provide a pro

Cited by 4SourceScholar
2023

Can Language Models Make Fun? A Case Study in Chinese Comical Crosstalk

ACL 2023long

Language is the principal tool for human communication, in which humor is one of the most attractive parts. Producing natural language like humans using computers, a.k.a, Natural Language Generation (NLG), has been widely used for dialogue systems, chatbots, machine translation, as well as computer-…

2023

Effective Open Intent Classification with K-center Contrastive Learning and Adjustable Decision Boundary

AAAI 2023technical

Open intent classification, which aims to correctly classify the known intents into their corresponding classes while identifying the new unknown (open) intents, is an essential but challenging task in dialogue systems. In this paper, we introduce novel K-center contrastive learning and adjustable d…

2023

HuatuoGPT, Towards Taming Language Model to Be a Doctor

EMNLP 2023long findings

In this paper, we present HuatuoGPT, a Large Language Model (LLM) for medical consultation. The core recipe of HuatuoGPT is to leverage both distilled data from **ChatGPT** and real-world data from **doctors** in the supervised fine-tuning stage. This is not only because purely using **ChatGPT**-di…

Cited by 0SourcecodeScholar
2022

A Centaur System for Assisting Human Walking with Load Carriage

IROS 2022poster

Walking with load is a common task in daily life and disaster rescue. Long-term load carriage may cause irreversible damage to the human body. Although remarkable progress has been made in the field of wearable robots, it is still far from avoiding interference to human legs, which will lead to ener…

Cited by 8SourceScholar
2020

Natural Scene Facial Expression Recognition with Dimension Reduction Network

ICRA 2020poster

As an external manifestation of human emotions, expression recognition plays an important role in human-computer interaction. Although existing expression recognition methods performs perfectly on constrained frontal faces, there are still many challenges in expression recognition in natural scenes…

Cited by 1SourceScholar
2017

12,000-fps Multi-object detection using HOG descriptor and SVM classifier

IROS 2017poster

This paper describes a high-frame-rate (HFR) vision system that can detect multiple objects in an image of 512 × 512 pixels at 12,000 frames per seconds (fps). An optimized algorithm is proposed based on conventional Histograms of Oriented Gradient (HOG) descriptor and Support Vector Machine (SVM) c…

Cited by 13SourceScholar