← Search

Xiaofeng Shi

3 accepted papers

2026

MechVQA: Benchmarking and Enhancing Multimodal LLMs on Comprehensive Mechanical Drawing Understanding

ICML 2026poster

Multimodal Large Language Models (MLLMs) have demonstrated significant achievements in general visual question answering (VQA) tasks. However, they remain brittle on mechanical engineering drawings, where high annotation density and weak domain knowledge, compounded by unreliable spatial relation re…

Cited by 0SourceScholar
2025

CareBot: A Pioneering Full-Process Open-Source Medical Language Model

AAAI 2025technical

Recently, both closed-source and open-source LLMs have made significant strides, outperforming humans in various general domains. However, their performance in specific professional domains such as medicine, especially within the open-source community, remains suboptimal due to the complexity of med…

2022

VarMAE: Pre-training of Variational Masked Autoencoder for Domain-adaptive Language Understanding

EMNLP 2022finding

Pre-trained language models have been widely applied to standard benchmarks. Due to the flexibility of natural language, the available resources in a certain domain can be restricted to support obtaining precise representation. To address this issue, we propose a novel Transformer-based language mod…

Cited by 9SourcePDFScholar