← Search

Run Shao

2 accepted papers

2026

Asking like Socrates: Socrates helps VLMs understand remote sensing images

CVPR 2026

Recent multimodal reasoning models, inspired by DeepSeek-R1, have significantly advanced vision-language systems. However, in remote sensing (RS) tasks, we observe widespread pseudo reasoning: models narrate the process of reasoning rather than genuinely reason toward the correct answer based on vis

Cited by 0SourcecodeScholar
2025

Select to Know: An Internal-External Knowledge Self-Selection Framework for Domain-Specific Question Answering

EMNLP 2025

Large Language Models (LLMs) perform well in general QA but often struggle in domain-specific scenarios. Retrieval-Augmented Generation (RAG) introduces external knowledge but suffers from hallucinations and latency due to noisy retrievals. Continued pretraining internalizes domain knowledge but is

Cited by 0SourcePDFScholar