← Search

Zhenni Bi

3 accepted papers

2025

Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts

AAAI 2025technical

Multimodal vision language models (VLMs) have made significant progress with the support of continuously increasing model sizes and data volumes. Running VLMs on edge devices has become a challenge for their widespread application. There are several efficient VLM efforts, but they often sacrifice li…

2025

Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning

ICML 2025poster

Large Language Models (LLMs) have demonstrated remarkable abilities across various language tasks, but solving complex reasoning problems remains a significant challenge. While existing methods, such as Chain-of-Thought (CoT) and Tree-of-Thought (ToT), enhance reasoning by decomposing problems or st…

2024

An Empirical Study of Scaling Law for Scene Text Recognition

CVPR 2024poster

The laws of model size data volume computation and model performance have been extensively studied in the field of Natural Language Processing (NLP). However the scaling laws in Scene Text Recognition (STR) have not yet been investigated. To address this we conducted comprehensive studies that invol…