← Search

Zhihan Zhou

15 accepted papers

2026

AncientBench: Towards Comprehensive Evaluation on Excavated and Transmitted Chinese Corpora

AAAI 2026technical

Comprehension of ancient texts plays an important role in archaeology and understanding of Chinese history and civilization. The rapid development of large language models needs benchmarks that can evaluate their comprehension of ancient characters. Existing Chinese benchmarks are mostly targeted at

Cited by 0SourcePDFScholar
2026

Genome-Factory: A Library for Tuning, Deploying, and Interpreting Genomic Foundation Models

ICML 2026poster

We introduce Genome-Factory, the first integrated Python library for tuning, deploying, and interpreting genomic foundation models. Our core contribution is to simplify and unify the workflow for genomic model development: data collection, model tuning, inference, benchmarking, and interpretability.…

Cited by 0SourceScholar
2026

InteChar: A Unified Oracle Bone Character List for Ancient Chinese Language Modeling

AAAI 2026technical

Constructing historical language models (LMs) plays a crucial role in aiding archaeological provenance studies and understanding ancient cultures. However, existing resources present major challenges for training effective LMs on historical texts. First, the scarcity of historical language samples r

Cited by 0SourcePDFScholar
2026

Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language Models

ICLR 2026poster

Large Language Models (LLMs) achieve impressive performance across many tasks but remain prone to hallucination, especially in long-form generation where redundant retrieved contexts and lengthy reasoning chains amplify factual errors. Recent studies highlight a critical phenomenon: the closer key i…

Cited by 0SourceScholar
2025

Fast and Low-Cost Genomic Foundation Models via Outlier Removal

ICML 2025poster

To address the challenge of scarce computational resources in genomic modeling, we introduce GERM, a genomic foundation model optimized for accessibility and adaptability. GERM improves upon models like DNABERT-2 by eliminating outliers that hinder low-rank adaptation and post-training quantization,…

2025

Learning to Instruct for Visual Instruction Tuning

NeurIPS 2025poster

We propose L2T, an advancement of visual instruction tuning (VIT). While VIT equips Multimodal LLMs (MLLMs) with promising multimodal capabilities, the current design choices for VIT often result in overfitting and shortcut learning, potentially degrading performance. This gap arises from an overemp…

Cited by 7SourcecodeScholar
2024

DNABERT-2: Efficient Foundation Model and Benchmark For Multi-Species Genomes

ICLR 2024poster

Decoding the linguistic intricacies of the genome is a crucial problem in biology, and pre-trained foundational models such as DNABERT and Nucleotide Transformer have made significant strides in this area. Existing works have largely hinged on k-mer, fixed-length permutations of A, T, C, and G, as t…

2024

EmoPrompt-ECPE: Emotion Knowledge-aware Prompt-tuning for Emotion-Cause Pair Extraction

COLING 2024main

Emotion-cause pair extraction (ECPE) main focus is on extracting all potential emotion clauses and corresponding cause clauses from unannotated documents. Existing methods achieve promising results with the help of fine-tuning and prompt paradigms, but they present three downsides. First, most appro…

2024

On Harmonizing Implicit Subpopulations

ICLR 2024poster

Machine learning algorithms learned from data with skewed distributions usually suffer from poor generalization, especially when minority classes matter as much as, or even more than majority ones. This is more challenging on class-balanced data that has some hidden imbalanced subpopulations, since…

Cited by 8SourcePDFScholar
2024

POP-CEE: Position-oriented Prompt-tuning Model for Causal Emotion Entailment

ACL 2024findings

The objective of the Causal Emotion Entailment (CEE) task is to identify the causes of the target emotional utterances in a given conversation. Most existing studies have focused on a fine-tuning paradigm based on a pretrained model, e.g., the BERT model. However, there are gaps between the pretrain…

2023

Combating Representation Learning Disparity with Geometric Harmonization

NeurIPS 2023spotlight

Self-supervised learning (SSL) as an effective paradigm of representation learning has achieved tremendous success on various curated datasets in diverse scenarios. Nevertheless, when facing the long-tailed distribution in real-world applications, it is still hard for existing methods to capture tra…

2023

Long-Tailed Partial Label Learning via Dynamic Rebalancing

ICLR 2023poster

Real-world data usually couples the label ambiguity and heavy imbalance, challenging the algorithmic robustness of partial label learning (PLL) and long-tailed learning (LT). The straightforward combination of LT and PLL, i.e., LT-PLL, suffers from a fundamental dilemma: LT methods build upon a give…

2022

Contrastive Learning with Boosted Memorization

ICML 2022spotlight

Self-supervised learning has achieved a great success in the representation learning of visual and textual data. However, the current methods are mainly validated on the well-curated datasets, which do not exhibit the real-world long-tailed distribution. Recent attempts to consider self-supervised l…

2022

Learning Dialogue Representations from Consecutive Utterances

NAACL 2022long

Learning high-quality dialogue representations is essential for solving a variety of dialogue-oriented tasks, especially considering that dialogue systems often suffer from data scarcity. In this paper, we introduce Dialogue Sentence Embedding (DSE), a self-supervised contrastive learning method tha…

2018

Joint Speaker Diarization and Recognition Using Convolutional and Recurrent Neural Networks

ICASSP 2018accepted

Speaker diarization (detecting who-spoke-when using relative identity labels) and speaker recognition (detecting absolute identity labels without timing) are different but related tasks that often need to be completed simultaneously in many scenarios. Traditional methods, however, address them indep…

Cited by 0SourceScholar