← Search

Jidong Ge

8 accepted papers

2026

CHASE: Contextual History for Adaptive and Simple Exploitation in Large Language Model Jailbreaking

AAAI 2026technical

We propose Contextual History for Adaptive and Simple Exploitation (CHASE), a novel multi-turn method for Large Language Model (LLM) jailbreaking. Rather than directly attack an LLM that may be difficult to jailbreak, CHASE first collects jailbroken histories from an easy-to-jailbreak LLM and then t

Cited by 0SourcePDFScholar
2025

InternLM-Law: An Open-Sourced Chinese Legal Large Language Model

COLING 2025main

We introduce InternLM-Law, a large language model (LLM) tailored for addressing diverse legal tasks related to Chinese laws. These tasks range from responding to standard legal questions (e.g., legal exercises in textbooks) to analyzing complex real-world legal situations. Our work contributes to Ch…

2025

LawShift: Benchmarking Legal Judgment Prediction Under Statute Shifts

NeurIPS 2025poster

Legal Judgment Prediction (LJP) seeks to predict case outcomes given available case information, offering practical value for both legal professionals and laypersons. However, a key limitation of existing LJP models is their limited adaptability to statutory revisions. Current SOTA models are neithe…

Cited by 0SourceScholar
2024

CMDL: A Large-Scale Chinese Multi-Defendant Legal Judgment Prediction Dataset

ACL 2024findings

Legal Judgment Prediction (LJP) has attracted significant attention in recent years. However, previous studies have primarily focused on cases involving only a single defendant, skipping multi-defendant cases due to complexity and difficulty. To advance research, we introduce CMDL, a large-scale rea…

2024

LJPCheck: Functional Tests for Legal Judgment Prediction

ACL 2024findings

Legal Judgment Prediction (LJP) refers to the task of automatically predicting judgment results (e.g., charges, law articles and term of penalty) given the fact description of cases. While SOTA models have achieved high accuracy and F1 scores on public datasets, existing datasets fail to evaluate sp…

Cited by 0SourcePDFScholar
2024

LawBench: Benchmarking Legal Knowledge of Large Language Models

EMNLP 2024main

We present LawBench, the first evaluation benchmark composed of 20 tasks aimed to assess the ability of Large Language Models (LLMs) to perform Chinese legal-related tasks. LawBench is meticulously crafted to enable precise assessment of LLMs’ legal capabilities from three cognitive levels that corr…

2021

Delving into Variance Transmission and Normalization: Shift of Average Gradient Makes the Network Collapse

AAAI 2021technical

Normalization operations are essential for state-of-the-art neural networks and enable us to train a network from scratch with a large learning rate (LR). We attempt to explain the real effect of Batch Normalization (BN) from the perspective of variance transmission by investigating the relationship…

2021

Don’t Miss the Potential Customers! Retrieving Similar Ads to Improve User Targeting

EMNLP 2021finding

User targeting is an essential task in the modern advertising industry: given a package of ads for a particular category of products (e.g., green tea), identify the online users to whom the ad package should be targeted. A (ad package specific) user targeting model is typically trained using histori…

Cited by 1SourcePDFScholar