← Search

Haibo Jin

9 accepted papers

2026

Agent Primitives: Reuseable Latent Building Blocks for Multi-Agent Systems

ICML 2026poster

While existing multi-agent systems (MAS) can handle complex problems by enabling collaboration among multiple agents, they are often highly task-specific, relying on manually crafted agent roles and interaction prompts, which leads to increased architectural complexity and limited reusability across…

Cited by 0SourceScholar
2025

Evaluating the Inductive Abilities of Large Language Models: Why Chain-of-Thought Reasoning Sometimes Hurts More Than Helps

NeurIPS 2025poster

Large Language Models (LLMs) have shown remarkable progress across domains, yet their ability to perform inductive reasoning—inferring latent rules from sparse examples—remains limited. It is often assumed that chain-of-thought (CoT) prompting, as used in Large Reasoning Models (LRMs), enhances suc…

Cited by 0SourceScholar
2025

Exploring the Vulnerability of the Content Moderation Guardrail in Large Language Models via Intent Manipulation

EMNLP 2025

Intent detection, a core component of natural language understanding, has considerably evolved as a crucial mechanism in safeguarding large language models (LLMs). While prior work has applied intent detection to enhance LLMs’ moderation guardrails, showing a significant success against content-leve

Cited by 0SourcePDFScholar
2025

Revolve: Optimizing AI Systems by Tracking Response Evolution in Textual Optimization

ICML 2025poster

Recent advancements in large language models (LLMs) have significantly enhanced the ability of LLM-based systems to perform complex tasks through natural language processing and tool interaction. However, optimizing these LLM-based systems for specific tasks remains challenging, often requiring manu…

2024

CatchBackdoor: Backdoor Detection via Critical Trojan Neural Path Fuzzing

ECCV 2024poster

"The success of deep neural networks (DNNs) in real-world applications has benefited from abundant pre-trained models. However, the backdoored pre-trained models can pose a significant trojan threat to the deployment of downstream DNNs. Numerous backdoor detection methods have been proposed but are…

Cited by 2SourcePDFScholar
2024

EditShield: Protecting Unauthorized Image Editing by Instruction-guided Diffusion Models

ECCV 2024poster

"Text-to-image diffusion models have emerged as an evolutionary for producing creative content in image synthesis. Based on the impressive generation abilities of these models, instruction-guided diffusion models can edit images with simple instructions and input images. While they empower users to…

Cited by 12SourcePDFScholar
2024

Jailbreaking Large Language Models Against Moderation Guardrails via Cipher Characters

NeurIPS 2024poster

Large Language Models (LLMs) are typically harmless but remain vulnerable to carefully crafted prompts known as ``jailbreaks'', which can bypass protective measures and induce harmful behavior. Recent advancements in LLMs have incorporated moderation guardrails that can filter outputs, which trigger…

Cited by 14SourcePDFScholar
2024

PromptMRG: Diagnosis-Driven Prompts for Medical Report Generation

AAAI 2024technical

Automatic medical report generation (MRG) is of great research value as it has the potential to relieve radiologists from the heavy burden of report writing. Despite recent advancements, accurate MRG remains challenging due to the need for precise clinical understanding and disease identification. M…

2022

RePFormer: Refinement Pyramid Transformer for Robust Facial Landmark Detection

IJCAI 2022poster

This paper presents a Refinement Pyramid Transformer (RePFormer) for robust facial landmark detection. Most facial landmark detectors focus on learning representative image features. However, these CNN-based feature representations are not robust enough to handle complex real-world scenarios due to…

Cited by 19SourcePDFScholar