← Search

Jiamu Bai

4 accepted papers

2026

Towards Better Optimization For Listwise Preference in Diffusion Models

ICLR 2026poster

Reinforcement learning from human feedback (RLHF) has proven effectiveness for aligning text-to-image (T2I) diffusion models with human preferences. Although Direct Preference Optimization (DPO) is widely adopted for its computational efficiency and avoidance of explicit reward modeling, its applica…

Cited by 0SourceScholar
2025

FlowerTune: A Cross-Domain Benchmark for Federated Fine-Tuning of Large Language Models

NeurIPS 2025poster

Large Language Models (LLMs) have achieved state-of-the-art results across diverse domains, yet their development remains reliant on vast amounts of publicly available data, raising concerns about data scarcity and the lack of access to domain-specific, sensitive information. Federated Learning (FL)…

Cited by 0SourceScholar
2024

Federated Fine-tuning of Large Language Models under Heterogeneous Tasks and Client Resources

NeurIPS 2024poster

Federated Learning (FL) has recently been applied to the parameter-efficient fine-tuning of Large Language Models (LLMs). While promising, it raises significant challenges due to the heterogeneous resources and data distributions of clients.This study introduces FlexLoRA, a simple yet effective agg…

Cited by 27SourcePDFScholar
2023

Among Us: Adversarially Robust Collaborative Perception by Consensus

ICCV 2023poster

Multiple robots could perceive a scene (e.g., detect objects) collaboratively better than individuals, although easily suffer from adversarial attacks when using deep learning. This could be addressed by the adversarial defense, but its training requires the often-unknown attacking mechanism. Differ…

Cited by 33PDFcodeScholar