← Search

Zhuowen Yuan

7 accepted papers

2025

MMDT: Decoding the Trustworthiness and Safety of Multimodal Foundation Models

ICLR 2025poster

Multimodal foundation models (MMFMs) play a crucial role in various applications, including autonomous driving, healthcare, and virtual assistants. However, several studies have revealed vulnerabilities in these models, such as generating unsafe content by text-to-image models. Existing benchmarks o…

2024

Fair Federated Learning via the Proportional Veto Core

ICML 2024poster

Previous work on fairness in federated learning introduced the notion of *core stability*, which provides utility-based fairness guarantees to any subset of participating agents. However, these guarantees require strong assumptions on agent utilities that render them impractical. To address this sho…

Cited by 7SourcePDFScholar
2024

RigorLLM: Resilient Guardrails for Large Language Models against Undesired Content

ICML 2024poster

Recent advancements in Large Language Models (LLMs) have showcased remarkable capabilities across various tasks in different domains. However, the emergence of biases and the potential for generating harmful content in LLMs, particularly under malicious inputs, pose significant challenges. Current m…

2023

FedGame: A Game-Theoretic Defense against Backdoor Attacks in Federated Learning

NeurIPS 2023poster

Federated learning (FL) provides a distributed training paradigm where multiple clients can jointly train a global model without sharing their local data. However, recent studies have shown that FL offers an additional surface for backdoor attacks. For instance, an attacker can compromise a subset o…

2023

Incentives in Federated Learning: Equilibria, Dynamics, and Mechanisms for Welfare Maximization

NeurIPS 2023poster

Federated learning (FL) has emerged as a powerful scheme to facilitate the collaborative learning of models amongst a set of agents holding their own private data. Although the agents benefit from the global model trained on shared data, by participating in federated learning, they may also incur c…

Cited by 14SourcePDFScholar
2022

SecretGen: Privacy Recovery on Pre-trained Models via Distribution Discrimination

ECCV 2022poster

"Transfer learning through the use of pre-trained models has become a growing trend for the machine learning community. Consequently, numerous pre-trained models are released online to facilitate further research. However, it raises extensive concerns on whether these pre-trained models would leak p…