← Search

Pratik Joshi

2 accepted papers

2025

Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning Attack

ICML 2025poster

Safety aligned Large Language Models (LLMs) are vulnerable to harmful fine-tuning attacks -- a few harmful data mixed in the fine-tuning dataset can break the LLMs's safety alignment. While several defenses have been proposed, our evaluation shows that existing defenses fail \textit{when some specif…

Cited by 0SourcePDFScholar
2023

Continual Learning for Personalized Co-speech Gesture Generation

ICCV 2023poster

Co-speech gestures are a key channel of human communication, making them important for personalized chat agents to generate. In the past, gesture generation models assumed that data for each speaker is available all at once, and in large amounts. However in practical scenarios, speaker data comes se…

Cited by 6PDFScholar