← Search

Hakim Sidahmed

4 accepted papers

2025

Improving Neutral Point-of-View Generation with Data- and Parameter-Efficient RL

EMNLP 2025

The paper shows that parameter-efficient reinforcement learning (PE-RL) is a highly effective training regime to improve large language models’ (LLMs) ability to answer queries on sensitive topics with a Neutral Point of View (NPOV), i.e. to provide significantly more informative, diverse and impart

Cited by 0SourcePDFScholar
2024

Faithful Persona-based Conversational Dataset Generation with Large Language Models

ACL 2024findings

High-quality conversational datasets are essential for developing AI models that can communicate with users.One way to foster deeper interactions between a chatbot and its user is through *personas*, aspects of the user’s character that provide insights into their personality, motivations, and behav…

2021

Federated Reconstruction: Partially Local Federated Learning

NeurIPS 2021poster

Personalization methods in federated learning aim to balance the benefits of federated and local training for data availability, communication cost, and robustness to client heterogeneity. Approaches that require clients to communicate all model parameters can be undesirable due to privacy and commu…

2020

Unsupervised Anomaly Detection for Self-flying Delivery Drones

ICRA 2020poster

We propose a novel anomaly detection framework for a fleet of hybrid aerial vehicles executing high-speed package pickup and delivery missions. The detection is based on machine learning models of normal flight profiles, trained on millions of flight log measurements of control inputs and sensor rea…

Cited by 34SourceScholar