← Search

Seyed Mohammad Hadi Hosseini

3 accepted papers

2026

Efficient Adversarial Attacks on High-dimensional Offline Bandits

ICLR 2026poster

Bandit algorithms have recently emerged as a powerful tool for evaluating machine learning models, including generative image models and large language models, by efficiently identifying top-performing candidates without exhaustive comparisons. These methods typically rely on a reward model---often…

Cited by 0SourceScholar
2026

SUSD: Structured Unsupervised Skill Discovery through State Factorization

ICLR 2026poster

Unsupervised Skill Discovery (USD) aims to autonomously learn a diverse set of skills without relying on extrinsic rewards. One of the most common USD approaches is to maximize the Mutual Information (MI) between skill latent variables and states. However, MI-based methods tend to favor simple, stat…

Cited by 0SourcecodeScholar
2025

CER: Confidence Enhanced Reasoning in LLMs

ACL 2025long

Ensuring the reliability of Large Language Models (LLMs) in complex reasoning tasks remains a formidable challenge, particularly in scenarios that demand precise mathematical calculations and knowledge-intensive open-domain generation. In this work, we introduce an uncertainty-aware framework design…