← Search

Karim Bouyarmane

6 accepted papers

2026

Learning to Reason Efficiently with Discounted Reinforcement Learning

ICLR 2026poster

Large reasoning models (LRMs) often consume excessive tokens, inflating computational cost and latency. We challenge the assumption that longer responses improve accuracy. By penalizing the reasoning tokens using a discounted reinforcement-learning setup (interpretable as a small per-token cost) and…

Cited by 0SourcecodeScholar
2026

Universal Guideline-Driven Image Clustering via a Hybrid LLM Agent

CVPR 2026

Unifying image clustering across different clustering scenarios remains challenging due to fundamental gaps among tasks. We introduce a Guideline-Driven Image Clustering Agent, the first universal framework that bridges these gaps through textual guidelines. To incorporate complex guidelines without

Cited by 0SourceScholar
2025

Zero-Shot Composed Image Retrieval via Dual-Stream Instruction-Aware Distillation

ICCV 2025poster

Composed Image Retrieval (CIR) targets the retrieval of images conditioned on a reference image and a textual modification, but constructing labeled triplets (reference image, textual modification, target image) is inherently challenging. Existing Zero-Shot CIR (ZS-CIR) approaches often rely on well…

Cited by 0SourcePDFScholar
2024

Structured Object Language Modeling (SO-LM): Native Structured Objects Generation Conforming to Complex Schemas with Self-Supervised Denoising

EMNLP 2024industry

In this paper, we study the problem of generating structured objects that conform to a complex schema, with intricate dependencies between the different components (facets) of the object. The facets of the object (attributes, fields, columns, properties) can be a mix of short, structured facts, or l…

Cited by 0SourcePDFScholar
2018

Generating Assistive Humanoid Motions for Co-Manipulation Tasks with a Multi-Robot Quadratic Program Controller

ICRA 2018poster

Human-humanoid collaborative tasks require that the robot take into account the goals of the task, interaction forces with the human, and its own balance. We present a formulation for a real-time humanoid controller which allows the robot to keep itself balanced, while also assisting the human in ac…

Cited by 26SourceScholar