← Search

Rujikorn Charakorn

5 accepted papers

2026

Doc-to-LoRA: Learning to Instantly Internalize Contexts

ICML 2026poster

Long input sequences are central to in-context learning, document understanding, and multi-step reasoning of Large Language Models (LLMs). However, the quadratic attention cost of Transformers makes inference memory-intensive and slow. While context distillation (CD) can transfer information into mo…

Cited by 0SourceScholar
2025

Text-to-LoRA: Instant Transformer Adaption

ICML 2025poster

While Foundation Models provide a general tool for rapid content creation, they regularly require task-specific adaptation. Traditionally, this exercise involves careful curation of datasets and repeated fine-tuning of the underlying model. Fine-tuning techniques enable practitioners to adapt found…

2024

Cleanba: A Reproducible and Efficient Distributed Reinforcement Learning Platform

ICLR 2024poster

Distributed Deep Reinforcement Learning (DRL) aims to leverage more computational resources to train autonomous agents with less training time. Despite recent progress in the field, reproducibility issues have not been sufficiently explored. This paper first shows that the typical actor-learner fram…

2024

Diversity Is Not All You Need: Training A Robust Cooperative Agent Needs Specialist Partners

NeurIPS 2024poster

Partner diversity is known to be crucial for training a robust generalist cooperative agent. In this paper, we show that partner specialization, in addition to diversity, is crucial for the robustness of a downstream generalist agent. We propose a principled method for quantifying both the diversity…

Cited by 1SourcePDFScholar
2023

Generating Diverse Cooperative Agents by Learning Incompatible Policies

ICLR 2023top-25%

Training a robust cooperative agent requires diverse partner agents. However, obtaining those agents is difficult. Previous works aim to learn diverse behaviors by changing the state-action distribution of agents. But, without information about the task's goal, the diversified agents are not guided…

Cited by 34SourcePDFScholar