← Search

Andrew Wei Tung Siah

2 accepted papers

2025

Data Mixture Optimization: A Multi-fidelity Multi-scale Bayesian Framework

NeurIPS 2025poster

Careful curation of data sources can significantly improve the performance of LLM pre-training, but predominant approaches rely heavily on intuition or costly trial-and-error, making them difficult to generalize across different data domains and downstream tasks. Although scaling laws can provide a…

Cited by 0SourcecodeScholar
2025

PersonalLLM: Tailoring LLMs to Individual Preferences

ICLR 2025poster

As LLMs become capable of complex tasks, there is growing potential for personalized interactions tailored to the subtle and idiosyncratic preferences of the user. We present a public benchmark, PersonalLLM, focusing on adapting LLMs to provide maximal benefits for a particular user. Departing from…