← Search

Duc Thien Nguyen

5 accepted papers

2025

Mastering the Craft of Data Synthesis for CodeLLMs

NAACL 2025long

Large language models (LLMs) have shown impressive performance in code understanding and generation, making coding tasks a key focus for researchers due to their practical applications and value as a testbed for LLM evaluation. Data synthesis and filtering techniques have been widely adopted and sho…

2024

Multi-Linear Kernel Regression and Imputation VIA Manifold Learning: the Dynamic MRI Case

ICASSP 2024accepted

This paper introduces an efficient multi-linear nonparametric (kernel-based) approximation framework for data regression and imputation. Data features are assumed to reside in or close to a smooth and userunknown manifold embedded in a reproducing kernel Hilbert space. Landmark points are identified…

Cited by 0SourceScholar
2022

Neural-progressive hedging: Enforcing constraints in reinforcement learning with stochastic programming

UAI 2022poster

We propose a framework, called neural-progressive hedging (NP), that leverages stochastic programming during the online phase of executing a reinforcement learning (RL) policy. The goal is to ensure feasibility with respect to constraints and risk-based objectives such as conditional value-at-risk…

2018

Credit Assignment For Collective Multiagent RL With Global Rewards

NeurIPS 2018poster

Scaling decision theoretic planning to large multiagent systems is challenging due to uncertainty and partial observability in the environment. We focus on a multiagent planning model subclass, relevant to urban settings, where agent interactions are dependent on their ``collective influence'' on ea…

Cited by 133SourcePDFScholar
2017

Policy Gradient With Value Function Approximation For Collective Multiagent Planning

NeurIPS 2017poster

Decentralized (PO)MDPs provide an expressive framework for sequential decision making in a multiagent system. Given their computational complexity, recent research has focused on tractable yet practical subclasses of Dec-POMDPs. We address such a subclass called CDec-POMDP where the collective behav…

Cited by 73SourcePDFScholar