← Search

Dawei Feng

14 accepted papers

2025

Complementary Learning System Theory-based Active Learning for Audio Classification

ICASSP 2025accepted

Deep learning has significantly advanced the audio classification, achieving remarkable results. However, these successes often rely on extensive manual annotation of audio, a labor-intensive and costly process. Active Learning (AL) presents a promising solution by minimizing the required amount of…

Cited by 0SourceScholar
2025

Enhancing Decision-Making for LLM Agents via Step-Level Q-Value Models

AAAI 2025technical

Agents significantly enhance the capabilities of standalone Large Language Models (LLMs) by perceiving environments, making decisions, and executing actions. However, LLM agents still face challenges in tasks that require multiple decision-making steps. Estimating the value of actions in specific ta…

Cited by 8SourcePDFScholar
2025

Scaling Bioacoustic Signal Pre-training with Million Samples Via Mask-Modeling

ICASSP 2025accepted

Deep learning-based bioacoustic audio analysis holds immense potential across various applications. However, existing studies in bioacoustics often focus on a limited number of species, potentially hindering the transferability of models across different species. Furthermore, the manual annotation o…

Cited by 0SourceScholar
2025

V-Pilot: A Velocity Vector Control Agent for Fixed-Wing UAVs from Imperfect Demonstrations

ICRA 2025

This paper addresses the challenge of Velocity Vector Control (VVC) for fixed-wing UAVs using Reinforcement Learning (RL) in the presence of imperfect demonstrations. The multi-objective and long-horizon nature of VVC introduces significant spatial and temporal complexities, complicating RL's explor

Cited by 0SourceScholar
2024

Optimistic Model Rollouts for Pessimistic Offline Policy Optimization

AAAI 2024technical

Model-based offline reinforcement learning (RL) has made remarkable progress, offering a promising avenue for improving generalization with synthetic model rollouts. Existing works primarily focus on incorporating pessimism for policy optimization, usually via constructing a Pessimistic Markov Decis…

Cited by 1SourcePDFScholar
2024

Transformer-Inspired Lightweight Model for Efficient Time Series Forecasting

ICASSP 2024accepted

Accuracy and efficiency are pivotal considerations in the field of time series forecasting. Through the integration of meticulously designed temporal components, the Transformer-based models have significantly enhanced the accuracy of time series prediction. However, due to the utilization of attent…

Cited by 0SourceScholar
2023

Complementary Learning System Based Intrinsic Reward in Reinforcement Learning

ICASSP 2023accepted

Deep reinforcement learning has achieved encouraging performance in many realms. However, one of its primary challenges is the sparsity of extrinsic rewards, which is still far from solved. Complementary learning system theory suggests that effective human learning relies on two complementary learni…

Cited by 0SourceScholar
2023

Diversifying Message Aggregation in Multi-Agent Communication Via Normalized Tensor Nuclear Norm Regularization

ICASSP 2023accepted

The use of graph attention networks (GAT) in communication-enhanced multi-agent reinforcement learning (Comm-MARL) has become prevalent. While successful, GAT can lead to homogeneity in the strategies of message aggregation, which can severely limit multi-agent coordination. To address this challeng…

Cited by 0SourceScholar
2023

Progressive Diversifying Policy for Multi-Agent Reinforcement Learning

ICASSP 2023accepted

Multi-Agent Reinforcement Learning (MARL) has recently achieved promising performance in many collaborative decision making tasks. However, one of the main bottleneck challenges for MARL is the sparsity of the team reward, which can lead to the homogenization of agents’ behaviors. To address these i…

Cited by 0SourceScholar
2022

FINT: Field-Aware Interaction Neural Network for Click-Through Rate Prediction

ICASSP 2022accepted

As a critical component for online advertising and marketing, click-through rate (CTR) prediction has drawn lots of attention from both industry and academia. Recently, deep learning has become the mainstream methodological choice for CTR. Despite sustainable efforts have been made, existing approac…

Cited by 0SourceScholar
2021

Diversity and Consistency: Exploring Visual Question-Answer Pair Generation

EMNLP 2021finding

Although showing promising values to downstream applications, generating question and answer together is under-explored. In this paper, we introduce a novel task that targets question-answer pair generation from visual images. It requires not only generating diverse question-answer pairs but also ke…

2019

Denoising Convolutional Autoencoder Based B-mode Ultrasound Tongue Image Feature Extraction

ICASSP 2019accepted

B-mode ultrasound tongue imaging is widely used in the speech production field. However, efficient interpretation is in a great need for the tongue image sequences. Inspired by the recent success of unsupervised deep learning approach, we explore unsupervised convolutional network architecture for t…

Cited by 0SourceScholar