← Search

Boshi Tang

9 accepted papers

2025

pFedGPA: Diffusion-based Generative Parameter Aggregation for Personalized Federated Learning

AAAI 2025technical

Federated Learning (FL) offers a decentralized approach to model training, where data remains local and only model parameters are shared between the clients and the central server. Traditional methods, such as Federated Averaging (FedAvg), linearly aggregate these parameters which are usually traine…

Cited by 0SourcePDFScholar
2024

DreamTime: An Improved Optimization Strategy for Diffusion-Guided 3D Generation

ICLR 2024poster

Text-to-image diffusion models pre-trained on billions of image-text pairs have recently enabled 3D content creation by optimizing a randomly initialized differentiable 3D representation with score distillation. However, the optimization process suffers slow convergence and the resultant 3D models o…

Cited by 22SourcePDFScholar
2024

Enhancing Expressiveness in Dance Generation Via Integrating Frequency and Music Style Information

ICASSP 2024accepted

Dance generation, as a branch of human motion generation, has attracted increasing attention. Recently, a few works attempt to enhance dance expressiveness, which includes genre matching, beat alignment, and dance dynamics, from certain aspects. However, the enhancement is quite limited as they lack…

Cited by 0SourceScholar
2024

Explore 3D Dance Generation via Reward Model from Automatically-Ranked Demonstrations

AAAI 2024technical

This paper presents an Exploratory 3D Dance generation framework, E3D2, designed to address the exploration capability deficiency in existing music-conditioned 3D dance generation models. Current models often generate monotonous and simplistic dance sequences that misalign with human preferences bec…

Cited by 4SourcePDFScholar
2024

Multi-View Midivae: Fusing Track- and Bar-View Representations for Long Multi-Track Symbolic Music Generation

ICASSP 2024accepted

Variational Autoencoders (VAEs) constitute a crucial component of neural symbolic music generation, among which some works have yielded outstanding results and attracted considerable attention. Nevertheless, previous VAEs still encounter issues with overly long feature sequences and generated result…

Cited by 0SourceScholar
2024

SimCalib: Graph Neural Network Calibration Based on Similarity between Nodes

AAAI 2024technical

Graph neural networks (GNNs) have exhibited impressive performance in modeling graph data as exemplified in various applications. Recently, the GNN calibration problem has attracted increasing attention, especially in cost-sensitive scenarios. Previous work has gained empirical insights on the issue…

Cited by 6SourcePDFScholar
2024

SongCreator: Lyrics-based Universal Song Generation

NeurIPS 2024poster

Music is an integral part of human culture, embodying human intelligence and creativity, of which songs compose an essential part. While various aspects of song generation have been explored by previous works, such as singing voice, vocal composition and instrumental arrangement, etc., generating so…

2024

TOSS: High-quality Text-guided Novel View Synthesis from a Single Image

ICLR 2024poster

In this paper, we present TOSS, which introduces text to the task of novel view synthesis (NVS) from just a single RGB image. While Zero123 has demonstrated impressive zero-shot open-set NVS capabilities, it treats NVS as a pure image-to-image translation problem. This approach suffers from the cha…

Cited by 18SourcePDFScholar
2024

Unsupervised Human Activity Recognition Via Large Language Models and Iterative Evolution

ICASSP 2024accepted

Human activity recognition (HAR) is crucial for health monitoring and disease diagnosis in Internet-of-Things environments. However, existing HAR approaches either suffer from poor accuracy or achieve high accuracy at the expense of costly manual annotations. To overcome the challenge above, we prop…

Cited by 0SourceScholar