← Search

Guoqing Ma

5 accepted papers

2026

Capturing Dynamic User Interests Under Modality Imbalance for Multimodal Sequential Recommendation

AAAI 2026technical

Multimodal sequential recommender systems leverage diverse modal inputs to enhance the accuracy and relevance of personalized recommendations. However, existing fusion strategies often struggle to capture intricate cross-modal interactions, especially under the evolving dynamics of user intent. More

Cited by 0SourcePDFScholar
2026

Orthogonal Weight Modification Enhances Learning Scalability and Convergence Efficiency without Gradient Backpropagation

ICASSP 2026poster

Recognizing the substantial computational cost of backpropagation (BP), non-BP methods have emerged as attractive alternatives for efficient learning on emerging neuromorphic systems. However, existing non-BP approaches still face critical challenges in efficiency and scalability. Inspired by neural…

Cited by 0SourcePDFScholar
2026

SetPO: Set-Level Policy Optimization for Diversity-Preserving LLM Reasoning

ICML 2026poster

Reinforcement learning with verifiable rewards has shown notable effectiveness in enhancing large language models (LLMs) reasoning performance, especially in mathematics tasks. However, such improvements often come with reduced outcome diversity, where the model concentrates probability mass on a na…

Cited by 0SourceScholar
2025

Efficient Reinforcement Learning Through Adaptively Pretrained Visual Encoder

AAAI 2025technical

While Reinforcement Learning (RL) agents can successfully learn to handle complex tasks, effectively generalizing acquired skills to unfamiliar settings remains a challenge. One of the reasons behind this is the visual encoder used are task-dependent, preventing effective feature extraction in diffe…

Cited by 0SourcePDFScholar
2025

Generative Pre-trained Autoregressive Diffusion Transformer

NeurIPS 2025poster

In this work, we present GPDiT, a Generative Pre-trained Autoregressive Diffusion Transformer that unifies the strengths of diffusion and autoregressive modeling for long-range video synthesis, within a continuous latent space. Instead of predicting discrete tokens, GPDiT autoregressively predicts f…

Cited by 0SourceScholar