← Search

Yuanyuan Xu

5 accepted papers

2026

Unlocking Multi-Modal Potentials for Link Prediction on Dynamic Text-Attributed Graphs

AAAI 2026technical

Dynamic Text-Attributed Graphs (DyTAGs) are a novel graph paradigm that captures evolving temporal events (edges) alongside rich textual attributes. Existing studies can be broadly categorized into TGNN-driven and LLM-driven approaches, both of which encode textual attributes and temporal structures

Cited by 0SourcePDFScholar
2024

LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens

ICML 2024poster

Large context window is a desirable feature in large language models (LLMs). However, due to high fine-tuning costs, scarcity of long texts, and catastrophic values introduced by new token positions, current extended context windows are limited to around 128k tokens. This paper introduces LongRoPE t…

2024

TECA: A Two-stage Approach with Controllable Attention Soft Prompt for Few-shot Nested Named Entity Recognition

COLING 2024main

Few-shot nested named entity recognition (NER), identifying named entities that are nested with a small number of labeled data, has attracted much attention. Recently, a span-based method based on three stages ( focusing, bridging and prompting) has been proposed for few-shot nested NER. However, su…

Cited by 3SourcePDFScholar
2023

Focusing, Bridging and Prompting for Few-shot Nested Named Entity Recognition

ACL 2023findings

Few-shot named entity recognition (NER), identifying named entities with a small number of labeled data, has attracted much attention. Frequently, entities are nested within each other. However, most of the existing work on few-shot NER addresses flat entities instead of nested entities. To tackle n…

Cited by 4SourcePDFScholar
2022

FOV-Based Coding Optimization for 360-Degree Virtual Reality Videos

ICASSP 2022accepted

Panoramic or 360-degree virtual reality videos have high resolution, frame rate, and visual quality that demand efficient coding. Although a user watching a 360-degree video can switch viewing angles, only a portion of the video in the user’s Field of View (FoV) is displayed at any time. In this pap…

Cited by 0SourceScholar