← Search

Guo Yang

3 accepted papers

2023

Dynamic Stashing Quantization for Efficient Transformer Training

EMNLP 2023short findings

Large Language Models (LLMs) have demonstrated impressive performance on a range of Natural Language Processing (NLP) tasks. Unfortunately, the immense amount of computations and memory accesses required for LLM training makes them prohibitively expensive in terms of hardware cost, and thus challeng…

Cited by 0SourceScholar
2022

Incremental Context Aware Attentive Knowledge Tracing

ICASSP 2022accepted

Knowledge Tracing is the prediction of the future performance of a learner, given the past performance. The existing knowledge tracing models represent the training data and does not generalize when there is a drift in the data distribution. We first empirically demonstrate an evolving Knowledge Tra…

Cited by 0SourceScholar
2022

Online Continual Learning Using Enhanced Random Vector Functional Link Networks

ICASSP 2022accepted

We propose an online continual learning algorithm based on an enhanced Random Vector Functional Link Network (OCL-eRVFL), that learns a sequence of tasks continually, where each task is defined by streaming data with each sample arriving once and only once. As data for a new task in domain increment…

Cited by 0SourceScholar