← Search

Masaki Asada

5 accepted papers

2025

A Video-grounded Dialogue Dataset and Metric for Event-driven Activities

AAAI 2025technical

This paper presents VDAct, a dataset for a Video-grounded Dialogue on Event-driven Activities, alongside VDEval, a session-based context evaluation metric specially designed for the task. Unlike existing datasets, VDAct includes longer and more complex video sequences that depict a variety of event-…

2025

Addressing the Training-Inference Discrepancy in Discrete Diffusion for Text Generation

COLING 2025main

This study addresses the discrepancy between training and inference in discrete diffusion models for text generation. We propose two novel strategies: (1) a training schema that considers two-step diffusion processes, allowing the model to use its own predicted output as input for subsequent steps d…

Cited by 0SourcePDFScholar
2025

ELAINE-medLLM: Lightweight English Japanese Chinese Trilingual Large Language Model for Bio-medical Domain

COLING 2025main

We propose ELAINE (EngLish-jApanese-chINesE)-medLLM, a trilingual (English, Japanese, Chinese) large language model adapted for the bio-medical domain based on Llama-3-8B. The training dataset was carefully curated in terms of volume and diversity to adapt to the biomedical domain and endow trilingu…

2025

Improving Relation Extraction by Sequence-to-sequence-based Dependency Parsing Pre-training

COLING 2025main

Relation extraction is a crucial natural language processing task that extracts relational triplets from raw text. Syntactic dependencies information has shown its effectiveness for relation extraction tasks. However, in most existing studies, dependency information is used only for traditional enco…

Cited by 0SourcePDFScholar
2025

ProMQA: Question Answering Dataset for Multimodal Procedural Activity Understanding

NAACL 2025long

Multimodal systems have great potential to assist humans in procedural activities, where people follow instructions to achieve their goals. Despite diverse application scenarios, systems are typically evaluated on traditional classification tasks, e.g., action recognition or temporal action localiza…