← Search

Yiming Qian

23 accepted papers

2025

Unveiling Maternity and Infant Care Conversations: A Chinese Dialogue Dataset for Enhanced Parenting Support

IJCAI 2025

The rapid development of large language models has greatly advanced human-computer dialogue research. However, applying these models to specialized fields like maternity and infant care often leads to subpar performance due to a lack of domain-specific datasets. To address this problem, we have crea

2024

DPA-Net: Structured 3D Abstraction from Sparse Views via Differentiable Primitive Assembly

ECCV 2024poster

"We present a differentiable rendering framework to learn structured 3D abstractions in the form of primitive assemblies from sparse RGB images capturing a 3D object. By leveraging differentiable volume rendering, our method does not require 3D supervision. Architecturally, our network follows the g…

Cited by 4SourcePDFScholar
2024

DiReCT: Diagnostic Reasoning for Clinical Notes via Large Language Models

NeurIPS 2024poster

Large language models (LLMs) have recently showcased remarkable capabilities, spanning a wide range of tasks and applications, including those in the medical domain. Models like GPT-4 excel in medical question answering but may face challenges in the lack of interpretability when handling complex ta…

2024

Harnessing the Power of Large Language Model for Uncertainty Aware Graph Processing

COLING 2024main

Handling graph data is one of the most difficult tasks. Traditional techniques, such as those based on geometry and matrix factorization, rely on assumptions about the data relations that become inadequate when handling large and complex graph data. On the other hand, deep learning approaches demons…

2024

MHGRL: An Effective Representation Learning Model for Electronic Health Records

COLING 2024main

Electronic health records (EHRs) serve as a digital repository storing comprehensive medical information about patients. Representation learning for EHRs plays a crucial role in healthcare applications. In this paper, we propose a Multimodal Heterogeneous Graph-enhanced Representation Learning, deno…

2024

SEER: Facilitating Structured Reasoning and Explanation via Reinforcement Learning

ACL 2024long

Elucidating the reasoning process with structured explanations from question to answer is crucial, as it significantly enhances the interpretability, traceability, and trustworthiness of question-answering (QA) systems. However, structured explanations demand models to perform intricately structured…

2023

HAL3D: Hierarchical Active Learning for Fine-Grained 3D Part Labeling

ICCV 2023poster

We present the first active learning tool for fine-grained 3D part labeling, a problem which challenges even the most advanced deep learning (DL) methods due to the significant structural variations among the intricate parts. For the same reason, the necessary effort to annotate training data is tre…

Cited by 1PDFScholar
2023

MPrompt: Exploring Multi-level Prompt Tuning for Machine Reading Comprehension

EMNLP 2023long findings

The large language models have achieved superior performance on various natural language tasks. One major drawback of such approaches is they are resource-intensive in fine-tuning new datasets. Soft-prompt tuning presents a resource-efficient solution to fine-tune the pre-trained language models (PL…

Cited by 0SourcecodeScholar
2023

Point-TTA: Test-Time Adaptation for Point Cloud Registration Using Multitask Meta-Auxiliary Learning

ICCV 2023poster

We present Point-TTA, a novel test-time adaptation framework for point cloud registration (PCR) that improves the generalization and the performance of registration models. While learning-based approaches have achieved impressive progress, generalization to unknown testing environments remains a maj…

Cited by 21PDFScholar
2023

TCRA-LLM: Token Compression Retrieval Augmented Large Language Model for Inference Cost Reduction

EMNLP 2023long findings

Since ChatGPT released its API for public use, the number of applications built on top of commercial large language models (LLMs) increase exponentially. One popular usage of such models is leveraging its in-context learning ability and generating responses given user queries leveraging knowledge ob…

Cited by 0SourceScholar
2022

A Reliable Online Method for Joint Estimation of Focal Length and Camera Rotation

ECCV 2022poster

"Linear perspective cues deriving from regularities of the built environment can be used to recalibrate both intrinsic and extrinsic camera parameters online, but these estimates can be unreliable due to irregularities in the scene, uncertainties in line segment estimation and background clutter. He…

2022

HEAT: Holistic Edge Attention Transformer for Structured Reconstruction

CVPR 2022poster

This paper presents a novel attention-based neural network for structured reconstruction, which takes a 2D raster image as an input and reconstructs a planar graph depicting an underlying geometric structure. The approach detects corners and classifies edge candidates between corners in an end-to-en…

Cited by 42PDFcodeScholar
2022

Single User WiFi Structure from Motion in the Wild

ICRA 2022poster

This paper proposes a novel motion estimation algorithm using WiFi networks and IMU sensor data in large uncontrolled environments, dubbed “WiFi Structure-from-Motion” (WiFi SfM). Given smartphone sensor data through day-to-day activities from a single user over a month, our WiFi SfM algorithm estim…

Cited by 4SourceScholar
2021

Fusion-DHL: WiFi, IMU, and Floorplan Fusion for Dense History of Locations in Indoor Environments

ICRA 2021poster

The paper proposes a multi-modal sensor fusion algorithm that fuses WiFi, IMU, and floorplan information to infer an accurate and dense location history in indoor environments. The algorithm uses 1) an inertial navigation algorithm to estimate a relative motion trajectory from IMU sensor data; 2) a…

Cited by 29SourcecodeScholar
2021

Roof-GAN: Learning To Generate Roof Geometry and Relations for Residential Houses

CVPR 2021poster

This paper presents Roof-GAN, a novel generative adversarial network that generates structured geometry of residential roof structures as a set of roof primitives and their relationships. Given the number of primitives, the generator produces a structured roof model as a graph, which consists of 1)…

Cited by 21PDFcodeScholar
2020

3D Human Shape Reconstruction from a Polarization Image

ECCV 2020poster

This paper tackles the problem of estimating 3D body shape of clothed humans from single polarized 2D images, i.e. polarization images. Polarization images are known to be able to capture polarized reflected lights that preserve rich geometric cues of an object, which has motivated its recent applic…

Cited by 56SourcePDFScholar
2020

Learning Pairwise Inter-Plane Relations for Piecewise Planar Reconstruction

ECCV 2020poster

This paper proposes a novel single-image piecewise planar reconstruction technique that infers and enforces inter-plane relationships. Our approach takes a planar reconstruction result from an existing system, then utilizes convolutional neural network (CNN) to (1) classify if two planes are orthogo…

2018

Simultaneous 3D Reconstruction for Water Surface and Underwater Scene

ECCV 2018poster

This paper presents the first approach for simultaneously recovering the 3D shape of both the wavy water surface and the moving underwater scene. A portable camera array system is constructed, which captures the scene from multiple viewpoints above the water. The correspondences across these cameras…

Cited by 43SourcePDFScholar