← Search

Tianfu Wang

15 accepted papers

2026

A Computational Framework for Evaluating Human-likeness in LLMs' Open-ended Human Behaviors

ICML 2026poster

Large Language Models (LLMs) have found widespread application and research in scenarios such as role-playing and sociological simulations. Despite the growing use of LLM-based agents to simulate human activities, the extent to which their behaviors resemble human behavior remains underexplored. As …

Cited by 0SourceScholar
2026

LaVR: Scene Latent Conditioned Generative Video Trajectory Re-Rendering using Large 4D Reconstruction Models

CVPR 2026

Given a monocular video, the goal of video re-rendering is to generate views of the scene from a novel camera trajectory. Existing methods face two distinct challenges. Geometrically unconditioned models lack spatial awareness, leading to drift and deformation under viewpoint changes. On the other h

Cited by 0SourceScholar
2026

PolarDepth: Polarization-Guided Monocular Depth for Visual Odometry

RA-L 2026

Glass surfaces remain challenging for indoor robot perception. Depth sensors and RGB-only monocular depth estimation often fail because of reflections, refractions, and low-texture regions. To this end, we present PolarDepth, a polarization-enhanced monocular depth framework for glass-dominant envir

Cited by 0SourceScholar
2026

Virne: A Comprehensive Benchmark for RL-based Network Resource Allocation in NFV

ICLR 2026poster

Resource allocation (RA) is critical to efficient service deployment in Network Function Virtualization (NFV), a transformative networking paradigm. This task is termed NFV-RA. Recently, deep Reinforcement Learning (RL)-based methods have been showing promising potential to address this combinatoria…

Cited by 0SourcecodeScholar
2025

CoderAgent: Simulating Student Behavior for Personalized Programming Learning with Large Language Models

IJCAI 2025

Personalized programming tutoring, such as exercise recommendation, can enhance learners' efficiency, motivation, and outcomes, which is increasingly important in modern digital education. However, the lack of sufficient and high-quality programming data, combined with the mismatch between offline e

2025

Explaining Length Bias in LLM-Based Preference Evaluations

EMNLP 2025

The use of large language models (LLMs) as judges, particularly in preference comparisons, has become widespread, but this reveals a notable bias towards longer responses, undermining the reliability of such evaluations. To better understand such bias, we propose to decompose the preference evaluati

Cited by 0SourcePDFScholar
2025

Flash-Split: 2D Reflection Removal with Flash Cues and Latent Diffusion Separation

CVPR 2025poster

Transparent surfaces, such as glass, create complex reflections that obscure images and challenge downstream computer vision applications. We introduce Flash-Split, a robust framework for separating transmitted and reflected light using a single (potentially misaligned) pair of flash/no-flash images…

2025

MMFN: Multi-Feature Multi-Modal Fusion Network for Diagnosis of Superficial Lymph Node Disease

ICASSP 2025accepted

The difficulty in identifying lymph node malignancies, including lymphoma and metastatic tumors, pose a diagnostic challenge at their primary sites. Given the heterogeneity of lymph node structures across different regions and the difficulty in distinguishing them from surrounding tissues, accurate…

Cited by 0SourceScholar
2025

Repurposing Pre-trained Video Diffusion Models for Event-based Video Interpolation

CVPR 2025poster

Video Frame Interpolation aims to recover realistic missing frames between observed frames, generating a high-frame-rate video from a low-frame-rate video. However, without additional guidance, large motion between frames makes this problem ill-posed. Event-based Video Frame Interpolation (EVFI) add…

Cited by 4SourcePDFScholar
2025

TokenSelect: Efficient Long-Context Inference and Length Extrapolation for LLMs via Dynamic Token-Level KV Cache Selection

EMNLP 2025

Rapid advances in Large Language Models (LLMs) have spurred demand for processing extended context sequences in contemporary applications. However, this progress faces two challenges: performance degradation due to sequence lengths out-of-distribution, and excessively long inference times caused by

2025

Unveiling the Learning Mind of Language Models: A Cognitive Framework and Empirical Study

NeurIPS 2025poster

Large language models (LLMs) have shown impressive capabilities across tasks such as mathematics, coding, and reasoning, yet their learning ability, which is crucial for adapting to dynamic environments and acquiring new knowledge, remains underexplored. In this work, we address this gap by introduc…

Cited by 0SourceScholar
2024

DGInStyle: Domain-Generalizable Semantic Segmentation with Image Diffusion Models and Stylized Semantic Control

ECCV 2024poster

"Large, pretrained latent diffusion models (LDMs) have demonstrated an extraordinary ability to generate creative content, specialize to user data through few-shot fine-tuning, and condition their output on other modalities, such as semantic maps. However, are they usable as large-scale data generat…

2024

DGR: A General Graph Desmoothing Framework for Recommendation via Global and Local Perspectives

IJCAI 2024poster

Graph Convolutional Networks (GCNs) have become pivotal in recommendation systems for learning user and item embeddings by leveraging the user-item interaction graph's node information and topology. However, these models often face the famous over-smoothing issue, leading to indistinct user and item…

2024

FlagVNE: A Flexible and Generalizable Reinforcement Learning Framework for Network Resource Allocation

IJCAI 2024poster

Virtual network embedding (VNE) is an essential resource allocation task in network virtualization, aiming to map virtual network requests (VNRs) onto physical infrastructure. Reinforcement learning (RL) has recently emerged as a promising solution to this problem. However, existing RL-based VNE met…

2024

Object Pose Estimation via the Aggregation of Diffusion Features

CVPR 2024highlight

Estimating the pose of objects from images is a crucial task of 3D scene understanding and recent approaches have shown promising results on very large benchmarks. However these methods experience a significant performance drop when dealing with unseen objects. We believe that it results from the li…