← Search

Demetri Terzopoulos

10 accepted papers

2026

AniMimic: Imitating 3D Animation from Video Priors

CVPR 2026

Creating realistic 3D animation remains a time-consuming and expertise-dependent process, requiring manual rigging, keyframing, and fine-tuning of complex motions. Meanwhile, video diffusion models have recently demonstrated remarkable 2D motion imagination, generating dynamic and visually coherent

Cited by 0SourceScholar
2025

CFSum: A Transformer-Based Multi-Modal Video Summarization Framework With Coarse-Fine Fusion

ICASSP 2025accepted

Video summarization, by selecting the most informative and/or user-relevant parts of original videos to create concise summary videos, has high research value and consumer demand in today’s video proliferation era. Multi-modal video summarization that accomodates user input has become a research hot…

Cited by 0SourceScholar
2025

Inverse Attention Agents for Multi-Agent Systems

ICLR 2025poster

A major challenge for Multi-Agent Systems (MAS) is enabling agents to adapt dynamically to diverse environments in which opponents and teammates may continually change. Agents trained using conventional methods tend to excel only within the confines of their training cohorts; their performance drops…

2025

Wonderland: Navigating 3D Scenes from a Single Image

CVPR 2025poster

This paper addresses a challenging question: how can we efficiently create high-quality, wide-scope 3D scenes from a single arbitrary image?Existing methods face several constraints, such as requiring multi-view data, time-consuming per-scene optimization, low visual quality, and distorted reconstru…

Cited by 12SourcePDFScholar
2024

MindAgent: Emergent Gaming Interaction

NAACL 2024findings

Large Foundation Models (LFMs) can perform complex scheduling in a multi-agent system and can coordinate agents to complete sophisticated tasks that require extensive collaboration.However, despite the introduction of numerous gaming frameworks, the community lacks adequate benchmarks that support t…

Cited by 100SourcePDFScholar
2023

ARNOLD: A Benchmark for Language-Grounded Task Learning with Continuous States in Realistic 3D Scenes

ICCV 2023poster

Understanding the continuous states of objects is essential for task learning and planning in the real world. However, most existing task learning benchmarks assume discrete (e.g., binary) object states, which poses challenges for learning complex tasks and transferring learned policy from the simul…

Cited by 29PDFcodeScholar
2023

mBEST: Realtime Deformable Linear Object Detection Through Minimal Bending Energy Skeleton Pixel Traversals

RA-L 2023

Robotic manipulation of deformable materials is a challenging task that often requires realtime visual feedback. This is especially true for deformable linear objects (DLOs) or “rods”, whose slender and flexible structures make proper tracking and detection nontrivial. To address this challenge, we

Cited by 28SourcecodeScholar
2020

Deep Learning of Neuromuscular and Visuomotor Control of a Biomimetic Simulated Humanoid

RA-L 2020

We present a biomimetic framework for human neuromuscular and visuomotor control that promises to be of value to researchers developing humanoid robots. Our framework features a biomechanically simulated human musculoskeletal model, actuated by numerous skeletal muscles, with realistic eyes driven b

Cited by 2SourceScholar
2020

End-to-End Trainable Deep Active Contour Models for Automated Image Segmentation: Delineating Buildings in Aerial Imagery

ECCV 2020poster

The automated segmentation of buildings in remote sensing imagery is a challenging task that requires the accurate delineation of multiple building instances over typically large image areas. Manual methods are often laborious and current deep-learning-based approaches fail to delineate all building…

Cited by 70SourcePDFScholar
2016

Inferring Forces and Learning Human Utilities From Videos

CVPR 2016oral

We propose a notion of affordance that takes into account physical quantities generated when the human body interacts with real-world objects, and introduce a learning framework that incorporates the concept of human utilities, which in our opinion provides a deeper and finer-grained account not onl…

Cited by 113PDFScholar