← Search

Ana C. Murillo

13 accepted papers

2024

CLIPSwarm: Generating Drone Shows from Text Prompts with Vision-Language Models

IROS 2024poster

This paper introduces CLIPSwarm, a new algorithm designed to automate the modeling of swarm drone formations based on natural language. The algorithm begins by enriching a provided word, to compose a text prompt that serves as input to an iterative approach to find the formation that best matches th…

Cited by 4SourceScholar
2024

SpectralWaste Dataset: Multimodal Data for Waste Sorting Automation

IROS 2024poster

The increase in non-biodegradable waste is a worldwide concern. Recycling facilities play a crucial role, but their automation is hindered by the complex characteristics of waste recycling lines like clutter or object deformation. In addition, the lack of publicly available labeled data for these en…

Cited by 1SourceScholar
2023

A Framework for Fast Prototyping of Photo-realistic Environments with Multiple Pedestrians

ICRA 2023poster

Robotic applications involving people often require advanced perception systems to better understand complex real-world scenarios. To address this challenge, photo-realistic and physics simulators are gaining popularity as a means of generating accurate data labeling and designing scenarios for eval…

Cited by 2SourceScholar
2023

CineTransfer: Controlling a Robot to Imitate Cinematographic Style from a Single Example

IROS 2023poster

This work presents CineTransfer, an algorithmic framework that drives a robot to record a video sequence that mimics the cinematographic style of an input video. We propose features that abstract the aesthetic style of the input video, so the robot can transfer this style to a scene with visual deta…

Cited by 2SourceScholar
2022

CineMPC: Controlling Camera Intrinsics and Extrinsics for Autonomous Cinematography

ICRA 2022poster

We present CineMPC, an algorithm to autonomously control a UAV-borne video camera in a nonlinear Model Predicted Control (MPC) loop. CineMPC controls both the position and orientation of the camera-the camera extrinsics-as well as the lens focal length, focal distance, and aperture-the camera intrin…

Cited by 8SourceScholar
2021

Semi-Supervised Semantic Segmentation With Pixel-Level Contrastive Learning From a Class-Wise Memory Bank

ICCV 2021poster

This work presents a novel approach for semi-supervised semantic segmentation. The key element of this approach is our contrastive learning module that enforces the segmentation network to yield similar pixel-level feature representations for same-class samples across the whole dataset. To achieve t…

Cited by 288PDFcodeScholar
2020

3D-MiniNet: Learning a 2D Representation From Point Clouds for Fast and Efficient 3D LIDAR Semantic Segmentation

RA-L 2020

LIDAR semantic segmentation is an essential task that provides 3D semantic information about the environment to robots. Fast and efficient semantic segmentation methods are needed to match the strong computational and temporal restrictions of many real-world robotic applications. This work presents

Cited by 162SourcecodeScholar
2019

Enhancing V-SLAM Keyframe Selection with an Efficient ConvNet for Semantic Analysis

ICRA 2019poster

Selecting relevant visual information from a video is a challenging task on its own and even more in robotics, due to strong computational restrictions. This work proposes a novel keyframe selection strategy based on image quality and semantic information, which boosts strategies currently used in V…

Cited by 19SourceScholar
2017

A multimodal dataset for object model learning from natural human-robot interaction

IROS 2017poster

Learning object models in the wild from natural human interactions is an essential ability for robots to perform general tasks. In this paper we present a robocentric multimodal dataset addressing this key challenge. Our dataset focuses on interactions where the user teaches new objects to the robot…

Cited by 20SourceScholar