← Search

Masakazu Yoshimura

8 accepted papers

2026

Online Data Curation for Object Detection via Marginal Contributions to Dataset-level Average Precision

CVPR 2026

High-quality data has become a primary driver of progress under scale laws, with curated datasets often outperforming much larger unfiltered ones at lower cost. Online data curation extends this idea by dynamically selecting training samples based on the model's evolving state. While effective in cl

Cited by 0SourceScholar
2026

SF-Mamba: Rethinking State Space Model for Vision

ICML 2026poster

The realm of Mamba for vision has been advanced in recent years to strike for the alternatives of Vision Transformers (ViTs) that suffer from the quadratic complexity. While the recurrent scanning mechanism of Mamba offers computational efficiency, it inherently limits non-causal interactions betwee…

Cited by 0SourceScholar
2025

Beyond RGB: Adaptive Parallel Processing for RAW Object Detection

ICCV 2025poster

Object detection models are typically applied to standard RGB images processed through Image Signal Processing (ISP) pipelines, which are designed to enhance sensor-captured RAW images for human vision. However, these ISP functions can lead to a loss of critical information that may be essential in…

2025

MambaPEFT: Exploring Parameter-Efficient Fine-Tuning for Mamba

ICLR 2025poster

An ecosystem of Transformer-based models has been established by building large models with extensive data. Parameter-efficient fine-tuning (PEFT) is a crucial technology for deploying these models to downstream tasks with minimal cost while achieving effective performance. Recently, Mamba, a State…

2023

DynamicISP: Dynamically Controlled Image Signal Processor for Image Recognition

ICCV 2023poster

Image Signal Processors (ISPs) play important roles in image recognition tasks as well as in the perceptual quality of captured images. In most cases, experts make a lot of effort to manually tune many parameters of ISPs, but the parameters are sub-optimal. In the literature, two types of techniques…

Cited by 21PDFScholar
2023

Rawgment: Noise-Accounted RAW Augmentation Enables Recognition in a Wide Variety of Environments

CVPR 2023poster

Image recognition models that work in challenging environments (e.g., extremely dark, blurry, or high dynamic range conditions) must be useful. However, creating training datasets for such environments is expensive and hard due to the difficulties of data collection and annotation. It is desirable i…

Cited by 22SourcePDFScholar
2021

MBAPose: Mask and Bounding-Box Aware Pose Estimation of Surgical Instruments with Photorealistic Domain Randomization

IROS 2021poster

Surgical robots are usually controlled using a priori models based on the robots’ geometric parameters, which are calibrated before the surgical procedure. One of the challenges in using robots in real surgical settings is that those parameters can change over time, consequently deteriorating contro…

Cited by 7SourceScholar
2020

Single-Shot Pose Estimation of Surgical Robot Instruments’ Shafts from Monocular Endoscopic Images

ICRA 2020poster

Surgical robots are used to perform minimally invasive surgery and alleviate much of the burden imposed on surgeons. Our group has developed a surgical robot to aid in the removal of tumors at the base of the skull via access through the nostrils. To avoid injuring the patients, a collision-avoidanc…

Cited by 23SourceScholar