← Search

Harsh Singh

4 accepted papers

2025

All Languages Matter: Evaluating LMMs on Culturally Diverse 100 Languages

CVPR 2025highlight

Existing Large Multimodal Models (LMMs) generally focus on only a few regions and languages. As LMMs continue to improve, it is increasingly important to ensure they understand cultural contexts, respect local sensitivities, and support low-resource languages, all while effectively integrating corr…

2025

MALMM: Multi-Agent Large Language Models for Zero-Shot Robotic Manipulation

IROS 2025

Large Language Models (LLMs) have demonstrated remarkable planning abilities across various domains, including robotic manipulation and navigation. While recent work in robotics deploys LLMs for high-level and low-level planning, existing methods often face challenges with failure recovery and suffe

Cited by 22SourceScholar
2025

Self-Improvement in Multimodal Large Language Models: A Survey

EMNLP 2025

Recent advancements in self-improvement for Large Language Models (LLMs) have efficiently enhanced model capabilities without significantly increasing costs, particularly in terms of human effort. While this area is still relatively young, its extension to the multimodal domain holds immense potenti

2022

Data Efficient Support Vector Machine Training Using the Minimum Description Length Principle

ICASSP 2022accepted

Support vector machines (SVMs) are established as highly successful classifiers in a broad range of applications, including numerous medical ones. Nevertheless, their current employment is restricted by a limitation in the manner in which they are trained, most often the training-validation-test or…

Cited by 0SourceScholar