← Search

Nikolaos Livathinos

3 accepted papers

2025

SmolDocling: An ultra-compact vision-language model for end-to-end multi-modal document conversion

ICCV 2025poster

We introduce SmolDocling, an ultra-compact vision-language model targeting end-to-end document conversion. Our model comprehensively processes entire pages by generating DocTags, a new universal markup format that captures all page elements in their full context with location. Unlike existing approa…

2024

ESG Accountability Made Easy: DocQA at Your Service

AAAI 2024technical

We present Deep Search DocQA. This application enables information extraction from documents via a question-answering conversational assistant. The system integrates several technologies from different AI disciplines consisting of document conversion to machine-readable format (via computer vision),…

2022

TableFormer: Table Structure Understanding With Transformers

CVPR 2022poster

Tables organize valuable content in a concise and compact representation. This content is extremely valuable for systems such as search engines, Knowledge Graph's, etc, since they enhance their predictive capabilities. Unfortunately, tables come in a large variety of shapes and sizes. Furthermore, t…

Cited by 91PDFcodeScholar