← Search

Inbal Lavi

4 accepted papers

2024

VisFocus: Prompt-Guided Vision Encoders for OCR-Free Dense Document Understanding

ECCV 2024poster

"In recent years, notable advancements have been made in the domain of visual document understanding, with the prevailing architecture comprising a cascade of vision and language models. The text component can either be extracted explicitly with the use of external OCR models in OCR-based approaches…

2022

GLASS: Global to Local Attention for Scene-Text Spotting

ECCV 2022poster

"In recent years, the dominant paradigm for text spotting is to combine the tasks of text detection and recognition into a single end-to-end framework. Under this paradigm, both tasks are accomplished by operating over a shared global feature map extracted from the input image. Among the main challe…

2022

Towards Weakly-Supervised Text Spotting Using a Multi-Task Transformer

CVPR 2022poster

Text spotting end-to-end methods have recently gained attention in the literature due to the benefits of jointly optimizing the text detection and recognition components. Existing methods usually have a distinct separation between the detection and recognition branches, requiring exact annotations f…

Cited by 77PDFScholar
2020

Can You Read Me Now? Content Aware Rectification using Angle Supervision

ECCV 2020poster

The ubiquity of smartphone cameras has led to more and more documents being captured by cameras rather than scanned. Unlike flatbed scanners, photographed documents are often folded and crumpled, resulting in large local variance in text structure. The problem of document rectification is fundamenta…

Cited by 35SourcePDFScholar