Modern OCR in Resource-Constrained Environments

Infos
- WannDo., 17. Sept. 14:00 - 14:30 Uhr
- WoOnline-EventLink sichtbar nach Anmeldung.
- FormatVortrag
- ExpertiseInteressierte
- kostenlos
Optical Character Recognition (OCR) is crucial for automating scanned document processing, yet designing high-quality systems under resource constraints remains challenging. In this talk, we share insights from building OCR solutions for real-world applications. We trace the evolution of OCR pipelines from traditional character recognition approaches, which rely on dictionaries and rule-based post-processing, to modern solutions leveraging open-source frameworks, vision models, and LLMs for contextual correction, thereby enhancing extraction quality and robustness. Through concrete examples, we compare different approaches, discussing their strengths, limitations, and performance in practical scenarios. We present engineering trade-offs required to maintain efficiency, reliability, and maintainability under these conditions. Furthermore, we explore methods for combining visual and textual signals to extract structured information while ensuring a consistent and reliable user experience
Speaker:innen
Olmo BarberisOlmo Barberis is a product manager and software engineer at Karakun AG in Basel. He has experience in automated document processing and semantic search engine development, specializing in AI technologies. Olmo holds an MSc in Engineering from SUPSI with a focus on natural language processing and human-computer interaction. Previously, he was a research assistant at SUPSI, working on AI systems for document classification and information extraction.
