davanstrien maps OCR models across four Hub sub-collections
TL;DR
- davanstrien's OCR on the Hub groups models into four sub-collections covering documents, languages and scripts, handwriting and archives, and text recognition pipelines.
- The languages sub-collection covers Thai, Japanese, Vietnamese, Arabic, Korean and Devanagari; handwriting spans Swedish, Norwegian, German Kurrent, Tibetan, and Hebrew-script manuscripts.
- The text recognition sub-collection features Kraken and PaddleOCR pipelines for documents, manga and text in photographs.
davanstrien's OCR on the Hub collection groups OCR models on Hugging Face into four sub-collections: 17 for documents, 10 for languages and scripts, 10 for handwriting and archives, and 6 for text recognition and pipelines. Two experts in our Who's Who directory shared the collection.
The framing is deliberately practical. 'Curated OCR models for documents, languages, handwriting and text in images,' the overview reads. 'Browse four collections with short practical notes.'
The languages list covers Thai, Japanese, Vietnamese, Arabic, Korean and Devanagari. The handwriting list spans Swedish, Norwegian, German Kurrent, Tibetan and Hebrew-script manuscripts. The pipelines section features Kraken and PaddleOCR for documents, manga and text in photographs. Contributions came from stefan-it, maehr and sharonshorg.
Each sub-collection ships with a caveat. The documents note warns that 'some models need a full parsing pipeline.' The handwriting note is blunter: 'Many models need cropped lines.' The languages note tells users to 'Check each model's expected input and domain.'
Shared on Bluesky by 2 AI experts
-
OCR for Japanese manga, Swedish handwriting or Arabic print? There’s a growing range of OCR models on @hf.co: VLMs, dedicated text recognisers and complete OCR pipelines. I’ve gathered 41 models into four collections, …
View on Bluesky →
Originally reported by huggingface.co
Read the original article →Original headline: OCR on the Hub - a davanstrien Collection