DeepSeek OCR-2: vision-language OCR model on Hugging Face
DeepSeek published DeepSeek-OCR-2, a dedicated image-text-to-text OCR model on Hugging Face under the Apache-2.0 license. The repository uses custom transformer code and carries tags for multilingual support, vision-language, and feature extraction. Two associated arXiv papers are referenced.
PUBLISHED2026-01-27
OBSERVED2026-08-11
AGE7mo
SOURCES1
- Pipeline: image-text-to-text (vision-language OCR)
- Library: transformers with custom_code
- Tags: multilingual, vision-language, ocr, feature-extraction, deepseek_vl_v2
- License: Apache-2.0
- Referenced papers: arxiv:2601.20552, arxiv:2510.18234
- Repository created 2026-01-27, last modified 2026-02-03
COMMUNITY
No curated reactions recorded for this event. Facts and takes are kept in separate layers — community context is added by hand, never blended into the record above.