release aidev NEEDS REVIEW SIG 1/5

DeepSeek OCR-2: vision-language OCR model on Hugging Face

DeepSeek published DeepSeek-OCR-2, a dedicated image-text-to-text OCR model on Hugging Face under the Apache-2.0 license. The repository uses custom transformer code and carries tags for multilingual support, vision-language, and feature extraction. Two associated arXiv papers are referenced.

PUBLISHED2026-01-27
OBSERVED2026-08-11
AGE7mo
SOURCES1
  • Pipeline: image-text-to-text (vision-language OCR)
  • Library: transformers with custom_code
  • Tags: multilingual, vision-language, ocr, feature-extraction, deepseek_vl_v2
  • License: Apache-2.0
  • Referenced papers: arxiv:2601.20552, arxiv:2510.18234
  • Repository created 2026-01-27, last modified 2026-02-03

COMMUNITY

No curated reactions recorded for this event. Facts and takes are kept in separate layers — community context is added by hand, never blended into the record above.