Dots OCR

Dots OCR

Recognize Any Human Scripts and Symbols

8,875GitHub
❤️1
Description

A 3B-parameter multimodal model composed of a 1.2B vision encoder and a 1.7B language model. Designed for universal accessibility, it possesses the capability to recognize virtually any human script. Beyond achieving state-of-the-art (SOTA) performance in standard multilingual document parsing among models of comparable size, dots.ocr-1.5 excels at converting structured graphics (e.g., charts and diagrams) directly into SVG code, parsing web screens and spotting scene text. Furthermore, the model demonstrates competitive performance in general OCR, object grounding & counting tasks.

Screenshots
Screenshot 1
Screenshot 2
Screenshot 3
Mobile Screenshots
Mobile Screenshot 1
Mobile Screenshot 2
Mobile Screenshot 3
App Information
Version
1.0.0
Package Size
764.05 KB
Image Size
49.53 MB
Updated
March 12, 2026
Source Code
rednote-hilab
Platform Support
PCMobile
Keywords
dotsocrmodelapppdfprompt