Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
BarzanAI 's Collections
Vision Models
Text2Speech
ASR
Guardrails Models
Embedding Models for all Modalities
Models for ML Tasks
Tool-using LLMs
Coversational Models
OCR
Text2SQL

OCR

updated 2 days ago

A bunch of VL Models/OCR Engines for OCR Tasks

Upvote
-

  • PaddlePaddle/PaddleOCR-VL-1.6

    Image-Text-to-Text • 1.0B • Updated Jun 5 • 64.4k • 391

  • dots-studio/dots.mocr

    Image-Text-to-Text • 3B • Updated about 1 month ago • 209k • 156

    Note Best performing OCR Engine


  • dots-studio/dots.ocr

    Image-Text-to-Text • 3B • Updated Oct 31, 2025 • 384k • 1.32k

  • google/gemma-4-12B-it

    Any-to-Any • 12B • Updated 15 days ago • 3.03M • 1.39k

    Note Best performing Vision-Language Model. could be used with a OCR Engine as a judge.

Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs