Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
JBrightmanAI 's Collections
OCR
Benchmarking

OCR

updated Jul 18

State-of-the-art OCR models, datasets, and demos for document parsing.

Upvote
-

  • LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget

    Paper • 2607.14952 • Published Jul 16 • 90

  • VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding

    Paper • 2607.14935 • Published Jul 16 • 119

  • From Pixels to States: Rethinking Interactive World Models as Game Engines

    Paper • 2607.14076 • Published Jul 15 • 31

  • baidu/Unlimited-OCR

    Image-Text-to-Text • 3B • Updated Jul 29 • 2.11M • 4.28k

  • zai-org/GLM-OCR

    Image-to-Text • 1B • Updated 12 days ago • 1.68M • 2.07k

  • nvidia/nemotron-ocr-v2

    Image-to-Text • Updated May 22 • 1.41k • 258

  • uv-scripts/ocr

    Updated 6 days ago • 3.78k • 163

  • bevaya/pubmed-ocr

    Viewer • Updated Jan 22 • 1.55M • 3.38k • 72

  • unsloth/LaTeX_OCR

    Viewer • Updated Nov 21, 2024 • 76.3k • 2.75k • 92
Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs