Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
JBrightmanAI 's Collections
OCR
Benchmarking

OCR

updated Jul 18

State-of-the-art OCR models, datasets, and demos for document parsing.

Upvote
-

  • LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget

    Paper • 2607.14952 • Published Jul 16 • 212

  • VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding

    Paper • 2607.14935 • Published Jul 16 • 172

  • From Pixels to States: Rethinking Interactive World Models as Game Engines

    Paper • 2607.14076 • Published Jul 15 • 37

  • baidu/Unlimited-OCR

    Image-Text-to-Text • 3B • Updated Jul 29 • 3.18M • 4.17k

  • zai-org/GLM-OCR

    Image-Text-to-Text • 1B • Updated May 19 • 2.1M • • 2.01k

  • nvidia/nemotron-ocr-v2

    Image-to-Text • Updated May 22 • 1.45k • 247

  • uv-scripts/ocr

    Updated 4 days ago • 2.62k • 155

  • bevaya/pubmed-ocr

    Viewer • Updated Jan 22 • 1.55M • 4.08k • 72

  • unsloth/LaTeX_OCR

    Viewer • Updated Nov 21, 2024 • 76.3k • 2.98k • 91
Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs