Image-to-Text
MLX
Safetensors
mlx-weights
paddlepaddle-ocr
ppocrv5
ppocrv6
ppdoclayoutv3
pp-structure
apple-silicon
Instructions to use plaincompute/ppocr-mlx with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use plaincompute/ppocr-mlx with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] hf download plaincompute/ppocr-mlx --local-dir ppocr-mlx
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
|
Download uvdoc/README.md from plaincompute/ppocr-mlx: direct link, hf CLI and curl.
- Browser
- Download file 1.22 kB
-
https://huggingface.co/plaincompute/ppocr-mlx/resolve/main/uvdoc/README.md
- Command line
-
hf download hf://plaincompute/ppocr-mlx/uvdoc/README.md
-
curl -L -o README.md https://huggingface.co/plaincompute/ppocr-mlx/resolve/main/uvdoc/README.md
1.22 kB
metadata
license: apache-2.0
library_name: PaddleOCR
language:
- en
- zh
pipeline_tag: image-to-text
tags:
- OCR
- PaddlePaddle
- PaddleOCR
- doc_img_unwarping
UVDoc
Introduction
The main purpose of text image correction is to carry out geometric transformation on the image to correct the document distortion, inclination, perspective deformation and other problems in the image, so that the subsequent text recognition can be more accurate.
| Model | CER |
|---|---|
| UVDoc | 0.179 |
Note: Test data set: docunet benchmark data set.
Model Usage
import requests
from PIL import Image
from transformers import AutoImageProcessor, AutoModel
model_path = "PaddlePaddle/UVDoc_safetensors"
model = AutoModel.from_pretrained(model_path, device_map="auto")
image_processor = AutoImageProcessor.from_pretrained(model_path)
image = Image.open(requests.get("https://paddle-model-ecology.bj.bcebos.com/paddlex/imgs/demo_image/doc_test.jpg", stream=True).raw)
inputs = image_processor(images=image, return_tensors="pt").to(model.device)
outputs = model(**inputs)
result = image_processor.post_process_document_rectification(outputs.last_hidden_state, inputs["original_images"])
print(result)