catalog / Vision & Image / PaddleOCR
PipelineVision & ImageFree

PaddleOCR

Industrial-grade multilingual OCR toolkit supporting 80+ languages with text detection, recognition, and layout analysis.

Instalações375k
⟳ upstream v3.7.0 · updated 1mo ago
Repositório fonte
! Grade B · 75/100 · ReviewSecurity assessment
No compromise signals46capabilities surfaced1known CVE8of 20 OWASP controls clear
External endpoints declaredExternal endpoints declaredSuspicious network referencesExternal endpoints declared
scanned 18d agoosv · gitleaks · opengrep · picklescan + heuristicsfull breakdown in the Security tab ↓

PaddleOCR

PaddleOCR is PaddlePaddle's comprehensive OCR toolkit covering the complete pipeline from text detection to recognition and structured layout analysis. It supports 80+ languages, handles rotated and curved text, and ships production-ready server and edge-deploy builds.

Key Features

  • End-to-end pipeline: DB text detection → SVTR/PP-OCRv4 recognition → PP-Structure layout recovery in one call
  • 80+ languages: Latin, Chinese, Japanese, Korean, Arabic (RTL), Devanagari, and more
  • PP-Structure: table recognition, form extraction, and document layout analysis — outputs structured Excel/HTML
  • PP-ChatOCR: multimodal RAG pipeline that answers questions about document images
  • Model compression: INT8 quantisation and pruning for edge deployment (Raspberry Pi, phones)
  • Benchmark: PP-OCRv4 tops open-source OCR leaderboards on ICDAR and scene-text datasets

Quick Start

pip install paddlepaddle paddleocr
from paddleocr import PaddleOCR

ocr = PaddleOCR(use_angle_cls=True, lang='en')
result = ocr.ocr('invoice.jpg', cls=True)
for line in result[0]:
    print(line[1][0])  # recognised text
npx ai-supply add paddleocr-multilingual-text-recognition

Curated mirror of the open-source PaddleOCR (Apache-2.0). Get it from the source.

More from @ai-supply

View profile →
Agent
MetaGPT
Multi-agent framework that assigns GPT roles (PM, engineer, QA) to solve complex software tasks end-to-end.
1.0M
Connector
vLLM
High-throughput, memory-efficient LLM inference engine with PagedAttention and continuous batching.
892k
Connector
Meilisearch
Lightning-fast open-source search engine with typo-tolerance, semantic hybrid search, and sub-50ms response times.
811k
Eval
Weights & Biases (wandb)
ML experiment tracking and visualization — log metrics, hyperparameters, models, and media in real time.
784k