Computer Vision

tesseract

tesseract-ocr/tesseract

An open-source OCR engine for converting text in images into machine-readable text.

★ 76.3KStars
⑂ 10.8KForks
483Open issues
C++Language
Apache-2.0License
Q90Editorial score

Overview

An open-source OCR engine for converting text in images into machine-readable text. The repository is maintained under tesseract-ocr on GitHub. Its primary language is C++. See the [project README](https://github.com/tesseract-ocr/tesseract#readme) for the supported workflows.

Key features

  • An open-source OCR engine for converting text in images into machine-readable text.

Requirements, installation and quick start

Use the repository installation guide to select the runtime and any required model weights. Verify operating-system and GPU compatibility before starting.

[Read the upstream installation and quickstart instructions](https://github.com/tesseract-ocr/tesseract#readme).

Usage

Try the documented image or video example with your own test material, inspect the results, then integrate the relevant API or processing workflow.

[Usage examples and configuration reference](https://github.com/tesseract-ocr/tesseract#readme).

Model compatibility and use cases

Model compatibility is not stated in the repository metadata.

License and risk notes

GitHub reports Apache-2.0 for this repository. Review the upstream license file; model weights, datasets and dependencies may have separate terms.

Source review 2026-09-05: GitHub search metadata and repository README. No runtime benchmark performed. License metadata: Apache-2.0.

Release and maintenance

[View upstream releases](https://github.com/tesseract-ocr/tesseract/releases).

opencv

opencv/opencv

★ 90.7KC++

yolov5

ultralytics/yolov5

★ 58KPython

faceswap

deepfakes/faceswap

★ 57.5KPython

PaddleClas

PaddlePaddle/PaddleClas

★ 5.8KPython