Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.
github.com/JaidedAI/EasyOCRWhat each tool is actually for, described plainly. No ratings, no scores, no invented benchmarks.
Vision and OCR: Reading images rather than making them: text extraction, document parsing, segmentation and object detection.
Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.
github.com/JaidedAI/EasyOCRGrounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
github.com/IDEA-Research/Grounded-Segment-AnythingTurn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
github.com/PaddlePaddle/PaddleOCRThe repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
github.com/facebookresearch/segment-anythingWe write your reusable computer vision tools. 💜
github.com/roboflow/supervisionOCR, layout analysis, reading order, table recognition in 90+ languages
github.com/datalab-to/suryaTesseract Open Source OCR Engine (main repository)
github.com/tesseract-ocr/tesseractState-of-the-art Machine Learning for the web. Run 🤗 Transformers directly in your browser, with no need for a server!
github.com/huggingface/transformers.jsUltralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
github.com/ultralytics/ultralytics