ComfyUI
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
Written mainly in Python. Released under the GPL-3.0 licence.
What each tool is actually for, described plainly. No ratings, no scores, no invented benchmarks.
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
Written mainly in Python. Released under the GPL-3.0 licence.
GGUF Quantization support for native ComfyUI models
Written mainly in Python. Released under the Apache-2.0 licence.
Various custom nodes for ComfyUI
Written mainly in Python. Released under the GPL-3.0 licence.
ComfyUI-Manager is an extension designed to enhance the usability of ComfyUI. It offers management functions to install, remove, disable, and enable various custom nodes of ComfyUI. Furthermore, this extension provides a hub feature and convenience functions to access a wide range of information within ComfyUI.
Written mainly in Python. Released under the GPL-3.0 licence.
Let us control diffusion models!
Written mainly in Python. Released under the Apache-2.0 licence.
Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.
Written mainly in Python. Released under the Apache-2.0 licence.
Focus on prompting and generating
Written mainly in Python. Released under the GPL-3.0 licence.
GFPGAN aims at developing Practical Algorithms for Real-world Face Restoration.
Written mainly in Python.
Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
Written mainly in Jupyter Notebook. Released under the Apache-2.0 licence.
Hunyuan-DiT : A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Written mainly in Jupyter Notebook.
Open-source Python project
Open-source project on GitHub. Written mainly in Python.
Image inpainting tool powered by SOTA AI Model. Remove any unwanted object, defect, people from your pictures or erase and replace(powered by stable diffusion) any thing on your pictures.
Written mainly in Python. Released under the Apache-2.0 licence. The repository is archived and is no longer maintained.
Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.
Written mainly in Python. Released under the Apache-2.0 licence.
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
Written mainly in Python. Released under the Apache-2.0 licence.
Real-ESRGAN aims at developing Practical Algorithms for General Image/Video Restoration.
Written mainly in Python. Released under the BSD-3-Clause licence.
Official Code for Stable Cascade
Written mainly in Jupyter Notebook. Released under the MIT licence.
Image to prompt with BLIP and CLIP
Written mainly in Python. Released under the MIT licence.
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
Written mainly in Python. Released under the Apache-2.0 licence.
Materials for the Hugging Face Diffusion Models Course
Written mainly in Jupyter Notebook. Released under the Apache-2.0 licence.
Official inference repo for FLUX.1 models
Written mainly in Python. Released under the Apache-2.0 licence.
Generative Models by Stability AI
Written mainly in Python. Released under the MIT licence.
Streamlined interface for generating images with AI in Krita. Inpaint and outpaint with optional text prompt, no tweaking required.
Written mainly in Python. Released under the GPL-3.0 licence.
Open-source Python project
Open-source project on GitHub. Written mainly in Python. The repository is archived and is no longer maintained.
Rembg is a tool to remove images background
Written mainly in Python. Released under the MIT licence.
A custom script for AUTOMATIC1111/stable-diffusion-webui to implement a tiny template language for random prompt generation
Written mainly in Python. Released under the MIT licence.
AnimateDiff for AUTOMATIC1111 Stable Diffusion WebUI
Written mainly in Python.
WebUI extension for ControlNet
Written mainly in Python. Released under the GPL-3.0 licence.
Open-source Python project
Open-source project on GitHub. Written mainly in Python. Released under the MIT licence.
SD.Next: All-in-one WebUI for AI generative image and video creation, captioning and processing
Written mainly in Python. Released under the Apache-2.0 licence.
The repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
Written mainly in Jupyter Notebook. Released under the Apache-2.0 licence.
A latent text-to-image diffusion model
Written mainly in Jupyter Notebook.
Stable Diffusion web UI
Written mainly in Python. Released under the AGPL-3.0 licence.
Open-source Python project
Open-source project on GitHub. Written mainly in Python. Released under the AGPL-3.0 licence.
We write your reusable computer vision tools. 💜
Written mainly in Python. Released under the MIT licence.
OCR, layout analysis, reading order, table recognition in 90+ languages
Written mainly in Python. Released under the Apache-2.0 licence.
Tesseract Open Source OCR Engine (main repository)
Written mainly in C++. Released under the Apache-2.0 licence.
State-of-the-art Machine Learning for the web. Run 🤗 Transformers directly in your browser, with no need for a server!
Written mainly in JavaScript. Released under the Apache-2.0 licence.
Ultralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
Written mainly in Python. Released under the AGPL-3.0 licence.