The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
github.com/Comfy-Org/ComfyUIWhat each tool is actually for, described plainly. No ratings, no scores, no invented benchmarks.
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
github.com/Comfy-Org/ComfyUIGGUF Quantization support for native ComfyUI models
github.com/city96/ComfyUI-GGUFVarious custom nodes for ComfyUI
github.com/kijai/ComfyUI-KJNodesComfyUI-Manager is an extension designed to enhance the usability of ComfyUI. It offers management functions to install, remove, disable, and enable various custom nodes of ComfyUI. Furthermore, this extension provides a hub feature and convenience functions to access a wide range of information within ComfyUI.
github.com/Comfy-Org/ComfyUI-ManagerLet us control diffusion models!
github.com/lllyasviel/ControlNetReady-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.
github.com/JaidedAI/EasyOCRFocus on prompting and generating
github.com/lllyasviel/FooocusGFPGAN aims at developing Practical Algorithms for Real-world Face Restoration.
github.com/TencentARC/GFPGANGrounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything
github.com/IDEA-Research/Grounded-Segment-AnythingHunyuan-DiT : A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
github.com/Tencent-Hunyuan/HunyuanDiTImage inpainting tool powered by SOTA AI Model. Remove any unwanted object, defect, people from your pictures or erase and replace(powered by stable diffusion) any thing on your pictures.
github.com/Sanster/IOPaintInvoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.
github.com/invoke-ai/InvokeAITurn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
github.com/PaddlePaddle/PaddleOCRReal-ESRGAN aims at developing Practical Algorithms for General Image/Video Restoration.
github.com/xinntao/Real-ESRGANOfficial Code for Stable Cascade
github.com/Stability-AI/StableCascadeImage to prompt with BLIP and CLIP
github.com/pharmapsychotic/clip-interrogator🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
github.com/huggingface/diffusersMaterials for the Hugging Face Diffusion Models Course
github.com/huggingface/diffusion-models-classOfficial inference repo for FLUX.1 models
github.com/black-forest-labs/fluxGenerative Models by Stability AI
github.com/Stability-AI/generative-modelsStreamlined interface for generating images with AI in Krita. Inpaint and outpaint with optional text prompt, no tweaking required.
github.com/Acly/krita-ai-diffusionOpen-source Python project
github.com/krea-ai/open-promptsRembg is a tool to remove images background
github.com/danielgatis/rembgA custom script for AUTOMATIC1111/stable-diffusion-webui to implement a tiny template language for random prompt generation
github.com/adieyal/sd-dynamic-promptsAnimateDiff for AUTOMATIC1111 Stable Diffusion WebUI
github.com/continue-revolution/sd-webui-animatediffWebUI extension for ControlNet
github.com/Mikubill/sd-webui-controlnetSD.Next: All-in-one WebUI for AI generative image and video creation, captioning and processing
github.com/vladmandic/sdnextThe repository provides code for running inference with the SegmentAnything Model (SAM), links for downloading the trained model checkpoints, and example notebooks that show how to use the model.
github.com/facebookresearch/segment-anythingA latent text-to-image diffusion model
github.com/CompVis/stable-diffusionStable Diffusion web UI
github.com/AUTOMATIC1111/stable-diffusion-webuiOpen-source Python project
github.com/lllyasviel/stable-diffusion-webui-forgeWe write your reusable computer vision tools. 💜
github.com/roboflow/supervisionOCR, layout analysis, reading order, table recognition in 90+ languages
github.com/datalab-to/suryaTesseract Open Source OCR Engine (main repository)
github.com/tesseract-ocr/tesseractState-of-the-art Machine Learning for the web. Run 🤗 Transformers directly in your browser, with no need for a server!
github.com/huggingface/transformers.jsUltralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
github.com/ultralytics/ultralytics