Stable Diffusion Webui
Stable Diffusion web UI
Stable Diffusion web UI
The most powerful and modular diffusion model GUI, …
画质出众的文生图工具
字节出品的中文文生图
阿里文生图与图像编辑
开源文生图基座模型
高质量开源文生图模型
文字渲染准确的文生图
Adobe 出品的商用安全图库
节点式图像生成工作流
百度 AI 图像生成,支持角色与场景设定
🤗 Diffusers: State-of-the-art diffusion models for…
Invoke is a leading creative engine for Stable Diff…
Image-to-Image Translation in PyTorch
Image inpainting tool powered by SOTA AI Model. Rem…
GPT-Image-2 API and Prompts
stable diffusion webui colab
Diffusion Bee is the easiest way to run Stable Diff…
Software that can generate photos from paintings, …
A collection of resources and papers on Diffusion M…
AI绘画资料合集(包含国内外可使用平台、使用教程、参数教程、部署教程、业界新闻等等) Stable d…
Implementation of DALL-E 2, OpenAI's updated text-t…
OpenVINO™ is an open source toolkit for optimizing …
Convert AI papers to GUI,Make it easy and convenien…
Image-to-image translation with conditional adversa…
Production ready toolkit to run AI locally
Community interface for generative AI
Text-to-3D & Image-to-3D & Mesh Exportation with Ne…
[NeurIPS 2024 Best Paper Award][GPT beats diffusion…
Multi-Platform Package Manager for Stable Diffusion
Stable Diffusion built-in to Blender
Awesome curated collection of images and prompts ge…
Run Stable Diffusion on Mac natively
fast-stable-diffusion + DreamBooth
Implementation of Dreambooth (https://arxiv.org/abs…
Using Low-rank adaptation to quickly fine-tune diff…
OpenMMLab Multimodal Advanced, Generative, and Inte…
SD.Next: All-in-one WebUI for AI generative image a…
Easy Docker setup for Stable Diffusion with user-fr…
A repository of models, textual inversions, and more
An APP that integrates mainstream large language mo…
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...)…
🌻 一键拥有你自己的 ChatGPT+众多AI 网页服务 | One click access to…
This repository is a curated collection of links to…
notes for software engineers getting up to speed on…
SD-Trainer. LoRA & Dreambooth training scripts & GU…
SUPIR aims at developing Practical Algorithms for P…
Implementation / replication of DALL-E, OpenAI's Te…
中文在线图片编辑与抠图
🦙 LaMa Image Inpainting, Resolution-robust Large …
一键去除图片背景
Ultralytics YOLO26, YOLO11, YOLOv8 — object detecti…
Ultralytics YOLOv5 in PyTorch for object detection,…
We write your reusable computer vision tools. 💜
Label Studio is a multi-type data labeling and anno…
CVPR 2026 论文和开源项目合集
Computer Vision Annotation Tool (CVAT) is a leading…
Image annotation with Python. Supports polygon, rec…
Advanced AI Explainability for computer vision. Su…
cvpr2024/cvpr2023/cvpr2022/cvpr2021/cvpr2020/cvpr20…
Semantic segmentation models with 500+ pretrained c…
Tutorials, assignments, and competitions for MIT De…
X-AnyLabeling: A lightweight, efficient, and unifie…
[CVPR 2024 Oral] InternVL Family: A Pioneering Open…
The code for our newly accepted paper in Pattern Re…
A collection of tutorials on state-of-the-art compu…
Robust Video Matting in PyTorch, TensorFlow, Tensor…
RF-DETR is a real-time object detection and segment…
Hello AI World guide to deploying deep-learning inf…
Stanford NLP Python library for tokenization, sente…
Real-Time High-Resolution Background Matting
Gluon CV Toolkit
图片无损放大与降噪
a cross-platform image super-resolution tool
阿里出品的电商设计工具
Original reference implementation of "3D Gaussian S…
Instant neural graphics primitives: lightning fast …
A collaboration friendly studio for NeRFs
Desktop app to generate 3D models from images or pr…
open Multiple View Geometry library. Basis for 3D c…
Standardized Distributed Generative and Predictive …
Red Ink - A one-stop Xiaohongshu image-and-text gen…
Framework agnostic sliced/tiled inference + interac…
Remove visible and invisible AI watermarks and prov…
A PyTorch Library for Accelerating 3D Deep Learning…
Clarity AI | AI Image Upscaler & Enhancer - free an…
深度学习辅助漫画翻译工具, 支持一键机翻和简单的图像/文本编辑 | Yet another compu…
StableSwarmUI, A Modular Stable Diffusion Web-User-…
🔎 Super-scale your images and run experiments with…
Create 🔥 videos with Stable Diffusion by exploring…
AI 助手全套开源解决方案,自带运营管理后台,开箱即用。集成了 ChatGPT, Azure, Cha…
Official implementation of "Neuralangelo: High-Fide…
Webots Robot Simulator
《Pytorch实用教程》(第二版)无论是零基础入门,还是CV、NLP、LLM项目应用,或是进阶工程化…
[ECCV 2022] This is the official implementation of …
Curated tutorials and resources for Large Language …
Machine learning, computer vision, statistics and g…
SwarmUI (formerly StableSwarmUI), A Modular Stable …
:man: Code for "Large Pose 3D Face Reconstruction …
A lite C++ AI toolkit: 100+ models with MNN, ORT an…
A curated list of papers & resources linked to 3D r…
Simple command line tool for text to image generati…
A Python package for segmenting geospatial data wit…
3D ResNets for Action Recognition (CVPR 2018)
🛰️ List of satellite image training datasets with …
[CVPR 2024] 4D Gaussian Splatting for Real-Time Dyn…
Stable diffusion for real-time music generation
Outpainting with Stable Diffusion on an infinite ca…
🪩 Create Disco Diffusion artworks in one line
A flexible, high-performance 3D simulator for Embod…
[CVPR 2026] PromptEnhancer is a prompt-rewriting to…
Paper list and datasets for industrial image anomal…
Bringing stable diffusion models to web browsers. E…
The PyTorch improved version of TPAMI 2017 paper: F…
Train, inspect, edit, automate, and export 3D Gauss…
roop extension for StableDiffusion web-ui
[CVPR19/TPAMI23] SiamMask: A Framework for Fast Onl…
min(DALL·E) is a fast, minimal port of DALL·E Mini …
3D Computer Vision Framework
✨ Reverse-engineered Python API for Google Gemini w…
Making ComfyUI more comfortable!
Run Stable Diffusion on Android Devices with Snapdr…
Awesome work on hand pose estimation/tracking
Accelerated deep learning R&D
Implementation of Dreambooth (https://arxiv.org/abs…
[CVPR 2025] MASt3R-SLAM: Real-Time Dense SLAM with …
The official PyTorch implementation of Towards Fast…
OpenMMLab Model Deployment Framework
Zero-1-to-3: Zero-shot One Image to 3D Object (ICCV…
Enhances Tesseract OCR output using LLMs (local or …
Minkowski Engine is an auto-diff neural network lib…
A cross-platform video structuring (video analysis)…
AI comic and manga translator app/browser extension…
A general fine-tuning kit geared toward image/video…
ICCV 2025 论文和开源项目合集
Implementation of papers in 100 lines of code.
Kandinsky 2 — multilingual text2image latent diffus…
Images to inference with no labeling (use foundatio…
A playground to generate images from any text promp…
🔥 [ICCV 2025 Highlight] InfiniteYou: Flexible Phot…
dLLM: Simple Diffusion Language Modeling
Stable diffusion for real-time music generation (we…
Project Page for "LISA: Reasoning Segmentation via …
[IJCV2024] Exploiting Diffusion Prior for Real-Worl…
Just playing with getting VQGAN+CLIP running locall…
Automatically remove the mosaics in images and vide…
A simple command line tool for text to image genera…
This is the repo for our new project Highly Accurat…
Contrastive unpaired image-to-image translation, fa…
Semantic Segmentation Suite in TensorFlow. Implemen…
Lora beYond Conventional methods, Other Rank adapta…
A collection of awesome resources in Human Pose est…
One-step image-to-image with Stable Diffusion turbo…
A ROS/ROS2 Multi-robot Simulator for Autonomous Veh…
C++ implementation of Lie Groups using Eigen.
(ෆ`꒳´ෆ) A Survey on Text-to-Image Generation/Synthe…
SDK for interacting with stability.ai APIs (e.g. st…
Ray tracing and hybrid rasterization of Gaussian pa…
The collection of pre-trained, state-of-the-art AI …
Translate full-length books and documents with Olla…
Objectron is a dataset of short, object-centric vid…
Comflowyspace is an intuitive, user-friendly, open-…
ScanNet/ScanNet — AI 开源项目
[ICLR 2025] From anything to mesh like human artist…
ICCV2021/2019/2017 论文/代码/解读/直播合集,极市团队整理
:pencil2: Web-based image segmentation tool for obj…
ECCV 2026 论文和开源项目合集,同时欢迎各位大佬提交issue,分享ECCV 2026论文和开…
An open source library and framework for deep learn…
A curated list of the latest breakthroughs in AI by…
collection of diffusion model papers categorized by…
SplaTAM: Splat, Track & Map 3D Gaussians for Dense …
Labeling tool with SAM(segment anything model),supp…
[CVPR'24 Highlight & Best Demo Award] Gaussian Spla…
Your Automatic Prompt Engineering Assistant for Gen…
A sketch extractor for anime/illustration.
Index repo for Kimera code
A ComfyUI custom node designed for advanced image b…
🌍 WorldGen - Generate Any 3D Scene in Seconds
Lightweight inference library for ONNX files, writt…
Collaborate & label any type of data, images, text,…
A reading list for large models safety, security, a…
A list of StableDiffusion styles and some notes for…
Let us democratise high-resolution generation! (CVP…
CVNets: A library for training computer vision netw…
[ECCV 2022] XMem: Long-Term Video Object Segmentati…
A lightweight 3D Morphable Face Model library in mo…
Custom Diffusion: Multi-Concept Customization of Te…
Supercharged experience for multiple models such as…
How do we integrate AI generation tools into actual…
Autoregressive Model Beats Diffusion: 🦙 Llama for …
AI magics meet Infinite draw board.
:unlock: Lip Reading - Cross Audio-Visual Recogniti…
NCRF++, a Neural Sequence Labeling Toolkit. Easy us…
[ICCV 2023] Make-It-3D: High-Fidelity 3D Creation f…
🚀 Easier & Faster YOLO Deployment Toolkit for NVID…
Official code for "DPM-Solver: A Fast ODE Solver fo…
High Resolution Depth Maps for Stable Diffusion Web…
(IROS 2020, ECCVW 2020) Official Python Implementat…
This is a resouce list for low light image enhancem…
[ICML 2024] Mastering Text-to-Image Diffusion: Reca…
:art: Semantic segmentation models, datasets and lo…
A small C++11 header-only library for Lie theory.
Text-to-Image generation. The repo for NeurIPS 2021…
Collecting awesome papers of RAG for AIGC. We prop…
PyTorch Implementation of Fully Convolutional Netwo…
AIdea 是一款支持 GPT 以及国产大语言模型通义千问、文心一言等,支持 Stable Diff…
Laser Odometry and Mapping (Loam) is a realtime met…
[ECCV 2024] The official implementation of paper "B…
A Collection of Papers and Codes for CVPR2026/CVPR2…
Software and pre-trained models for automatic photo…
Declarative way to run AI models in React Native on…
Create characters in Unity with LLMs!
Tracking and collecting papers/projects/others rela…
Must-have resource for anyone who wants to experime…
An extensible, easy-to-use, and portable diffusion …
All-in-one training for vision models (YOLO, ViTs, …
:metal: TT-NN operator library, and TT-Metalium low…
Generate images from texts. In Russian
🐳Dockerfile for 🎨ComfyUI. | 容器镜像与启动脚本
Segment-Anything + 3D. Let's lift anything to 3D.
UniFace: A Unified Face Analysis Library for Python…
A compendium of informations regarding Stable Diffu…
Amica is an open source interface for interactive c…
[CVPR 2025 Oral]Infinity ∞ : Scaling Bitwise AutoRe…
Bundler Structure from Motion Toolkit
A comprehensive list of Implicit Representations an…
🔥RandLA-Net in Tensorflow (CVPR 2020, Oral & IEEE …
The official code repository for the second edition…
ComfyUI nodes for the Ultimate Stable Diffusion Ups…
A crowdsourced distributed cluster for AI art and t…
《深度学习与计算机视觉》配套代码
An open-source project for Windows developers to le…
A simple baseline for 3d human pose estimation in t…
[CVPR 2024] GaussianEditor: Swift and Controllable …
Unofficial implementation of InstantID for ComfyUI
A toolkit for time series machine learning and dee…
Python code to fuse multiple RGB-D images into a TS…
[ECCV 2022] SimpleRecon: 3D Reconstruction Without …
3D Procedural Game Engine Using OpenGL
收集 CVPR 最新的成果,包括论文、代码和demo视频等,欢迎大家推荐!Collect the la…
[CVPR 2021] Multi-Stage Progressive Image Restorati…
It's not AI that takes away your job, but the peopl…
A large-scale text-to-image prompt gallery dataset …
Application implementation with business use cases …
GeneticSharp is a fast, extensible, multi-platform …
Using Diffusion Models to Segment/Reconstruct Organ…
[ICCV 2025] 🔥🔥 UNO: A Universal Customization Me…
Turn any face into a video game character, pixel ar…
Scalable and memory-optimized training of diffusion…
🪐 Objaverse-XL is a Universe of 10M+ 3D Objects. C…
Tag manager and captioner for image datasets
A simple standalone viewer for reading prompts from…
Unofficial implementation of "Prompt-to-Prompt Imag…
Auto1111 extension implementing text2video diffusio…
Build, train, optimize, and run computer vision mod…
A clean and readable Pytorch implementation of Cycl…
The TypeScript library for building AI applications.
Research of DeepSeek Engram Architecture based on Q…
Custom nodes for SDXL and SD1.5 including Multi-Con…
Inpaint Anything extension performs stable diffusio…
An awesome face technology repository.
Pruna is a model optimization framework built for d…
Interactively explore unstructured datasets from yo…
Paper reading notes on Deep Learning and Machine Le…
Examples of programs built using Modal
This codebase demonstrates how to synthesize realis…
Stable Diffusion AI client app for Android and iOS
[CVPR2024, Highlight] Official code for DragDiffusi…
Python SDK for local and cloud background removal (…
An out-of-box human parsing representation extracto…
Paint by Example: Exemplar-based Image Editing with…
[ECCV`24&ICLR`25] CityGaussian Series for High-qual…
Universal Monocular Metric Depth Estimation
An open-source impl. of Large Reconstruction Models
Danbooru / NovelAI 标签超市
Collection of awesome resources on image-to-image t…
:robot: PaddleViT: State-of-the-art Visual Transfor…
Xtreme1 is an all-in-one data labeling and annotati…
Nodes for better inpainting with ComfyUI: Fooocus i…
⚡️The spatial perception framework for rapidly buil…
Stable Diffusion in Blender
GCNet: Non-local Networks Meet Squeeze-Excitation N…
Official PyTorch implementation of VoxFormer [CVPR …
Automatically find issues in image datasets and pra…
Unleash endless possibilities with ComfyUI and Stab…
The simplest way to serve AI/ML models in production
Fast computer vision library for SFM, calibration, …
Pytorch Repo for DeepGCNs (ICCV'2019 Oral, TPAMI'20…
Gaussian-SLAM: Photo-realistic Dense SLAM with Gaus…
PyTorch - FID calculation with proper image resizin…
Using a Model to generate prompts for Model applica…
a self-hosted webui for 30+ generative ai
Video classification tools using 3D ResNet
An application tool of edge-connect, which can do a…
Based on GroundingDino and SAM, use semantic string…
A collection of resources on controllable generatio…
[ECCV 2024] PowerPaint, a versatile image inpaintin…
CogView4, CogView3-Plus and CogView3(ECCV 2024)
👾 A Python API wrapper for Poe.com. With this, you…
Text2Room generates textured 3D meshes from a given…
[CVPR 2022 Oral] Official repository for "MAXIM: Mu…
Stable Diffusion implemented from scratch in PyTorch
Visit PixelLib's official documentation https://pi…
Face Editor for Stable Diffusion
Stable Diffusion in NCNN with c++, supported txt2im…
Official Pytorch Implementation for "MultiDiffusion…
ComfyUI docker images for use in GPU cloud and loca…
EntitySeg Toolbox: Towards Open-World and High-Qual…
半个神器👉一键文本转视频的工具
Reverse-engineered the official API for Jimeng/Drea…
A minimal solution to hand motion capture from a si…
Pure PyTorch Implementation of NVIDIA paper on Inst…
Image Test Time Augmentation with PyTorch!
Segment Anything in 3D with NeRFs (NeurIPS 2023 & I…
Algorithms and Publications on 3D Object Tracking
Easy NeRF synthetic dataset creation within Blender
Metadata-indexer and Viewer for AI-generated images
Lightweight models for real-time semantic segmentat…
Beautiful and Easy to use Stable Diffusion WebUI
Perception toolkit for sim2real training and valida…
The official implementation of Segment Any 3D GAuss…
Deep Image Matting
[CVPR 2023 Highlight] Neural Kernel Surface Reconst…
The idea of this list is to collect shared data and…
This repository implements a demo of the networks d…
[CVPR 2024] DeepCache: Accelerating Diffusion Model…
ICCV 2023-2025 Papers: Discover cutting-edge resear…
Stable diffusion webui based on diffusers.
CLI for using ComfyUI
online 3d openpose editor for stable diffusion and …
Get Deinked!!
official code repo for paper "CogView2: Faster and …
[ECCV 2026] Skyfall-GS: Synthesizing Immersive 3D U…
A real-time method that estimates the 3D human pose…
Gated-Shape CNN for Semantic Segmentation (ICCV 201…
This project is the implementation of FernRP Packag…
[CVPR 2024 Oral] Rethinking Inductive Biases for Su…
[CVPR 2022] FaceFormer: Speech-Driven 3D Facial Ani…
Implementation of Muse: Text-to-Image Generation vi…
A Kitti Road Segmentation model implemented in tens…
App showcasing multiple real-time diffusion models …
Fix AI pixel art and vector images right in your br…
[CVPR 2016] Unsupervised Feature Learning by Image …
3DMatch - a 3D ConvNet-based local geometric descri…
About This repository is a curated collection of th…
[CVPR 2020] CascadePSP: Toward Class-Agnostic and V…
Learning to Regress 3D Face Shape and Expression fr…
AI-powered tools to enhance Anki flashcards with ex…
PointFlow : 3D Point Cloud Generation with Continuo…
Fine-tune SAM (Segment Anything Model) for computer…
Official PyTorch implementation of "Camera Distance…
Implementation of MeshGPT, SOTA Mesh generation usi…
Learning to Adapt Structured Output Space for Seman…
Official repository accompanying a CVPR 2022 paper …
[ICCV 2021 Oral] PoinTr: Diverse Point Cloud Comple…
Instruct-NeRF2NeRF: Editing 3D Scenes with Instruct…
Python Computer Vision & Video Analytics Framework …
PyTorch implementation of multi-task learning archi…
A microframework on top of PyTorch with first-class…
[CVPR 2024] GaussianDreamer: Fast Generation from T…
Unofficial implementation of BRIA RMBG Model for Co…
Implementations of NeRF variants based on Taichi + …
Unofficial implementation of YOLO-World + Efficien…
The most easy-to-understand tutorial for using LoRA…
Example code for the FLAME 3D head model. The code …
Fuse multiple depth frames into a TSDF voxel volume.
😎 A list of awesome scene understanding papers.
A curated list of 3D Vision papers relating to Robo…
[CVPR 2022] "MonoScene: Monocular 3D Semantic Scene…
High-performance Vision library in Python. Scale yo…
PyTorch implementation of DeepLabV3, trained on the…
Video Frame Interpolation & Super Resolution using …
[ICCV 2023] "TF-ICON: Diffusion-Based Training-Free…
Unofficial implementation of PhotoMaker for ComfyUI
Code for APDrawingGAN: Generating Artistic Portrait…
This is a implementation of the 3D FLAME model in P…
Real-time 3D face tracking and reconstruction from …
YoloDotNet - A C# .NET 8.0 project for Classificati…
[CVPR 2024] Paint3D: Paint Anything 3D with Lightin…
[ECCV 2024 - Oral] ACE0 is a learning-based structu…
[ICCV 2025] LongSplat: Robust Unposed 3D Gaussian S…
[ECCV 2024 Oral 🔥] Arc2Face: A Foundation Model fo…
A lightweight tool for labeling 3D bounding boxes i…
Diffusion attentive attribution maps for interpreti…
Rich-Text-to-Image Generation
基于chatgpt-next-web,增加了midjourney绘画功能,支持mj-plus的ai换脸…
🎬 Generate images from any camera viewpoint via 3D…
Trankit is a Light-Weight Transformer-based Python …
A Tensorflow implementation of RetinexNet
Unsupervised Semantic Segmentation by Distilling Fe…
TensorFlow Implementation for Computing a Semantica…
Using Gemini in ComfyUI
Official code for MAMMA: Markerless Accurate Multi-…
Official PyTorch implementation of Revisiting Image…
[CVPR 2021] Anycost GANs for Interactive Image Synt…
基于JDK8 AI 聊天机器人!微信公众号 Midjourney画图、卡密兑换、web 支持ChatG…
AI-powered Text-to-Art Generator - Text2Art.com
Official Implementation for "Attend-and-Excite: Att…
A collection of awesome text-to-image generation st…
3D face swapping implemented in Python
ENFUGUE is an open-source web app for making studio…
An unofficial implementation of paper 3D Gaussian S…
[ICLR 2025] Official implementation of Posterior-Me…
[CVPR 2025] UniK3D: Universal Camera Monocular 3D E…
Mask3D predicts accurate 3D semantic instances achi…
Run the official Stable Diffusion releases in a Doc…
Generative fill in 3D.
Personalization for Stable Diffusion via Aesthetic …
[ECCV 2024] InstructIR: High-Quality Image Restorat…
SEAIT is a user-friendly application that simplifie…
ComfyUI as a serverless API on Runpod
This repository is a curated collection of the most…
:earth_africa: lightweight C++17 ply 3d mesh format…
[CVPR 2024 Highlight] DistriFusion: Distributed Par…
Animation oriented nodes pack for ComfyUI
This repository contains the source codes for the p…
[ECCV 2020] Learning Enriched Features for Real Ima…
Create images of a given character in different pos…
A tensorflow implementation of "Fast and Accurate I…
The most advanced Nano Banana image generator and e…
[Siggraph Asia 2013] Large-Scale, Real-Time 3D Reco…
A curated list of resources focused on Machine Lear…
A Python frontend and library for ComfyUI
Image augmentation for object detection, segmentati…
🎨 Infinite Drawboard in Python
🦀 Low-level 3D Computer Vision library in Rust
A Fast Deep Learning Model to Upsample Low Resoluti…
Raising the Cost of Malicious AI-Powered Image Edit…
Papers, code and datasets about deep learning for 3…
[ECCV] Swin2SR: SwinV2 Transformer for Compressed I…
Customizable Stable Diffusion frontend for ComfyUI
Editable, part-aware 3D generation from text or ref…
Real-time 3D multi-person pose estimation demo in P…
面向 GPT-image-2 的 AI 图片生成 WebUI 工作台,支持 Codex Respons…
AttentionGAN for Unpaired Image-to-Image Translatio…
[IROS 2021] BundleTrack: 6D Pose Tracking for Novel…
[ECCV'20] Structured3D: A Large Photo-realistic Dat…
[CVPR 2024 & NeurIPS 2024] EmbodiedScan: A Holistic…
HunyuanImage-2.1: An Efficient Diffusion Model for …
Nodes for using ComfyUI as a backend for external t…
FastAPI powered API for Fooocus
News: the 10k dataset is ready for download.
ViewComfy is a open source tool to help you create …
Official implementation of OneDiffusion paper (CVPR…
All-in-one Toolbox for Computer Vision Research.
Flash Diffusion — accelerating conditional diffusio…
EfficientSAM3 compresses SAM3 into lightweight, edg…
[ICCV 2023] A latent space for stochastic diffusion…
[CVPR2024] SeeSR: Towards Semantics-Aware Real-Worl…
A list of papers and resources (data,code,etc) for …
MICA - Towards Metrical Reconstruction of Human Fac…
This repository is a curated collection of the most…
Chat with ChatGPT (gpt-3.5 or newer),WeChat hook in…
基于Stable Diffusion优化的AI绘画模型。支持输入中英文文本,可生成多种现代艺术风格的高…
UPSNet: A Unified Panoptic Segmentation Network
Implementation of Paint-with-words with Stable Diff…
Real-time 3D full-body reconstruction from a single…
face-to-sticker
(Accepted by IJCV) Liquid: Language Models are Scal…
Wining solution and its improvement for MICCAI 2017…
👁️ + 💬 + 🎧 = 🤖 Curated list of top foundation …
All NeRF-related papers @ CVPR/ICCV/ECCV/NIPS/ICML/…
This repository contains a pure C++ ONNX implementa…
Official implementation for "Blended Latent Diffusi…
Official implementation of the CVPR 2022 Paper "Neu…
Keras beit,caformer,CMT,CoAtNet,convnext,davit,dino…
This is tensorflow implementation for paper "Deep I…
[ECCV 2020] Searching Efficient 3D Architectures wi…
For automating the creation of large batches of AI-…
Segmind Distilled diffusion
KeyForge3D is an app that turns a photo of a key in…
[CVPR 2024 Highlight] MIGC and [TPAMI 2024] MIGC++ …
Image Super-Resolution Using Deep Convolutional Net…
FlashAttention (Metal Port)
Python library for analysing faces using PyTorch
Chain together LLMs for reasoning & orchestrate mul…
Official implementation of the paper Plan2Scene.
Self-Supervised Learning of 3D Human Pose using Mul…
This repository provides a comprehensive list of ra…
DeepLab resnet v2 model in pytorch
Generative Adversarial Text to Image Synthesis / Pl…
Evaluating text-to-image/video/3D models with VQASc…
Lightweight Stable Diffusion v 2.1 web UI: txt2img,…
Offline semantic Text-to-Image and Image-to-Image s…
🎨ComfyUI standalone pack with 40+ custom nodes. | …
AI Image Signal Processing and Computational Photog…
Yet another PyTorch implementation of Stable Diffus…
Photometric optimization code for creating the FLAM…
Estimate absolute 3D human poses from RGB images.
[CVPR 2024 - Oral] Matching 2D Images in 3D: Metric…
📚A curated list of Awesome Diffusion Inference Pap…
Official implementation for "Blended Diffusion for …
Official code for the CVPR 2025 paper "SemanticDraw…
ComfyUI adaptation of IDM-VTON for virtual try-on.
A PyTorch Implementation of Neural IMage Assessment
[CVPR 2023] Official code release of our paper "BiF…
118+ plug-and-play JSON style packs for Nano Banana…
Code for "Dense Object Nets: Learning Dense Visual…
[CVPR 2020--Oral] CycleISP: Real Image Restoration …
Open-source local AI SDK - run AI on-device with no…
[ICLR 2026] Trace Anything: Representing Any Video …
CVPR 2024: Language Guided Generation of 3D Embodie…
[NeurIPS 2021] Rethinking Space-Time Networks with …
:video_game: Unity SDK to use the IBM Watson servic…
Official Code for ECCV 2024 paper — One-Shot Diffus…
Prompt Visualization | Art Gallery
Panoptic Segmentation Resources List
Python library for YOLO small object detection and …
Real-time inference for Stable Diffusion - 0.88s la…
Official implementation of Würstchen: Efficient Pre…
A Versatile and Robust SDXL-ControlNet Model for Ad…
Train and block edit and save LoRAs directly insid…
Node Creative Coding / 3D / Image Processing tool i…
基于u-net,cv2以及cnn的中文车牌定位,矫正和端到端识别软件,其中unet和cv2用于车牌定位…
[NeurIPS 2024] PointMamba: A Simple State Space Mod…
[CVPR 2024 Highlight] XCube: Large-Scale 3D Generat…
T2F: text to face generation using Deep Learning
comfyui colabs templates new nodes
Official Implementation of Self-Supervised Street G…
[ICLR'26] Rethinking High-Quality Aesthetic Poster …
Next-generation Albumentations: dual-licensed for o…
Tiled Diffusion, MultiDiffusion, Mixture of Diffuse…
TernausNetV2: Fully Convolutional Network for Insta…
A list of popular deep learning models related to c…
[CVPR 2022] StyleSwin: Transformer-based GAN for Hi…
Virtual Clothing Assistant a custom unique implemen…
A Large-Scale Multimodal Car Dataset with Computati…
Implementation of Parti, Google's pure attention-ba…
Support for miscellaneous image models. Currently s…
Code for "Neural 3D Scene Reconstruction with the M…
ConvMAE: Masked Convolution Meets Masked Autoencode…
Official Code for ICCV 2021 paper "Towards Flexible…
This project is dedicated to the implementation and…
[NeurIPS 2025 Spotlight] A Unified Tokenizer for Vi…
Official implementation for "Break-A-Scene: Extract…
Contrast Enhancement Techniques for low-light images
LEDNet: A Lightweight Encoder-Decoder Network for R…
Current state of supervised and unsupervised depth …
One-click Portable Windows installation of 'AI-Tool…
Official repo for DAD-3DHeads: A Large-scale Dense,…
RESTai is an AIaaS (AI as a Service) open-source pl…
Adversarial Learning for Semi-supervised Semantic S…
Computer vision utils for Blender (generate instanc…
[ECCV 2022 Oral] Perspective Transformer on 3D Lane…
Papers and resources on Controllable Generation usi…
A mega collection of all resources and news related…
[SIGGRAPH Asia 2024] ReVersion: Diffusion-Based Rel…
[CVPR 2023] Official repository for downloading, pr…
🧙🏻♂️A list of papers curated for you to dive int…
Official repository for the paper F, B, Alpha Matti…
PITI: Pretraining is All You Need for Image-to-Imag…
Person Image Synthesis via Denoising Diffusion Mode…
[ECCV 2022] Official repository for "MaxViT: Multi-…
Multi-threaded GUI manager for mass creation of AI-…
Official code for "HumanRF: High-Fidelity Neural Ra…
PyTorch implementation for 3D Bounding Box Estimati…
Official PyTorch implementation of DD3D: Is Pseudo-…
Official PyTorch implementation of "Camera Distance…
Diffusion Classifier leverages pretrained diffusion…
[CVPR 2024✨Highlight] Official repository for HOLD,…
Implementation of the KinectFusion approach in mode…
Fast, offline OCR for Node.js & C++. PP-OCRv6 with …
Diffusion Explainer: Visual Explanation for Text-to…
🧬 Generative modeling of regulatory DNA sequences …
[CVPR 2021] Modular Interactive Video Object Segmen…
Code for PCN: Point Completion Network in 3DV'18 (O…
The First Place Solution of Kaggle iMaterialist (Fa…
LLM-grounded Diffusion: Enhancing Prompt Understand…
Tensorflow framework for the FLAME 3D head model. T…
Official Pytorch implementation of "StyleKeeper: Pr…
Amazing Semantic Segmentation on Tensorflow && Kera…
[CVPR'22 & IJCV'24] Semi-Supervised Semantic Segmen…
Deep Learning for Seismic Imaging and Interpretation
🎨 精选 3000+ Gemini Nano Banana Pro 高质量提示词与生成案例 | 涵盖…
🤖️ 基于 Golang + Vue3 + NaiveUI 的全新的个人、团队、企业私有化AIGC平…
This repo contains the updated version of all the a…
Official implementation for the paper "Deep ViT Fea…
StyleShot: A SnapShot on Any Style. 一款可以迁移任意风格到任意内容…
Convert from Basel Face Model (BFM) to the FLAME he…
[ICCV21] Self-Calibrating Neural Radiance Fields
Enables use of ChatGPT directly from the UI
cuCIM - RAPIDS GPU-accelerated image processing lib…
🏘️ Scaling Embodied AI by Procedurally Generating …
CVPR 2022 papers with code (论文及代码)
[CVPR 2019 Oral] Multi-Channel Attention Selection …
[ACM MM 2023] - LaDI-VTON: Latent Diffusion Textual…
Official implementation of AsymFlow, pi-Flow, GMFlow
Instant-angelo: Build high-fidelity Digital Twin wi…
[ ICLR 2024 ] Official Codebase for "InstructCV: In…
A 3D vision library from 2D keypoints: monocular an…
A CLI tool/python module for generating images from…
a GUI application, which uses YOLOs (YOLOv8, YOLO11…
Intrinsic3D - High-Quality 3D Reconstruction by Joi…
Convert your Stable Diffusion checkpoints quickly a…
基于SpringBoot3开发的Ai平台 含双端 网页以及小程序 包含各类Ai模型 和绘图 ,含支付 …
Pytorch implementation of Structure-Preserving Supe…
CVPR 2023-2024 Papers: Dive into advanced research …
An unified model that seamlessly integrates multimo…
Mixture of Diffusers for scene composition and high…
Accelerate your Stable Diffusion inference with the…
A 3D virtual head control system for VTuber in Unit…
Implicit Geometric Regularization for Learning Shap…
[ICCV 2023] This is the official repository for the…
[CVPR 2023] Collaborative Diffusion
Awesome Monocular 3D detection
A PyTorch Implementation of Fast-SCNN: Fast Semanti…
Choose your diffusion models and spin up a WebUI on…
A collection of resources (including the papers and…
ComfyUI node that let you pick the way in which pro…
[ICCV 2021 Oral] NerfingMVS: Guided Optimization of…
Generate, downscale, change palletes and restore pi…
Vision AI Solution Accelerator
🖼 A collection of high-quality anime faces.
Open-AI's DALL-E for large scale training in mesh-t…
[ACM Computing Surveys] The collection of awesome p…
If your image was a pizza and the CFG the temperatu…
HugAi是由Springboot Vue2 elementUI集成各大AI大模型平台开发的智能问答助…
Official PyTorch implementation of the paper: "Deep…
Relation-Shape Convolutional Neural Network for Poi…
[IROS 2020] se(3)-TrackNet: Data-driven 6D Pose Tra…
Distributed and Graph-based Structure from Motion. …
A framework for Medical Image Segmentation with Con…
Export and run SAM, MobileSAM, EfficientSAM, SAM 2/…
Code and models for our ICCV 2021 paper "MINE: Towa…
Ensembling Off-the-shelf Models for GAN Training (C…
Code for ICCV2021 paper PARE: Part Attention Regres…
💄 Lipstick ain't enough: Beyond Color-Matching for…
3DV 2021: Synergy between 3DMM and 3D Landmarks for…
[ICCV 2021] Instances as Queries
End-to-End SLAM with camera calibration, monocular …
[ICLR 2026] Official repo of paper "Reconstruction …
🚘 Easiest Fully Convolutional Networks
A License Plate Image Reconstruction Project in Ten…
An SDK/Python library for Automatic 1111 to run sta…
Implementation of Newcombe et al. CVPR 2015 Dynamic…
Tiled samplers for ComfyUI
A collection of one-click self-hosted AI
Pytorch implementation of Generative Adversarial Te…
[NeurIPS'23] "MagicBrush: A Manually Annotated Data…
Benchmark diffusion models faster. Automate evals, …
Stable Diffusion Houdini Toolset
Better version for BiRefNet in ComfyUI | Both img &…
attention map tools for huggingface/diffusers
Official implementation for "Stable Flow: Vital Lay…
Extension for Automatic1111 and ComfyUI to automati…
Make use of Intel Arc Series GPU to Run Ollama, Sta…
Officail Implementation for "Cross-Image Attention …
Stable Diffusion Browser for Windows, Mac, and Linux
SD-WEBUI-DISCORD is a Discord bot developed in Go l…
End-to-end recipes for optimizing diffusion models …
Animation via tick. Wave-based parameter modulation…
cutoff implementation for ComfyUI
[CVPR 2024] "MACE: Mass Concept Erasure in Diffusio…
Learn to serve Stable Diffusion models on cloud inf…
Official code base for MinD-Video
BentoDiffusion: A collection of diffusion models se…
Just playing with getting CLIP Guided Diffusion run…
Unofficial implementation of APISR for ComfyUI
ELLA nodes for ComfyUI
[ICCV 2023] Q-Diffusion: Quantizing Diffusion Model…
Implementation of Segformer, Attention + MLP neural…
GPU-ready Dockerfile to run Stability.AI stable-dif…
Using StableDiffusion webui on Colab
RobustSAM: Segment Anything Robustly on Degraded Im…
Official GitHub repository for FLUX.1 Krea [dev].
UI interface for experimenting with multimodal (tex…
Create butter-smooth transitions between prompts, p…
[CVPR 2025 Highlight] Material Anything: Generating…
A list of AI Art courses, tools, libraries, people,…
Python package for segmenting LiDAR data using Segm…
The AI-powered 3D CAD IDE — edit code, visualize in…
Official Repository of the paper "Trajectory Consis…
🤖 Awesome AI
Generate a picture book from a single prompt using …
diffusers implementation for node.js and browser
3D-Adapter: Geometry-Consistent Multi-View Diffusio…
[Neurips 2023 & TPAMI] T2I-CompBench (++) for Compo…
Official implementation of the NeurIPS 2023 paper "…
[ICLR 2025] Official Implementation of Meissonic: R…
A distributed backend AI pipeline server
This is a ChatGPT based prompt generation model for…
A simple web UI for interactive text-guided image t…
[NeurIPS 2024] MeshXL: Neural Coordinate Field for …
This is a Go language version of the SDK based on s…
High-quality Text-to-Audio Generation with Efficien…
A collection of ControlNet poses
Custom nodes for ComfyUI such as CLIP Text Encode++
DiffSeg is an unsupervised zero-shot segmentation m…
Official code repository of < CBGBench: Fill in the…
A small neural network to provide interoperability …
[CVPR2022 oral] A Simple and Effective Baseline for…
AlignProp uses direct reward backpropogation for th…
MIT-Princeton Vision Toolbox for Robotic Pick-and-P…
Pytorch code for ICRA'22 paper: "Single-Shot Multi-…
Low-rank adaptation for Erasing COncepts from diffu…
Colab notebook for Stable Diffusion Hyper-SDXL.
Implementation of Encoder-based Domain Tuning for F…
6 nodes for ComfyUI that allows for more control an…
A Compressed Stable Diffusion for Efficient Text-to…
🔥ICLR 2025 (Spotlight) One-Prompt-One-Story: Free-…
Ovis-Image is a 7B text-to-image model specifically…
Satellite Image Classification using semantic segme…
Local-first AI image organizer and generative media…
Krea 2, MiniMax & Klein 9B LoRA - LoKR Studio — tra…
[CVPR 2023, Highlight] "NeuralLift-360: Lifting An …
Getting the latest versions of Disco Diffusion to w…
stable diffusion multi-user django server code with…
MinImagen: A minimal implementation of the Imagen t…
Bria RMBG 2.0 - image background removee
MIT-Princeton Vision Toolbox for the Amazon Picking…
Joy Caption is a ComfyUI node using the LLaVA model…
A neat Discord bot for old Stable Diffusion Web UIs
🧑🎨 Soothing pastel theme for Stable Diffusion We…
FASHN VTON v1.5: Efficient Maskless Virtual Try-On …
🔥 [ICCV 2025 Highlight] Official ComfyUI native no…
[CVPR 2023] LayoutDM: Discrete Diffusion Model for …
PhotoMaker for ComfyUI
Implementation of a U-net complete with efficient a…
Discord AI chatbot using Ollama and Stable Diffusion
[ICML2025] An 8-step inversion and 8-step editing p…
Better Aligning Text-to-Image Models with Human Pre…
[ICCV 2023] Official PyTorch implementation of the …
Code release for "i1: A Simple and Fully Open Recip…
ENet - A Neural Net Architecture for real time Sema…
A simple Windows / Xbox app for generating AI image…
witcherofresearch/Forgedit — AI 开源项目
[ICCV 2025] VisualCloze: A universal image generati…
Discord bot and Interface for Stable Diffusion
[CVPR 2024] Official repository of "Material Palett…
A gradio web UI demo for Stable Diffusion XL 1.0, w…
Inpaint Anything performs stable diffusion inpainti…
Anti-DreamBooth: Protecting users from personalized…
A paper collection of recent diffusion models for t…
a CLI utility/library for AnimateDiff stable diffus…
[CVPR 2025] Aesthetic Post-Training Diffusion Model…
[CVPR 2025] FaithDiff for Classic Film Rejuvenation…
Official implementation for "Story2Board: A Trainin…
Tiny Dream - An embedded, Header Only, Stable Diffu…
Outfit Anyone in the Wild: Get rid of Annoying Rest…
Officail Implementation for "ReNoise: Real Image In…
Merges two latent diffusion models at a user-define…
[CVPR 2025] MVPaint: Synchronized Multi-View Diffus…
[ECCV 2024, Oral] FMBoost: Boosting Latent Diffusio…
Transcribe audio and add subtitles to videos using …
Stable Diffusion 3 via API in ComfyUI
多模型同时对话、文生图,纯前端。Multi-model simultaneous chat、text-…
A collection of Post Processing Nodes for ComfyUI, …
a powerful stable-diffusion-webui client for android
🔥 A frontend for generating images with Stable Dif…
Official Implementation for "ConceptLab: Creative G…
"4DGen: Grounded 4D Content Generation with Spatial…
This repo implements a Stable Diffusion model in Py…
Yet another multi-purpose Colab Notebook
This is a WhatsApp AI bot that uses various AI mode…
[ICLR 2025] Rectified Diffusion: Straightness Is No…
[CVPR2023] A faster, smaller, and better text-to-im…
Code for instruction-tuning Stable Diffusion.
Your fully proficient, AI-powered and local chatbot…
支持安装、下载、启动和管理多种 AI WebUI / 训练工具的一体化项目,提供 Installer、…
2-4x faster ComfyUI Image Upscaling using Tensorrt ⚡
(CVPR 2025) Code of "Chat2SVG: Vector Graphics Gene…
🤗 Official implementation for "CC-Pan: Channel-wis…
AI 作图知识库
Inference Stable Diffusion with C# and ONNX Runtime
A flexible UI script to help create and expand on p…
Self-hosted, one-tab workbench for the whole traini…
🔥 On-Policy Self-Distillation in Diffusion Models
An extension to AUTOMATIC1111 WebUI for stable diff…
Mac和Windows一键安装Stable Diffusion WebUI,LamaCleaner,S…
Official repository for "CFG++: manifold-constraine…
Turn emoji into amazing artwork via AI
FreeFuse: Multi-Subject LoRA Fusion via Adaptive To…
A playground for creative exploration that uses SDX…
🪢 A reactflow base stable diffusion GUI as ComfyUI…
Implementation of Key-Locked Rank One Editing, from…
C# Stable Diffusion using ONNX Runtime
🐳 | Dockerfiles for the Runpod container images u…
A powerful anti-burn allowing much higher CFG scale…
web UI for GPU-accelerated ONNX pipelines like Stab…
a node for comfyui for restore/edit/enchance faces …
[NeurIPS-2023] Annual Conference on Neural Informat…
A collection of tutorials about training and genera…
The ultimate list of resources to teach yourself ho…
Text to Image Latent Diffusion using a Transformer …
Rebuild the Stable Diffusion Model in a single pyth…
Crie imagens suas usando IA de forma fácil
Codebase for performing various experiments with St…
A Node.js client for Midjourney/Openjourney on Repl…
Create GIFs and Videos using Stable Diffusion
Deforum based on flux-dev by XLabs-AI
[AAAI'2024] IT3D: Improved Text-to-3D Generation wi…
Python library for solving reinforcement learning (…
[CVPR'2024] Official implementation of the paper "E…
Upscale your videos up to 4k on free google colab u…
QWen-VL-Plus & QWen-VL-Max in ComfyUI
Modern AI image generator with multi-provider suppo…
A curated list of awesome Diffusion notebooks, tool…
Pickle Scanner GUI
An awesome & curated list of cool tools for ComfyUI.
Official Code Release for [SIGGRAPH 2024] DiLightNe…
A collection of arbitrary text to image papers with…
A discord bot to generate AI art from prompts using…
收集有关so-vits-svc、TTS、SD、LLMs的各种模型、应用以及文字、声音、图片、视频有关的…
A set of custom ComfyUI nodes for performing post-p…
A Python library for efficient image generation usi…
Code for the paper "pix2gestalt: Amodal Segmentatio…
Run Stable Diffusion using Core ML on iOS within yo…
A Toolkit for OpenAI's Consistency Models.
Playing around with stable diffusion. Generated ima…
Diffusers / Stable Diffusion in docker with a REST …
[AAAI 2026 Oral] LiDARCrafter: Dynamic 4D World Mod…
[ICLR 2025] IterComp: Iterative Composition-Aware F…
Official code for our ICCV2025 paper "SDMatte: Graf…
Generate image from anything with ImageBind and Sta…
Official repo for VGGHeads: 3D Multi Head Alignment…
Deploy Your Own Stable Diffusion Service
Custom node for ComfyUI/Stable Diffustion
Live2Diff: A Pipeline that processes Live video str…