image-generation

99 progetti condividono questo topic GitHub

image-generation — diffusers ★34.1kimage-generationInvokeAI — ★27.6kPixelle-Video — ★25.7ksatori — ★13.7kdream-textures — ★8.2kvllm-omni — ★5.6kmultidiffusion-upscaler-for-automatic1111 — ★5kDragGAN — ★5kStableSwarmUI — ★4.9ksnappy — ★4.5kSwarmUI — ★4.4kOmniGen — ★4.3kAnyDoor — ★4.2kImage-Super-Resolution-via-Iterative-Refinement — ★3.9kDreambooth-Stable-Diffusion — ★3.2kHunyuanImage-3.0 — ★3.2kgpt_image_playground — ★3.1kKandinsky-2 — ★2.8kAwesome-Text-to-Image — ★2.4kComfyUI-to-Python-Extension — ★2.4kDreamOmni2 — ★2kLlamaGen — ★2kganspace — ★1.8kawesome-segment-anything — ★1.7kopendream — ★1.7kru-dalle — ★1.6kVideoClaw — ★1.6kMiniMax-MCP — ★1.5kdiffusiondb — ★1.4kUNO — ★1.4kstable-diffusion-prompt-reader — ★1.3kFireRed-Image-Edit — ★1.3kdata-efficient-gans — ★1.3kLance — ★1.3kMV-Adapter — ★1.3kKnpSnappyBundle — ★1.2kawesome-image-translation — ★1.2kMagicDrive — ★1.2kSDEdit — ★1.2kVITON-HD — ★1.2kWebAI2API — ★1.2kInvokeAI★ 27.6kPixelle-Video★ 25.7ksatori★ 13.7kdream-textures★ 8.2kvllm-omni★ 5.6kmultidiffusion-upscaler-…★ 5kDragGAN★ 5kStableSwarmUI★ 4.9ksnappy★ 4.5kSwarmUI★ 4.4kOmniGen★ 4.3kAnyDoor★ 4.2kImage-Super-Resolution-v…★ 3.9kDreambooth-Stable-Diffus…★ 3.2kHunyuanImage-3.0★ 3.2kgpt_image_playground★ 3.1kKandinsky-2★ 2.8kAwesome-Text-to-Image★ 2.4kComfyUI-to-Python-Extens…★ 2.4kDreamOmni2★ 2kLlamaGen★ 2kganspace★ 1.8kawesome-segment-anything★ 1.7kopendream★ 1.7kru-dalle★ 1.6kVideoClaw★ 1.6kMiniMax-MCP★ 1.5kdiffusiondb★ 1.4kUNO★ 1.4kstable-diffusion-prompt-…★ 1.3kFireRed-Image-Edit★ 1.3kdata-efficient-gans★ 1.3kLance★ 1.3kMV-Adapter★ 1.3kKnpSnappyBundle★ 1.2kawesome-image-translatio…★ 1.2kMagicDrive★ 1.2kSDEdit★ 1.2kVITON-HD★ 1.2kWebAI2API★ 1.2k

Le linee collegano membri che sono misurabilmente correlati tra loro. La dimensione dei punti riflette le stelle.

🧬 Membri
diffusers
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
★ 34.1k
InvokeAI
Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and…
★ 27.6k
Pixelle-Video
🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine
★ 25.7k
satori
Enlightened library to convert HTML and CSS to SVG
★ 13.7k
dream-textures
Stable Diffusion built-in to Blender
★ 8.2k
vllm-omni
A framework for efficient model inference with omni-modality models
★ 5.6k
multidiffusion-upscaler-for-automatic1111
Tiled Diffusion and VAE optimize, licensed under CC BY-NC-SA 4.0
★ 5k
DragGAN
Unofficial Implementation of DragGAN - "Drag Your GAN: Interactive Point-based Manipulation on the Generative…
★ 5k
StableSwarmUI
StableSwarmUI, A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily…
★ 4.9k
snappy
PHP library allowing thumbnail, snapshot or PDF generation from a url or a html page. Wrapper for…
★ 4.5k
SwarmUI
SwarmUI (formerly StableSwarmUI), A Modular Stable Diffusion Web-User-Interface, with an emphasis on making…
★ 4.4k
OmniGen
OmniGen: Unified Image Generation. https://arxiv.org/pdf/2409.11340
★ 4.3k
AnyDoor
Official implementations for paper: Anydoor: zero-shot object-level image customization
★ 4.2k
Image-Super-Resolution-via-Iterative-Refinement
Unofficial implementation of Image Super-Resolution via Iterative Refinement by Pytorch
★ 3.9k
Dreambooth-Stable-Diffusion
Implementation of Dreambooth (https://arxiv.org/abs/2208.12242) by way of Textual Inversion…
★ 3.2k
HunyuanImage-3.0
HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generation
★ 3.2k
gpt_image_playground
基于 OpenAI gpt-image-2 API 的图片生成与编辑工具
★ 3.1k
Kandinsky-2
Kandinsky 2 — multilingual text2image latent diffusion model
★ 2.8k
Awesome-Text-to-Image
(ෆ`꒳´ෆ) A Survey on Text-to-Image Generation/Synthesis.
★ 2.4k
ComfyUI-to-Python-Extension
A powerful tool that translates ComfyUI workflows into executable Python code.
★ 2.4k
DreamOmni2
This project is the official implementation of 'DreamOmni2: Multimodal Instruction-based Editing and…
★ 2k
LlamaGen
Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation
★ 2k
ganspace
Discovering Interpretable GAN Controls [NeurIPS 2020]
★ 1.8k
awesome-segment-anything
Tracking and collecting papers/projects/others related to Segment Anything.
★ 1.7k
opendream
An extensible, easy-to-use, and portable diffusion web UI 👨‍🎨
★ 1.7k
ru-dalle
Generate images from texts. In Russian
★ 1.6k
VideoClaw
🚀 AI 全自动化视频生成员工 | Your First AIGC Coworker. Chat an Idea. Get a Film. 🦞
★ 1.6k
MiniMax-MCP
Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech,…
★ 1.5k
diffusiondb
A large-scale text-to-image prompt gallery dataset based on Stable Diffusion
★ 1.4k
UNO
[ICCV 2025] 🔥🔥 UNO: A Universal Customization Method for Both Single and Multi-Subject Conditioning
★ 1.4k
stable-diffusion-prompt-reader
A simple standalone viewer for reading prompts from Stable Diffusion generated image outside the webui.
★ 1.3k
FireRed-Image-Edit
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art…
★ 1.3k
data-efficient-gans
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
★ 1.3k
Lance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and…
★ 1.3k
MV-Adapter
[ICCV 2025] Official impl. of "MV-Adapter: Multi-view Consistent Image Generation Made Easy"
★ 1.3k
KnpSnappyBundle
Easily create PDF and images in Symfony by converting html using webkit
★ 1.2k
awesome-image-translation
Collection of awesome resources on image-to-image translation.
★ 1.2k
MagicDrive
[ICLR24] Official implementation of the paper “MagicDrive: Street View Generation with Diverse 3D Geometry…
★ 1.2k
SDEdit
PyTorch implementation for SDEdit: Image Synthesis and Editing with Stochastic Differential Equations
★ 1.2k
VITON-HD
Official PyTorch implementation of "VITON-HD: High-Resolution Virtual Try-On via Misalignment-Aware…
★ 1.2k
WebAI2API
★ 1.2k
Bernini
Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner…
★ 1.1k
CogView4
CogView4, CogView3-Plus and CogView3(ECCV 2024)
★ 1.1k
MultiDiffusion
Official Pytorch Implementation for "MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation"…
★ 1.1k
image-extender
Seamlessly extend any image in any direction with AI. Open-source web app powered by Gemini via OpenRouter,…
★ 1.1k
story-iter
[ICLR 2026] A Training-free Iterative Framework for Long Story Visualization
★ 961
HR-VITON
Official PyTorch implementation for the paper High-Resolution Virtual Try-On with Misalignment and…
★ 916
malnyun_faces
침착한 생성모델 학습기
★ 901
domain-transfer-network
TensorFlow Implementation of Unsupervised Cross-Domain Image Generation
★ 860
awesome-text-to-video
A Survey on Text-to-Video Generation/Synthesis.
★ 736
OMG
[ECCV 2024] OMG: Occlusion-friendly Personalized Multi-concept Generation In Diffusion Models
★ 701
ChronoEdit
[ICLR 2026] ChronoEdit: Towards Temporal Reasoning for Image Editing and World Simulation
★ 697
nanobanana-trending-prompts
1,400+ curated trending AI image prompts from X, ranked by engagement. Works with NanoBanana, GPT Image 2,…
★ 677
HunyuanImage-2.1
HunyuanImage-2.1: An Efficient Diffusion Model for High-Resolution (2K) Text-to-Image Generation​
★ 672
Awesome-CVPR2024-Low-Level-Vision
A Collection of Papers and Codes in CVPR2023/2022 about low level vision
★ 660
NOVA
[ICLR 2025] Autoregressive Video Generation without Vector Quantization
★ 655
Uncensored-Local-Studio
Uncensored local AI studio for Windows, Linux, and macOS. Zero-setup GUI for Image Generation, GGUF LLMs,…
★ 630
Flow-Factory
A unified framework for easy reinforcement learning in Flow-Matching models
★ 629
Few-Shot-Patch-Based-Training
The official implementation of our SIGGRAPH 2020 paper Interactive Video Stylization Using Few-Shot…
★ 625
ai-art-generator
For automating the creation of large batches of AI-generated artwork locally.
★ 621
stable-diffusion-2-gui
Lightweight Stable Diffusion v 2.1 web UI: txt2img, img2img, depth2img, inpaint and upscale4x.
★ 600
stable-diffusion-pytorch
Yet another PyTorch implementation of Stable Diffusion (probably easy to read)
★ 593
character_select_stand_alone_app
Character Select Stand Alone App with AI prompt and ComfyUI/WebUI API support for wai-il model
★ 585
storyteller
Multimodal AI Story Teller, built with Stable Diffusion, GPT, and neural text-to-speech
★ 537
ComfyUI-TiledDiffusion
Tiled Diffusion, MultiDiffusion, Mixture of Diffusers, and optimized VAE
★ 536
AI-Visual-Prompt-Cookbook
Curated collection of reusable JSON prompt templates & style references for AI image generation. Updated…
★ 504
ReVersion
[SIGGRAPH Asia 2024] ReVersion: Diffusion-Based Relation Inversion from Images
★ 504
PITI
PITI: Pretraining is All You Need for Image-to-Image Translation
★ 502
Bonsai-Image-Demo
Generate images locally
★ 502
BlendGAN
Official PyTorch implementation of "BlendGAN: Implicitly GAN Blending for Arbitrary Stylized Face Generation"…
★ 498
dream-factory
Multi-threaded GUI manager for mass creation of AI-generated art with support for multiple GPUs.
★ 496
AmazingZImageWorkflow
Z-Image workflow with predefined styles for high-quality image generation and a user-friendly experience.…
★ 490
LLM-groundedDiffusion
LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language…
★ 482
FreezeG
Freezing generator for pseudo image translation
★ 473
clip-guided-diffusion
A CLI tool/python module for generating images from text using guided diffusion and CLIP from OpenAI.
★ 459
AutoStudio
[CVPRW 2026] AutoStudio: Crafting Consistent Subjects in Multi-turn Interactive Image Generation
★ 450
Qihoo-T2X
Efficient DiT architecture for text2any tasks, ICLR2025
★ 446
comfyui-api
A simple API server to make ComfyUI easy to scale horizontally. Get outputs directly in the response, or…
★ 440
pytorch-generative
Easy generative modeling in PyTorch
★ 437
Awesome-Try-On-Models
A repository for organizing papers, codes and other resources related to Virtual Try-on Models
★ 436
inpainting_gmcnn
Image Inpainting via Generative Multi-column Convolutional Neural Networks, NeurIPS2018
★ 435
FreeDrag
[CVPR 2024] Official implementation of FreeDrag: Feature Dragging for Reliable Point-based Image Editing
★ 418
Text-to-Image-Synthesis
Pytorch implementation of Generative Adversarial Text-to-Image Synthesis paper
★ 411
MagicBrush
[NeurIPS'23] "MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing".
★ 411
SpatialGen
[3DV 2026] SpatialGen: Layout-guided 3D Indoor Scene Generation
★ 402
Awesome-Generation-Acceleration
📚 Collection of awesome generation acceleration resources.
★ 401
LLMGA
This project is the official implementation of 'LLMGA: Multimodal Large Language Model based Generation…
★ 396
VINE
[ICLR 2025] "Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to…
★ 389
RelaCtrl
Efficient controlnet for DiTs
★ 387
Nexior
Consumer AI app for chat, image generation, video generation, and music creation powered by Ace Data Cloud…
★ 379
stable-diffusion-nvidia-docker
GPU-ready Dockerfile to run Stability.AI stable-diffusion model v2 with a simple web interface. Includes…
★ 370
ZenCtrl
In-context subject-driven image generation while preserving foreground fidelity
★ 351
Voost
[SIGGRAPH Asia 25] Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and…
★ 340
OmniTokenizer
[NeurIPS 2024]OmniTokenizer: one model and one weight for image-video joint tokenization.
★ 325
ComfyUI-ZImagePowerNodes
A set of ComfyUI nodes designed specifically for the Z-Image / Z-Image Turbo model.
★ 319
Ovis-Image
Ovis-Image is a 7B text-to-image model specifically optimized for high-quality text rendering, designed to…
★ 318
the-og
A pure PHP OpenGraph Image Generator
★ 305
Pulse-of-Motion
The Pulse of Motion: Measuring Physical Frame Rate from Visual Dynamics
★ 71
openvenice
Open-source, customizable frontend for Venice AI. Chat, image gen, audio, video, embeddings + visual…
★ 55
🔗 Famiglie affini

Misurato dai temi di GitHub condivisi da entrambi i progetti, ponderato in base a quanto è raro ciascun tema.