awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

163 مستودعات

Awesome GitHub RepositoriesImage Generation

Tools and interfaces for creating visual content from text prompts.

Distinguishing note: Focuses on the core image generation capability.

Explore 163 awesome GitHub repositories matching artificial intelligence & ml · Image Generation. Refine with filters or upvote what's useful.

Awesome Image Generation GitHub Repositories

اعثر على أفضل المستودعات باستخدام الذكاء الاصطناعي.سنبحث عن أفضل المستودعات المطابقة باستخدام الذكاء الاصطناعي.
  • comfyanonymous/comfyuiالصورة الرمزية لـ comfyanonymous

    comfyanonymous/ComfyUI

    117,322عرض على GitHub↗

    ComfyUI is a modular generative AI workflow orchestrator and node-based GUI for designing and executing complex diffusion model pipelines. It functions as both a visual interface for building generative logic graphs and a programmable backend API that exposes diffusion model operations for external integration. The system distinguishes itself through a graph-based execution model that supports differential workflow execution, re-running only modified nodes to reduce computation. It features dynamic model offloading to manage memory between system RAM and GPU VRAM and utilizes metadata-embedde

    Supports advanced image modification through inpainting, outpainting, and masking using diffusion models.

    Python
    عرض على GitHub↗117,322
  • lllyasviel/fooocusالصورة الرمزية لـ lllyasviel

    lllyasviel/Fooocus

    50,260عرض على GitHub↗

    Fooocus is a generative image interface designed to simplify the creation of high-quality visual content from text descriptions. It functions as a latent diffusion pipeline and model orchestrator, managing the complex interactions between neural network layers, mathematical samplers, and hardware resource allocation to produce professional-grade imagery. The project distinguishes itself through a sophisticated prompt engineering engine and modular style management. Users can dynamically modify output characteristics by injecting style adapters directly into prompts or by utilizing wildcards a

    Creates high-quality visual content from text descriptions using a streamlined interface.

    Python
    عرض على GitHub↗50,260
  • anthropics/anthropic-cookbookالصورة الرمزية لـ anthropics

    anthropics/anthropic-cookbook

    45,984عرض على GitHub↗

    This repository is a collection of guides, notebooks, and recipes for implementing advanced prompting techniques and workflow patterns with large language models. It serves as a prompt engineering guide, an evaluation suite for scoring prompt quality, and a framework for orchestrating agents and integrating external tools. The project provides implementation patterns for building applications with Claude, specifically focusing on coordinating multiple models to split complex tasks between high-reasoning and high-efficiency agents. It includes technical demonstrations for multimodal data proce

    Coordinates with external image generation models to create visual assets based on text descriptions.

    Jupyter Notebook
    عرض على GitHub↗45,984
  • heyputer/puterالصورة الرمزية لـ HeyPuter

    HeyPuter/puter

    42,318عرض على GitHub↗

    Puter is a browser-based desktop environment and cloud-native development platform that provides a virtualized graphical workspace. It enables developers to build and deploy full-stack web applications by integrating cloud storage, authentication, and serverless backend logic directly into the browser, eliminating the need for traditional server infrastructure. The platform distinguishes itself through a unified cloud storage layer and a distributed network runtime that facilitates peer-to-peer communication and cross-origin resource fetching. It features a sophisticated cross-window orchestr

    Creates images from text prompts by configuring model parameters like aspect ratio, quality, and seed.

    TypeScriptcloudcloud-oscloud-storage
    عرض على GitHub↗42,318
  • bin-huang/chatboxالصورة الرمزية لـ Bin-Huang

    Bin-Huang/chatbox

    40,509عرض على GitHub↗

    Chatbox is a desktop client and multi-provider chat interface for interacting with large language model APIs across various service providers and local installations. It functions as a local-first AI conversation manager that stores chat history and user settings directly on the device. The application provides a unified interface to connect multiple AI backends for text generation and image creation. It includes a specialized rendering system for AI responses that supports technical documentation through syntax highlighting, Markdown, and Latex mathematical notation. The platform manages pr

    Integrates capabilities to generate images from text descriptions via AI models.

    TypeScript
    عرض على GitHub↗40,509
  • quantumnous/new-apiالصورة الرمزية لـ QuantumNous

    QuantumNous/new-api

    39,722عرض على GitHub↗

    This project is an AI model API gateway and proxy server designed to provide a unified interface for interacting with diverse artificial intelligence service providers. It functions as a centralized middleware platform that routes, load balances, and translates API requests across multiple models, enabling developers to access text, image, audio, and video generation capabilities through a single, standardized integration. The gateway distinguishes itself through comprehensive administrative and financial controls, including event-driven usage accounting, real-time token consumption tracking,

    Creates visual content from text descriptions by routing prompts to configured image generation models.

    Goai-gatewayclaudedeepseek
    عرض على GitHub↗39,722
  • xingangpan/dragganالصورة الرمزية لـ XingangPan

    XingangPan/DragGAN

    35,822عرض على GitHub↗

    DragGAN is an interactive generative AI editor and GAN image editing tool designed for modifying the shape and structure of objects within images. It functions as a latent space manipulator that enables precise geometric and appearance editing by transforming images into editable latent codes. The system provides a web-based visual dashboard for real-time manipulation. Users can change the appearance of generated objects through an interactive point-based dragging interface, utilizing a process where source and target coordinates drive the optimization of the generative model. The project in

    Enables the transformation of real photographs into latent representations to allow interactive editing of non-generated imagery.

    Pythonartificial-intelligencegenerative-adversarial-networkgenerative-models
    عرض على GitHub↗35,822
  • microsoft/taskmatrixالصورة الرمزية لـ microsoft

    microsoft/TaskMatrix

    34,079عرض على GitHub↗

    TaskMatrix is a visual language model orchestration framework and modular visual pipeline designed to coordinate disparate foundation models. It functions as a multi-model workflow coordinator that sequences visual and textual models through logic paths to handle image processing tasks without requiring additional training. The system integrates large language models with visual foundation models to enable the exchange of image data during interactive chat sessions. It utilizes template-based orchestration to chain specialized models together for complex visual tasks. The framework supports

    Combines bounding boxes and segmentation masks with text-driven generative fills to modify specific image regions.

    Python
    عرض على GitHub↗34,079
  • chenfei-wu/taskmatrixالصورة الرمزية لـ chenfei-wu

    chenfei-wu/TaskMatrix

    34,082عرض على GitHub↗

    TaskMatrix is a multimodal AI chat interface and visual task orchestrator. It combines language models with visual recognition to enable the exchange, analysis, and modification of images within a conversational environment. The system coordinates multiple foundation models through orchestration pipelines that chain language, detection, and segmentation models. This allows for complex visual operations, such as using text instructions to guide image masking and executing modular inpainting workflows to edit specific image regions. The project includes a computer vision toolset for object det

    Generates precise pixel-level masks from natural language descriptions to guide image editing.

    Python
    عرض على GitHub↗34,082
  • microsoft/visual-chatgptالصورة الرمزية لـ microsoft

    microsoft/visual-chatgpt

    34,079عرض على GitHub↗

    Visual-ChatGPT is a visual orchestration framework and multimodal AI pipeline designed to coordinate large language models with visual foundation models. It functions as an integration layer that enables the exchange of text and images between different AI models to automate image analysis and editing tasks without requiring additional model training. The system differentiates itself through model-chain orchestration and prompt-based task dispatching, allowing natural language instructions to trigger specific vision models or tools. It utilizes coordinate-based region mapping and iterative ma

    Creates and modifies images using a combination of text-guided bounding boxes, segmentation masks, and inpainting.

    Python
    عرض على GitHub↗34,079
  • huggingface/diffusersالصورة الرمزية لـ huggingface

    huggingface/diffusers

    33,872عرض على GitHub↗

    Diffusers is a PyTorch-based library and generative AI framework used to build, train, and deploy diffusion pipelines for producing multi-modal media. It provides a suite of tools for generating images, video, and audio from natural language descriptions, as well as specialized systems for text-to-image generation. The project differentiates itself through a modular architecture that separates noise schedulers, pretrained model blocks, and pipeline compositions. This structure allows for the construction of custom generation workflows and the ability to swap individual components of the diffu

    Provides techniques for creating visual variations of a source image while maintaining subject and composition.

    Pythondeep-learningdiffusionflux
    عرض على GitHub↗33,872
  • danswer-ai/danswerالصورة الرمزية لـ danswer-ai

    danswer-ai/danswer

    30,552عرض على GitHub↗

    Danswer is an LLM application framework and RAG engine that provides a self-hosted interface for connecting large language models to private data. It serves as an enterprise AI chat interface and agent orchestrator, enabling the creation of specialized assistants with custom instructions and knowledge bases. The platform differentiates itself through an observability dashboard for tracking query history and token consumption, as well as a white-labeled interface for customized branding. It includes a multi-step research workflow for producing long-form reports and a sandboxed environment for

    Translates descriptive text prompts into original visual images.

    Python
    عرض على GitHub↗30,552
  • sgl-project/sglangالصورة الرمزية لـ sgl-project

    sgl-project/sglang

    29,079عرض على GitHub↗

    Sglang is a high-performance inference engine and serving system designed for large language and multimodal models. It provides a programmable interface for orchestrating complex generation workflows, enabling developers to coordinate multi-turn dialogues, tool invocations, and reasoning chains through a domain-specific language. The platform is built to support production-scale deployments, offering an OpenAI-compatible API that allows for integration with existing application ecosystems. The system distinguishes itself through a disaggregated architecture that separates compute-intensive pr

    Requests image generation from served models using text prompts and quality presets.

    Pythonattentionblackwellcuda
    عرض على GitHub↗29,079
  • black-forest-labs/fluxالصورة الرمزية لـ black-forest-labs

    black-forest-labs/flux

    25,637عرض على GitHub↗

    Flux is a diffusion model inference engine designed for text-to-image generation and image-to-image manipulation. It provides a system for executing open-weight models to transform natural language descriptions into visual imagery or to modify existing images. The project distinguishes itself through a flow-matching framework for image generation and a structural image controller. This controller allows for guided synthesis by using depth maps and Canny edge detection to constrain the geometry and composition of the output. The toolkit covers a broad range of image editing capabilities, incl

    Incorporates spatial and structural constraints like Canny edges and depth maps to preserve image composition.

    Python
    عرض على GitHub↗25,637
  • pytorch/examplesالصورة الرمزية لـ pytorch

    pytorch/examples

    23,752عرض على GitHub↗

    This repository serves as a comprehensive collection of reference implementations for the PyTorch machine learning library. It provides practical examples for building, training, and deploying deep learning models, functioning as a toolkit for developers to explore neural network architectures and training workflows. The project distinguishes itself by offering concrete demonstrations of complex machine learning operations, ranging from computer vision tasks like object detection and depth estimation to the training of large-scale transformer models. These examples illustrate how to implement

    Trains generative adversarial networks to produce high-resolution images through incremental resolution scaling.

    Python
    عرض على GitHub↗23,752
  • sanster/iopaintالصورة الرمزية لـ Sanster

    Sanster/IOPaint

    23,244عرض على GitHub↗

    IOPaint is an AI image editor and Stable Diffusion inpainting tool providing a web interface for removing objects and replacing image content. It utilizes latent diffusion image processing to synthesize high-resolution replacements for erased sections of an image. The project features a specialized AI background remover for isolating subjects and an AI image upscaler that employs super-resolution models for general photos and anime artwork. The software covers a broad range of capabilities including image segmentation for object isolation, face restoration for improving facial details, and t

    Modifies existing pictures by following natural language commands to perform visual modifications.

    Pythoninpaintinglamalatent-diffusion
    عرض على GitHub↗23,244
  • sanster/lama-cleanerالصورة الرمزية لـ Sanster

    Sanster/lama-cleaner

    23,235عرض على GitHub↗

    Lama Cleaner is an AI-powered image editing application focused on inpainting, object removal, and generative filling. It provides a suite of tools for erasing unwanted elements from photos and filling the resulting gaps using generative artificial intelligence. The project includes specialized capabilities for image outpainting to extend borders, background removal through object segmentation, and face restoration to fix visual defects. It also features an image upscaler to increase resolution and clarity via super-resolution AI, as well as a Stable Diffusion-based editor for replacing speci

    Swaps existing objects or elements within an image with new content using diffusion models.

    Python
    عرض على GitHub↗23,235
  • vercel/aiالصورة الرمزية لـ vercel

    vercel/ai

    21,885عرض على GitHub↗

    This project is a comprehensive framework for building AI-powered applications, providing a unified toolkit for orchestrating language models, autonomous agents, and interactive user interfaces. It serves as a central library for managing the entire lifecycle of AI interactions, from initial prompt generation and model provider abstraction to complex, multi-step reasoning and tool execution. The framework distinguishes itself through its deep integration with frontend development, specifically by enabling generative user interfaces that render dynamic components directly from model outputs. I

    Generates images from text descriptions with support for batching and provider-specific configurations.

    TypeScriptanthropicartificial-intelligencegemini
    عرض على GitHub↗21,885
  • anil-matcha/open-higgsfield-aiالصورة الرمزية لـ Anil-matcha

    Anil-matcha/Open-Higgsfield-AI

    20,529عرض على GitHub↗

    Open-Higgsfield-AI is a generative AI content studio and visual workflow orchestrator. It provides a unified interface for creating photorealistic images and videos, utilizing a node-based editor to chain multiple image, video, and audio models into automated content pipelines. The system functions as an AI video animation tool and local GPU inference engine, allowing users to run generative models on local hardware or remote servers. It includes specialized capabilities for audio-driven lip synchronization and cinematic camera controls to adjust virtual lens and focal settings. The platform

    Maintains a local cache of uploaded visual assets to serve as conditional inputs for generative models.

    JavaScriptai-art-generatorai-image-generationai-video-generation
    عرض على GitHub↗20,529
  • kwaivgi/liveportraitالصورة الرمزية لـ KwaiVGI

    KwaiVGI/LivePortrait

    18,632عرض على GitHub↗

    LivePortrait is a deep learning framework for portrait animation that transfers facial expressions from a driving video to a static image. It functions as an AI motion retargeting tool, mapping movements between different identities while preserving the unique features of the source portrait. The system includes specialized capabilities for cross-species portrait animation, adapting human-centric models to non-human subjects and animals. It also features a motion template generator that converts driving videos into portable files to accelerate inference and protect the identity of the origina

    Allows for generative modifications and pose adjustments to specific isolated areas of a portrait.

    Python
    عرض على GitHub↗18,632
السابق123456…9التالي
  1. Home
  2. Artificial Intelligence & ML
  3. Image Generation

استكشف الوسوم الفرعية

  • Architectural GenerationSpecialized image generation that preserves straight lines and geometric consistency for buildings and interiors. **Distinct from Image Generation:** Distinct from Image Generation: focuses specifically on architectural geometry and straight-line consistency.
  • Audio Style Transfers2 وسوم فرعيةTechniques for applying the stylistic properties of a reference source to an audio signal. **Distinct from Style Transfers:** Operates on audio signals and sonic characteristics rather than image aesthetics.
  • Blank Image CreationGeneration of empty image canvases with specified dimensions and colors. **Distinct from Image Generation:** Distinct from AI Image Generation: focuses on creating basic blank canvases rather than synthetic content via models
  • BotsAutomated agents that generate images from text prompts and perform reverse image searches. **Distinct from Image Generation:** Distinct from Image Generation: adds a bot interface for automated image creation and reverse image search, not just the core generation capability.
  • Composition-Controlled GeneratorsImage generators that incorporate spatial and structural constraints to manage subject placement and pose. **Distinct from Image Generation:** Image Generation is a broad capability; this specifies the addition of composition control.
  • Generative Adversarial ArchitecturesFrameworks for training generative adversarial networks to synthesize high-quality visual content. **Distinct from Image Generation:** Distinct from Image Generation: focuses on the specific GAN-based training methodology for resolution enhancement rather than general text-to-image generation.
  • Graphic Composition UtilitiesUtilities for composing visual content with support for external assets and custom font configuration. **Distinct from Image Generation:** Distinct from Image Generation: focuses on manual composition and layout rather than generative AI synthesis.
  • HEIF Image FilesCreates HEIF image files from video streams, with options for frame selection and tiled content splitting. **Distinct from Image Generation:** Distinct from general Image Generation: specifically creates HEIF container files from HEVC streams, not AI-generated or containerized images.
  • Image Concept IntegrationIntegrating encoded image-based concepts into the generation pipeline to influence the output. **Distinct from Image Generation:** Specifically about using images to influence a generative process, not general image generation
  • Image Editing21 وسوم فرعيةTools for modifying existing visual content using generative AI instructions. **Distinct from Image Generation:** Distinct from Image Generation: focuses on inpainting and image-to-image transformations rather than creating new images from scratch.
  • Image-Conditioned GenerationCreating new images using another image as a structural or stylistic reference. **Distinct from Image Generation:** Distinct from general generation by requiring an image as a primary prompt/reference.
  • Post-Generation RefinementsTools that enhance or correct images after initial AI generation, including upscaling and facial detail fixes. **Distinct from Image Generation:** Distinct from Image Generation: focuses on post-processing steps after the initial image is created, not the generation itself.
  • Reference-Conditioned Generation2 وسوم فرعيةGenerating images using a combination of text prompts and external visual reference files. **Distinct from Image Generation:** Specifically addresses using reference files to modify styles or backgrounds within the generation process.
  • Scripted GenerationAutomating the production of generative images via scripts and predefined constraints without a UI. **Distinct from Image Generation:** Focuses on head-less, scripted production based on constraints, unlike the general UI-driven image generation parent.
  • Style Transfers1 وسم فرعيApplying the aesthetic and stylistic properties of a reference source to a generated image. **Distinct from Image Generation:** Focuses specifically on aesthetic transfer rather than general creation of images from text.
  • Template-DrivenCreation of visual images based on structured data templates rather than AI prompts. **Distinct from Image Generation:** Distinct from AI image generation; focuses on programmatic generation from JSON templates
  • Unconditional GenerationSynthesizing data based on learned distributions without external conditioning signals. **Distinct from Image Generation:** Distinct from general Image Generation as it explicitly excludes text or label guidance