awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

26 रिपॉजिटरी

Awesome GitHub RepositoriesGenerative AI Workflows

Defined sequences of automated steps for creating and refining generative content through iterative processing.

Explore 26 awesome GitHub repositories matching artificial intelligence & ml · Generative AI Workflows. Refine with filters or upvote what's useful.

Awesome Generative AI Workflows GitHub Repositories

AI के साथ बेहतरीन रिपॉजिटरी खोजें।हम AI का उपयोग करके सबसे सटीक रिपॉजिटरी खोजेंगे।
  • comfy-org/comfyuiComfy-Org का अवतार

    Comfy-Org/ComfyUI

    117,227GitHub पर देखें↗

    ComfyUI is a node-based generative AI orchestration engine designed for constructing, testing, and executing complex image and video synthesis pipelines. By utilizing a directed acyclic graph execution model, the platform allows users to build reproducible workflows through modular, interconnected processing blocks without requiring manual code implementation. It serves as both a local environment for high-performance model inference and a production-ready server for deploying generative capabilities. The platform distinguishes itself through its focus on workflow portability and extensibilit

    Synthesizes video sequences from textual prompts by executing complex, multi-stage generative workflows.

    Pythonaicomfycomfyui
    GitHub पर देखें↗117,227
  • quantumnous/new-apiQuantumNous का अवतार

    QuantumNous/new-api

    39,722GitHub पर देखें↗

    This project is an AI model API gateway and proxy server designed to provide a unified interface for interacting with diverse artificial intelligence service providers. It functions as a centralized middleware platform that routes, load balances, and translates API requests across multiple models, enabling developers to access text, image, audio, and video generation capabilities through a single, standardized integration. The gateway distinguishes itself through comprehensive administrative and financial controls, including event-driven usage accounting, real-time token consumption tracking,

    Provides workflows for generating modified or extended versions of existing images via text instructions.

    Goai-gatewayclaudedeepseek
    GitHub पर देखें↗39,722
  • chenfei-wu/taskmatrixchenfei-wu का अवतार

    chenfei-wu/TaskMatrix

    34,082GitHub पर देखें↗

    TaskMatrix is a multimodal AI chat interface and visual task orchestrator. It combines language models with visual recognition to enable the exchange, analysis, and modification of images within a conversational environment. The system coordinates multiple foundation models through orchestration pipelines that chain language, detection, and segmentation models. This allows for complex visual operations, such as using text instructions to guide image masking and executing modular inpainting workflows to edit specific image regions. The project includes a computer vision toolset for object det

    Provides a modular workflow for modifying image regions using segmentation masks and diffusion models.

    Python
    GitHub पर देखें↗34,082
  • anil-matcha/open-higgsfield-aiAnil-matcha का अवतार

    Anil-matcha/Open-Higgsfield-AI

    20,529GitHub पर देखें↗

    Open-Higgsfield-AI is a generative AI content studio and visual workflow orchestrator. It provides a unified interface for creating photorealistic images and videos, utilizing a node-based editor to chain multiple image, video, and audio models into automated content pipelines. The system functions as an AI video animation tool and local GPU inference engine, allowing users to run generative models on local hardware or remote servers. It includes specialized capabilities for audio-driven lip synchronization and cinematic camera controls to adjust virtual lens and focal settings. The platform

    Designs automated sequences of image, video, and audio models to create generative content.

    JavaScriptai-art-generatorai-image-generationai-video-generation
    GitHub पर देखें↗20,529
  • thudm/cogvideoTHUDM का अवतार

    THUDM/CogVideo

    12,792GitHub पर देखें↗

    CogVideo is a generative video framework that uses diffusion models and transformer-based architectures to synthesize high-resolution video clips. It functions as both a text-to-video and image-to-video generator, converting textual descriptions or static images into temporal visual sequences. The system integrates large language model capabilities to expand short user prompts into detailed descriptions for better visual alignment. It supports the animation of static images through latent seeding and provides the ability to extend the length of existing video sequences. The project includes

    Integrates a large language model to expand short user prompts into detailed descriptions to improve video generation quality.

    Python
    GitHub पर देखें↗12,792
  • stability-ai/stablestudioStability-AI का अवतार

    Stability-AI/StableStudio

    9,045GitHub पर देखें↗

    StableStudio is a generative AI frontend and image interface designed for creating and editing visual content. It provides a web-based graphical interface that connects to generative AI models via API connections to facilitate image synthesis and modification. The project functions as a pluggable AI backend manager, using a modular system to standardize diverse AI provider APIs into a unified format. This architecture allows users to swap between different generative AI backends and providers to compare outputs and optimize production. The system manages long-running generation tasks through

    Provides specialized workflows for image-to-image manipulation and AI-driven artistic refinement.

    TypeScriptfrontendmlstability-ai
    GitHub पर देखें↗9,045
  • dusty-nv/jetson-inferencedusty-nv का अवतार

    dusty-nv/jetson-inference

    8,734GitHub पर देखें↗

    jetson-inference is a set of libraries and tools for executing optimized deep learning models on embedded GPU hardware. Its primary purpose is to enable real-time computer vision and AI inference at the edge with low latency and high throughput. The project distinguishes itself through high-performance streaming analytics and the ability to execute concurrent AI pipelines on auto-grade silicon. It provides specialized support for multi-sensor stream processing, utilizing zero-copy data transport to load camera frames directly into GPU memory. The codebase covers a broad surface of capabiliti

    Builds retrieval-augmented generation pipelines and agentic AI applications using hosted API endpoints.

    C++caffecomputer-visiondeep-learning
    GitHub पर देखें↗8,734
  • ml-explore/mlx-examplesml-explore का अवतार

    ml-explore/mlx-examples

    8,254GitHub पर देखें↗

    This repository provides a collection of reference implementations and code examples for training and deploying machine learning models using the MLX framework. It serves as a practical guide for executing distributed training, fine-tuning large language models, converting model weights, and implementing multimodal generative workflows. The project distinguishes itself through specialized examples for local hardware execution, featuring weight quantization to reduce memory usage and low-rank adaptation for parameter-efficient fine-tuning. It also includes scripts for transforming external mod

    Provides sequences of automated steps for creating images, video, music, and text using diffusion models and transformers.

    Pythonmlx
    GitHub पर देखें↗8,254
  • lykosai/stabilitymatrixLykosAI का अवतार

    LykosAI/StabilityMatrix

    7,544GitHub पर देखें↗

    StabilityMatrix is a centralized installer and orchestrator for Stable Diffusion web interfaces and their dependencies. It functions as a generative AI workspace and portable runtime, providing a unified interface to install and update AI image generation packages within isolated environments to prevent global system conflicts. The project distinguishes itself through a shared model manager that imports, organizes, and shares checkpoints across different installations. It utilizes a central model repository and filesystem mapping to allow multiple packages to access the same large binary asse

    Provides a workflow for creating visual assets via a dedicated interface and managed workspaces.

    C#aiautomatic1111avalonia
    GitHub पर देखें↗7,544
  • 0xplaygrounds/rig0xPlaygrounds का अवतार

    0xPlaygrounds/rig

    7,450GitHub पर देखें↗

    Rig is a framework for building large language model applications, featuring a multi-provider client and a workflow builder for retrieval-augmented generation systems. It serves as an orchestrator for creating autonomous agents that can maintain conversation state and execute complex tasks through custom prompting and plugins. The project provides standardized interfaces for both completion and embedding model providers, allowing for unified request and response patterns across different engines. It also includes a vector database integration layer that defines a common interface for indexing

    Enables the design of automated sequences that produce multimedia content using various generative AI model capabilities.

    Rustagentaiartificial-intelligence
    GitHub पर देखें↗7,450
  • abdbarho/stable-diffusion-webui-dockerAbdBarho का अवतार

    AbdBarho/stable-diffusion-webui-docker

    7,315GitHub पर देखें↗

    This project is a containerized deployment for running Stable Diffusion web interfaces. It provides a portable runtime for generative AI that manages dependencies and hardware acceleration to enable text-to-image generation and image-to-image transformations via a browser-based interface. The system uses hardware-specific image tags to support both GPU-accelerated synthesis and CPU-only execution. It ensures environment isolation across different operating systems while utilizing bind-mount data persistence to keep heavy model weights and generated outputs on the host machine. The deployment

    Implements a portable container setup for executing automated sequences of generative image refinement.

    Shell
    GitHub पर देखें↗7,315
  • civitai/civitaicivitai का अवतार

    civitai/civitai

    7,158GitHub पर देखें↗

    Civitai is a platform for generative media creation and AI model distribution. It provides a centralized service for producing images, videos, audio, and music, while serving as a repository where users can share, discover, and browse custom model weights and fine-tuned adaptations. The platform distinguishes itself through a provider-agnostic orchestration layer that manages multi-step generation pipelines and complex workflows across different backends. It integrates with autonomous AI agents and editors via the Model Context Protocol, allowing external tools to access generation pipelines

    Manages complex generation pipelines by submitting and controlling raw configurations.

    TypeScriptaisocial-networkstable-diffusion
    GitHub पर देखें↗7,158
  • skyworkai/skyreels-v2SkyworkAI का अवतार

    SkyworkAI/SkyReels-V2

    6,356GitHub पर देखें↗

    SkyReels-V2 is a video generation system that creates, extends, and refines video clips from text descriptions, images, or both. It operates as a diffusion-based video generation model that can produce videos of any duration by denoising frames sequentially, with each new frame conditioned on the ones that came before it. The system supports generating videos from scratch using text prompts, starting from a single image and producing subsequent frames, or constraining both the first and last frames to match user-provided images. What distinguishes SkyReels-V2 is its combination of infinite-le

    Expands brief text descriptions into detailed prompts using a language model for better video alignment.

    Python
    GitHub पर देखें↗6,356
  • mervinpraison/praisonaiMervinPraison का अवतार

    MervinPraison/PraisonAI

    5,592GitHub पर देखें↗

    PraisonAI is an autonomous AI agent platform that coordinates multiple LLM-powered agents for research, planning, and execution of complex workflows. It functions as a multi-agent orchestration framework, a workflow builder, and a Model Context Protocol server, while also providing retrieval-augmented generation through vector knowledge bases. Agents can interact via CLI, web, or standardized protocols with sandboxed code execution. The platform distinguishes itself with a rich set of agent communication protocols, including A2A, REST, WebSocket, voice and telephony integration, and MCP, allo

    Expands a brief input into a comprehensive, structured prompt using one of several expansion strategies.

    Pythonagentsaiai-agent-framework
    GitHub पर देखें↗5,592
  • bram2w/baserowbram2w का अवतार

    bram2w/baserow

    5,085GitHub पर देखें↗

    Baserow एक नो-कोड रिलेशनल डेटाबेस और एप्लिकेशन बिल्डर है जो उपयोगकर्ताओं को विज़ुअल इंटरफ़ेस के माध्यम से स्ट्रक्चर्ड डेटा टेबल और बिज़नेस टूल्स बनाने की अनुमति देता है। यह एक हेडलेस REST API डेटा बैकएंड और एक सेल्फ-होस्टेड डेटा वर्कस्पेस के रूप में कार्य करता है, जो डेटा रेसीडेंसी पर पूर्ण नियंत्रण बनाए रखते हुए सहयोगी डेटाबेस को मैनेज करने के लिए एक प्लेटफ़ॉर्म प्रदान करता है। यह प्लेटफ़ॉर्म LLM-संचालित डेटा प्लेटफ़ॉर्म के रूप में कार्य करने के लिए लार्ज लैंग्वेज मॉडल्स को एकीकृत करता है, जो प्राकृतिक भाषा से डेटाबेस संरचनाएं, रिकॉर्ड कंटेंट और तकनीकी वर्कफ़्लो उत्पन्न करने में सक्षम है। यह एक मॉडल कॉन्टेक्स्ट प्रोटोकॉल सर्वर के रूप में भी कार्य करता है, जो रिमोट AI एजेंट्स को प्रोग्रामेटिक रूप से स्ट्रक्चर्ड डेटाबेस रिकॉर्ड्स के साथ इंटरैक्ट करने में सक्षम बनाता है। अपनी मुख्य डेटाबेस क्षमताओं से परे, यह प्रोजेक्ट ब्रांडेड बाहरी पोर्टल्स, आंतरिक बिज़नेस एप्लिकेशन और इंटरैक्टिव डैशबोर्ड बनाने के लिए टूल्स प्रदान करता है। इसमें बिज़नेस प्रोसेस ऑटोमेशन के लिए एक इवेंट-ड्रिवन ऑटोमेशन इंजन शामिल है और यह वेबहुक्स, WebSocket इवेंट स्ट्रीमिंग और थर्ड-पार्टी डेटा सिंक्रोनाइज़ेशन सहित API एकीकरण की एक विस्तृत श्रृंखला को सपोर्ट करता है। यह सॉफ्टवेयर डेटा संप्रभुता और सुरक्षा सुनिश्चित करने के लिए प्राइवेट इंफ्रास्ट्रक्चर होस्टिंग और कंटेनराइज़्ड डिप्लॉयमेंट के लिए डिज़ाइन किया गया है।

    Creates complete technical workflows and database architectures using natural language processing.

    Python
    GitHub पर देखें↗5,085
  • nateraw/stable-diffusion-videosnateraw का अवतार

    nateraw/stable-diffusion-videos

    4,695GitHub पर देखें↗

    यह प्रोजेक्ट एक Stable Diffusion वीडियो जनरेटर है जो जेनरेटिव मॉडल के लेटेंट स्पेस के भीतर टेक्स्ट प्रॉम्प्ट्स के बीच इंटरपोलेट करके चलती हुई इमेजेस बनाता है। यह AI वीडियो जनरेशन और लेटेंट स्पेस इंटरपोलेशन के लिए एक टूल के रूप में कार्य करता है, जो वर्णनात्मक टेक्स्ट को विजुअल सीक्वेंस में बदलता है। यह सिस्टम विशेष रूप से ऑडियो फाइल की बीट और रिदम के साथ इमेज इंटरपोलेशन की दर को सिंक्रोनाइज़ करके ऑडियो-रिएक्टिव विजुअल्स को सक्षम बनाता है। यह इन सीक्वेंस को मॉर्फिंग वीडियो जनरेशन के माध्यम से तैयार करता है, जो अलग-अलग टेक्स्ट प्रॉम्प्ट्स के बीच सुचारू रूप से ट्रांज़िशन करता है। इस प्रोजेक्ट में एक ग्राफिकल यूजर इंटरफेस शामिल है जो टेक्स्ट-टू-वीडियो वर्कफ़्लो को मैनेज करने के लिए एक वेब-आधारित कंट्रोल इंटरफेस प्रदान करता है। यह मैनुअल पाइपलाइन कोड लिखे बिना जेनरेटिव प्रक्रिया के ऑर्केस्ट्रेशन की अनुमति देता है।

    Provides a text-to-video generation workflow that transforms descriptive prompts into moving imagery.

    Python
    GitHub पर देखें↗4,695
  • aidc-ai/comfyui-copilotAIDC-AI का अवतार

    AIDC-AI/ComfyUI-Copilot

    4,599GitHub पर देखें↗

    ComfyUI-Copilot is an AI-powered assistant and design tool for creating, optimizing, and automating node-based generative AI workflows within ComfyUI. It functions as an LLM workflow orchestrator that provides natural language guidance to help design and refine complex image and video generation pipelines. The system uses large language models to translate natural language prompts into structured node-based graph configurations. It incorporates state-aware prompt engineering and dynamic node mapping to ensure generated sequences match the specific requirements of the target AI nodes. The too

    Creates and optimizes complex node-based sequences for automating generative image and video production.

    TypeScriptagentaicomfy-ui
    GitHub पर देखें↗4,599
  • tencent-hunyuan/hunyuanvideo-1.5Tencent-Hunyuan का अवतार

    Tencent-Hunyuan/HunyuanVideo-1.5

    4,440GitHub पर देखें↗

    HunyuanVideo-1.5 is a video generation foundation model and text-to-video diffusion framework. It utilizes a latent video diffusion model and a spatio-temporal transformer architecture to generate high-definition video sequences from text descriptions and images. The project enables cinematic camera control for directing pans and tilts and provides image-to-video animation capabilities. It supports visual style adaptation through low-rank adaptation tuning and uses a language model for prompt refinement to improve visual alignment. The model covers high-resolution video upscaling via a super

    Enriches short user prompts into detailed descriptions using a large language model for better visual alignment.

    Pythonimage-to-videotext-to-videovideo-generation
    GitHub पर देखें↗4,440
  • showlab/tune-a-videoshowlab का अवतार

    showlab/Tune-A-Video

    4,364GitHub पर देखें↗

    Tune-A-Video is a text-to-video diffusion framework designed to convert pretrained text-to-image diffusion models into video generators. It utilizes a spatio-temporal attention mechanism and single text-video pair training to enable the synthesis of moving sequences from text prompts. The project provides tools for one-shot video personalization, allowing a model to be tuned on a single reference video to preserve specific characters or artistic styles across new generations. It also functions as a video editor that modifies subjects, backgrounds, and styles through noise-sampling prompt guid

    Synthesizes moving video sequences from textual prompts by repurposing image diffusion models.

    Python
    GitHub पर देखें↗4,364
  • futantan/opengptfutantan का अवतार

    futantan/OpenGpt

    3,902GitHub पर देखें↗

    OpenGpt एक एजेंट ऑर्केस्ट्रेशन प्लेटफ़ॉर्म और मल्टीमॉडल इंटरफ़ेस है जिसे विशेष AI पर्सना बनाने और तैनात करने के लिए डिज़ाइन किया गया है। यह उपयोगकर्ताओं को पेशेवर, रचनात्मक और तकनीकी वर्कफ़्लो को स्वचालित करने के लिए कस्टम सिस्टम प्रॉम्प्ट और व्यवहार संबंधी बाधाओं के साथ कार्य-उन्मुख एजेंट बनाने की अनुमति देता है। इस प्रोजेक्ट में एक प्रॉम्प्ट इंजीनियरिंग वर्कफ़्लो है जो मॉडल की सटीकता में सुधार करने के लिए सरल उपयोगकर्ता इनपुट को संरचित निर्देशों में बदलता है। यह चैट इंटरफ़ेस से वेक्टर डेटाबेस को जोड़कर रिट्रीवल-ऑगमेंटेड जनरेशन को एकीकृत करता है, जिससे निजी डेटासेट से संदर्भ-जागरूक प्रतिक्रियाएं सक्षम होती हैं। यह प्लेटफ़ॉर्म PDF और ऑडियो के लिए मल्टीमॉडल डेटा पार्सिंग, व्यक्तिगत कुंजियों के माध्यम से मल्टी-प्रोवाइडर API प्रबंधन, और पेशेवर डॉक्यूमेंट्स, कार्यात्मक कोड और विज़ुअल प्रॉम्प्ट जैसे विविध कंटेंट प्रकारों के निर्माण सहित क्षमताओं की एक विस्तृत श्रृंखला को कवर करता है। इसमें कंटेंट विश्लेषण, अनुवाद सेवाएं और Google OAuth के माध्यम से पहचान प्रबंधन के लिए टूल भी शामिल हैं।

    Rewrites basic text descriptions into detailed prompt ideas for better visual alignment in image generators.

    TypeScript
    GitHub पर देखें↗3,902
पिछला12अगला
  1. Home
  2. Artificial Intelligence & ML
  3. Generative AI Resources
  4. Diffusion & Visual Synthesis Models
  5. Generative AI Workflows

सब-टैग एक्सप्लोर करें

  • Example GalleriesCurated collections of generative workflows that showcase specific artistic techniques or model combinations. **Distinct from Generative AI Workflows:** Focuses on the curation and discovery of examples rather than the definition of the workflow itself.
  • Image Editing Workflows1 सब-टैगWorkflows specifically designed for image-to-image manipulation and editing.
  • Technical Workflow GenerationGenerative AI that produces sequences of automated technical steps and data structures. **Distinct from Generative AI Workflows:** Focuses on functional business workflows and database setups rather than iterative content or image refinement.
  • Text-to-Video Generation1 सब-टैगWorkflows that synthesize video sequences from textual prompts.
  • Workflow Consolidation ToolsUtilities that merge discrete functional nodes into unified components to reduce graph complexity. **Distinct from Image Editing Workflows:** Distinct from Image Editing Workflows by focusing on the architectural simplification of the node graph rather than the editing process itself.