awesome-repositories.com
المدونة
MCP
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعخادم MCPحولكيفية ترتيب النتائجالصحافة
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

9 مستودعات

Awesome GitHub RepositoriesGenerative Camera Controls

Controls for simulating camera movements and perspectives during the synthesis of video sequences from images.

Distinct from Camera Interaction Controllers: Candidates focus on hardware camera parameters or UI interaction, not generative video perspective control.

Explore 9 awesome GitHub repositories matching graphics & multimedia · Generative Camera Controls. Refine with filters or upvote what's useful.

Awesome Generative Camera Controls GitHub Repositories

اعثر على أفضل المستودعات باستخدام الذكاء الاصطناعي.سنبحث عن أفضل المستودعات المطابقة باستخدام الذكاء الاصطناعي.
  • comfyanonymous/comfyuiالصورة الرمزية لـ comfyanonymous

    comfyanonymous/ComfyUI

    117,322عرض على GitHub↗

    ComfyUI is a modular generative AI workflow orchestrator and node-based GUI for designing and executing complex diffusion model pipelines. It functions as both a visual interface for building generative logic graphs and a programmable backend API that exposes diffusion model operations for external integration. The system distinguishes itself through a graph-based execution model that supports differential workflow execution, re-running only modified nodes to reduce computation. It features dynamic model offloading to manage memory between system RAM and GPU VRAM and utilizes metadata-embedde

    Produces video sequences from images while applying specific camera movement and perspective controls.

    Python
    عرض على GitHub↗117,322
  • anil-matcha/open-higgsfield-aiالصورة الرمزية لـ Anil-matcha

    Anil-matcha/Open-Higgsfield-AI

    20,529عرض على GitHub↗

    Open-Higgsfield-AI is a generative AI content studio and visual workflow orchestrator. It provides a unified interface for creating photorealistic images and videos, utilizing a node-based editor to chain multiple image, video, and audio models into automated content pipelines. The system functions as an AI video animation tool and local GPU inference engine, allowing users to run generative models on local hardware or remote servers. It includes specialized capabilities for audio-driven lip synchronization and cinematic camera controls to adjust virtual lens and focal settings. The platform

    Adjusts virtual lens and focal settings to simulate photorealistic cinematic camera movements.

    JavaScriptai-art-generatorai-image-generationai-video-generation
    عرض على GitHub↗20,529
  • guoyww/animatediffالصورة الرمزية لـ guoyww

    guoyww/AnimateDiff

    12,144عرض على GitHub↗

    AnimateDiff is a latent diffusion video generator and text-to-video diffusion framework. It converts existing text-to-image diffusion models into animation generators by applying specialized motion modules, allowing for the creation of video sequences without modifying the original base model. The project provides an image-to-video animation framework that uses sparse RGB images, sketches, or structural keyframe constraints to guide generation. It further distinguishes itself with a motion adapter system that injects cinematic camera movements, such as zooming, panning, and tilting, into anim

    Provides controls for simulating cinematic camera movements like zooming and panning during video synthesis.

    Python
    عرض على GitHub↗12,144
  • ashawkey/stable-dreamfusionالصورة الرمزية لـ ashawkey

    ashawkey/stable-dreamfusion

    8,841عرض على GitHub↗

    This project is a diffusion-based 3D generator and image-to-3D reconstruction system. It translates natural language descriptions or two-dimensional images into three-dimensional assets using neural radiance fields and diffusion models. The system utilizes score-distillation sampling and diffusion-based guidance to refine 3D shapes without requiring 3D training data. It includes specialized tools for transforming neural representations into exportable meshes with texture and material data, as well as a pipeline for iterative optimization of geometry and textures. The project covers a broad r

    Provides controls for defining overhead and front angle borders to specify the perspective of generated 3D assets.

    Python
    عرض على GitHub↗8,841
  • nvlabs/sanaالصورة الرمزية لـ NVlabs

    NVlabs/Sana

    8,310عرض على GitHub↗

    Sana is a framework for high-resolution image and video synthesis based on a linear diffusion transformer. It provides a toolkit for the training, fine-tuning, and execution of text-to-image and text-to-video models, as well as a video generative world model capable of simulating physical environments with precise spatial control. The project is distinguished by its use of linear complexity layers to handle high resolutions and its support for long-form, minute-length video generation in real time. It implements a two-stage inference paradigm that separates structural generation from visual t

    Enables precise per-frame camera trajectory control using domain-specific action strings and matrices.

    Python
    عرض على GitHub↗8,310
  • gyroflow/gyroflowالصورة الرمزية لـ gyroflow

    gyroflow/gyroflow

    8,256عرض على GitHub↗

    Gyroflow is a gyroscope video stabilization software and IMU telemetry processor designed to remove camera shake from video files. It functions as a hardware-accelerated video renderer and lens calibration tool, utilizing embedded or external gyroscope and accelerometer data to perform pixel-level stabilization. The system is distinguished by its ability to integrate with professional non-linear video editing software via plugins, allowing stabilization to be applied directly to timelines without transcoding original footage. It supports diverse telemetry ingestion from camera brands, flight

    Translates gyroscope data into virtual 3D camera movements for use in compositing environments.

    Rustfpvgoprogpu
    عرض على GitHub↗8,256
  • tencent-hunyuan/hunyuanvideo-1.5الصورة الرمزية لـ Tencent-Hunyuan

    Tencent-Hunyuan/HunyuanVideo-1.5

    4,440عرض على GitHub↗

    HunyuanVideo-1.5 is a video generation foundation model and text-to-video diffusion framework. It utilizes a latent video diffusion model and a spatio-temporal transformer architecture to generate high-definition video sequences from text descriptions and images. The project enables cinematic camera control for directing pans and tilts and provides image-to-video animation capabilities. It supports visual style adaptation through low-rank adaptation tuning and uses a language model for prompt refinement to improve visual alignment. The model covers high-resolution video upscaling via a super

    Directs camera movements like pans, tilts, and orbits in generated videos through cinematography keywords.

    Pythonimage-to-videotext-to-videovideo-generation
    عرض على GitHub↗4,440
  • lightricks/comfyui-ltxvideoالصورة الرمزية لـ Lightricks

    Lightricks/ComfyUI-LTXVideo

    3,840عرض على GitHub↗

    ComfyUI-LTXVideo is a generative framework and ComfyUI custom node extension for synthesizing high-fidelity video. It utilizes a latent diffusion and transformer-based system to create cinematic clips from text, image, and audio inputs, providing a modular interface for precise control over subject behavior and temporal consistency. The tool distinguishes itself with production-grade capabilities, including the generation of High Dynamic Range video in linear formats such as ARRI LogC3. It supports multimodal synchronization for audio-driven animation and lip-syncing, and allows for the creat

    Directs camera behavior and character movement using depth-aware controls and pose-driven input.

    Pythoncomfyuidiffusion-modelsdit
    عرض على GitHub↗3,840
  • robbyant/lingbot-worldالصورة الرمزية لـ Robbyant

    Robbyant/lingbot-world

    2,915عرض على GitHub↗

    Lingbot-world is an interactive world simulator and framework for generating high-fidelity video environments from text and image prompts. It functions as a video generation system designed to create controllable simulations for applications such as robotics learning and gaming. The project includes a video motion controller that directs camera and object movement using transformation matrices and action strings. It utilizes a quantized inference engine to reduce memory usage and accelerate the generation of video sequences. The system covers a range of optimization techniques, including fou

    Governs camera and object movement in generated videos by applying transformation matrices to latent spatial representations.

    Pythonaigcimage-to-videolingbot-world
    عرض على GitHub↗2,915
  1. Home
  2. Graphics & Multimedia
  3. Generative Camera Controls

استكشف الوسوم الفرعية

  • Camera Motion ReplicationTranslating sensor telemetry into 3D camera paths for virtual environments. **Distinct from Generative Camera Controls:** Distinct from Generative Camera Controls by using actual telemetry data to replicate real-world movement rather than synthesizing new perspectives.
  • Face-Driven Camera MotionVirtual camera movements triggered by the detection of facial positions in a video stream. **Distinct from Generative Camera Controls:** Distinct from Generative Camera Controls as it uses real-time computer vision input rather than synthetic simulation.
  • Generative Camera ControlsTools for synthesizing and directing camera movement and character behavior in generative video. **Distinct from Face-Driven Camera Motion:** Covers synthetic camera path generation rather than real-time computer vision trigger-based motion.
  • Prompt-Based Camera Controls2 وسوم فرعيةDirects camera motion by parsing cinematography keywords in the text prompt to influence latent diffusion dynamics. **Distinct from Generative Camera Controls:** Distinct from Generative Camera Controls: focuses on controlling camera movement through text prompt keywords rather than general simulation of camera perspectives.