awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoAcerca deCómo clasificamosPrensaServidor MCP
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
browser-use avatar

browser-use/video-use

0
View on GitHub↗
9,743 estrellas·1,404 forks·Python·MIT·8 vistas

Video Use

This project is an AI video post-production suite that uses large language models and programmatic tools to automate editing, transcription, and subtitle generation. It functions as an AI editing agent that translates natural language instructions into shell commands, providing a programmatic interface for manipulating media via FFmpeg.

The toolkit includes a motion graphics engine that generates technical animations and visual overlays through code-driven rendering and mathematical definitions. It distinguishes itself by combining an AI-powered transcriber for word-level timestamps with an automated system for removing filler words, false starts, and dead space.

The system covers a broad range of post-production capabilities, including audio-based video cutting, cinematic color grading through filter chains, and the integration of synthetic AI voiceovers. It also provides observability tools such as timeline visualization through composite filmstrips and waveforms, as well as self-evaluation loops to validate rendered output for visual jumps or audio pops.

Session data and editing history are persisted in text files to maintain project continuity across different execution contexts.

Features

  • Video Editing Agents - Implements an AI agent that translates natural language instructions into actionable shell commands for automated video editing.
  • AI Video Editing Automation - Uses natural language and coding agents to automate cutting, color grading, and assembling video projects.
  • Audio Transcription - Converts spoken audio into text transcripts with word-level timestamps and speaker identification.
  • Automated Video Transcribers - Converts audio to word-level timestamps for precise cutting and automated subtitle burning.
  • Natural Language Automation - Translates natural language instructions into shell commands to automate video processing tasks.
  • Automated Subtitle Generators - Creates timed text overlays with configurable chunking and positioning using an automated transcription and embedding workflow.
  • Burned-In Subtitle Rendering - Overlays customizable text chunks directly onto the video based on transcribed audio timestamps.
  • Animation & Motion Graphics - Produces programmatic animations and overlays for diagrams, UI mockups, and kinetic typography.
  • Programmatic Motion Graphics - Ships a motion graphics engine that generates technical animations and visual overlays through code-driven rendering.
  • Animation Sequence Renderers - Processes programmatic motion code into sequenced frames for final video files.
  • Transcription-Driven Slicing - Determines precise cut points by mapping word-level transcriptions and silence gaps to video frame boundaries.
  • FFmpeg Wrappers - Provides a programmatic interface for manipulating media using FFmpeg filters and shell-based processing.
  • Filler Audio Removal - Identifies and cuts out filler words, false starts, and dead space to tighten video pacing.
  • Text-Driven Video Editing - Translates natural language instructions into shell commands to automate video cutting and color grading.
  • Programmatic Animations - Creates programmatic animations and visual effects using mathematical definitions and code-driven rendering.
  • Audio-Driven Cutting - Identifies optimal cut points based on word boundaries and silence gaps to maintain natural pacing.
  • AI Video Post-Production Suites - Provides a toolkit for automating transcription, filler word removal, and subtitle generation through large language models.
  • Text-to-Speech - Adds synthetic speech and narrated audio tracks to video content using text-to-speech services.
  • Color Grading - Applies cinematic visual styles by passing video segments through sequential image adjustment filters.
  • Final Episode Assembly - Processes extracted segments and adds overlays and subtitles to produce a final cohesive video export.
  • Technical Animation Generation - Produces mathematical and technical animated videos from text prompts by generating code and rendering scenes.
  • Text-to-Speech Engines - Creates audio narrations for video content by integrating with external text-to-speech services.
  • Analysis-Driven Editing Guides - Processes audio transcripts to generate visual composite filmstrips that guide editing decisions without processing raw frames.
  • Post-Production Environments - Manages the end-to-end process of removing filler audio, applying color grades, and exporting final renders.
  • Timeline Visualization Tools - Generates low-resolution visual and waveform previews to guide editing decisions.
  • Video Timeline Visualizers - Generates filmstrip and waveform images to assist in making precise editing and cutting decisions.
  • Visual Overlays - Provides the ability to generate visual overlays using specialized animation tools to enhance video content.
  • Render Quality Validation - Executes a self-evaluation loop that checks cut boundaries for visual jumps or audio pops before finalizing.

Historial de estrellas

Gráfico del historial de estrellas de browser-use/video-useGráfico del historial de estrellas de browser-use/video-use

Búsqueda con IA

Explora más repositorios increíbles

Describe lo que necesitas en lenguaje sencillo: la IA clasifica miles de proyectos open-source curados por relevancia.

Start searching with AI

Preguntas frecuentes

¿Qué hace browser-use/video-use?

This project is an AI video post-production suite that uses large language models and programmatic tools to automate editing, transcription, and subtitle generation. It functions as an AI editing agent that translates natural language instructions into shell commands, providing a programmatic interface for manipulating media via FFmpeg.

¿Cuáles son las características principales de browser-use/video-use?

Las características principales de browser-use/video-use son: Video Editing Agents, AI Video Editing Automation, Audio Transcription, Automated Video Transcribers, Natural Language Automation, Automated Subtitle Generators, Burned-In Subtitle Rendering, Animation & Motion Graphics.

¿Qué alternativas de código abierto existen para browser-use/video-use?

Las alternativas de código abierto para browser-use/video-use incluyen: yils-lin/short-video-factory — Short video factory is a local AI content generator and automated video editing tool. It provides a production… wyattblue/auto-editor — Auto-editor is a command-line automated video editor that uses FFmpeg to remove silence and inactive footage from… linyqh/narratoai — NarratoAI is an automated video production pipeline that uses large language models to generate scripts, voiceovers,… buxuku/smartsub — SmartSub is a cross-platform desktop application for AI-driven video transcription and subtitle generation. It… carykh/jumpcutter — Jumpcutter is an audio-based video cutter and automatic editor designed to eliminate dead air from video files. It… rayventura/shortgpt — ShortGPT is an automated short-form video creation framework that combines large language model-driven scripting with…

Alternativas open-source a Video Use

Proyectos open-source similares, clasificados según cuántas características comparten con Video Use.
  • yils-lin/short-video-factoryAvatar de YILS-LIN

    YILS-LIN/short-video-factory

    3,428Ver en GitHub↗

    Short video factory is a local AI content generator and automated video editing tool. It provides a production pipeline that uses large language models to transform text prompts into marketing scripts and rendered short-form videos. The system is designed for local-first execution, running all processing and asset management on the host machine to maintain data privacy. It distinguishes itself through a batch-processing workflow that can sequentially execute copywriting and rendering for multiple items using predefined presets. The software covers a broad range of media capabilities, includi

    TypeScriptaiautomaticautomation
    Ver en GitHub↗3,428
  • wyattblue/auto-editorAvatar de WyattBlue

    WyattBlue/auto-editor

    4,460Ver en GitHub↗

    Auto-editor is a command-line automated video editor that uses FFmpeg to remove silence and inactive footage from video files. It functions as a processing suite with specialized cut generators that identify segments to trim based on loudness thresholds, motion analysis, and speech-to-text transcription. The tool distinguishes itself by offering a flexible post-production workflow, allowing users to export automated cut timelines as XML or JSON files for use in professional non-linear editing software. Beyond simple deletion, it can perform dynamic playback adjustments, such as increasing the

    Nimaudioaudio-editingaudio-processing
    Ver en GitHub↗4,460
  • buxuku/smartsubAvatar de buxuku

    buxuku/SmartSub

    4,056Ver en GitHub↗

    SmartSub is a cross-platform desktop application for AI-driven video transcription and subtitle generation. It converts audio and video files into text subtitles using local AI models and incorporates hardware acceleration to increase processing speed. The tool features a subtitle translator that leverages large language models, such as OpenAI and DeepSeek, to convert subtitles between different languages. It includes a visual editor for proofreading and polishing transcribed text, paired with a video preview for frame-accurate synchronization. The software supports batch processing of multi

    TypeScriptdeepseekelectronnodejs
    Ver en GitHub↗4,056
  • linyqh/narratoaiAvatar de linyqh

    linyqh/NarratoAI

    8,091Ver en GitHub↗

    NarratoAI is an automated video production pipeline that uses large language models to generate scripts, voiceovers, and edited video commentary. It functions as a combined scriptwriter, voiceover generator, and video editor to streamline the creation of movie and television commentary content. The system automates the production workflow by converting input data into structured narrative scripts, synthesizing artificial speech for narration, and programmatically assembling video clips based on script timestamps. It also converts spoken audio from video files into written text for subtitles a

    Pythonaiagentaiopsgemini-api
    Ver en GitHub↗8,091
  • Ver las 30 alternativas a Video Use→