awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
SamurAIGPT avatar

SamurAIGPT/AI-Youtube-Shorts-Generator

0
View on GitHub↗
3,037 stars·528 forks·Python·mit·19 viewswww.vadoo.tv/ai-youtube-shorts-generator↗

AI Youtube Shorts Generator

This project is an AI-driven suite of tools designed to repurpose long-form video content into short-form clips. It integrates a speech-to-text engine for automated transcription, a highlighting system that ranks engaging segments based on emotional hooks, and a video processor that converts horizontal footage into vertical formats.

The system distinguishes itself through intelligent video cropping that utilizes face tracking and motion smoothing to keep subjects centered. It also employs an analysis system to extract viral highlights by scoring segments for engagement and practical value.

The software covers a broad range of media processing capabilities, including aspect ratio adjustment, short-form clip generation, and the export of transcripts and viral scores to structured JSON files for automation workflows.

Features

  • Short-Form Video Generation - Provides AI-driven generation of short-form social video clips from long-form source content.
  • Video Clip Extraction - Transforms long-form videos into multiple short, vertical clips optimized for social media distribution.
  • Audio and Video File Transcription - Generates timestamped text transcripts from video files using cloud or local speech-to-text engines.
  • Face Tracking - Uses facial coordinate tracking to keep the subject centered during the vertical cropping process.
  • Semantic Segment Scoring - Analyzes timestamped text for emotional hooks and high-value keywords to score and rank potential viral highlights.
  • Speech to Text Transcription - Provides a speech-to-text engine that generates timestamped transcripts to identify key highlights.
  • Automated Video Transcribers - Converts video audio into time-synced text transcripts to identify key quotes and organize editing.
  • Highlight Detection - Automatically identifies high-energy segments and engaging quotes from long videos to generate shareable clips.
  • Dynamic Cropping Windows - Calculates a dynamic cropping window to transform widescreen footage into vertical formats without losing the primary subject.
  • Intelligent Cropping Engines - Automates the conversion of horizontal videos to vertical formats using facial analysis and tracking.
  • Speech-to-Text Pipelines - Implements an automated pipeline that converts video audio streams into timestamped text transcripts.
  • Video Content Repurposing - Uses AI to identify viral moments in long videos and transform them into shareable short-form segments.
  • Content-Aware Vertical Cropping - Implements face tracking and motion smoothing to automatically center subjects when converting horizontal footage to vertical clips.
  • Subject-Aware Auto-Cropping - Converts horizontal footage to a vertical aspect ratio using face tracking and motion smoothing.
  • Social Media Content Creation - Generates high-engagement video clips tailored specifically to social media platform dimensions.
  • Asynchronous Task Processing - Offloads computationally heavy video transcoding and cropping tasks to background workers to maintain application responsiveness.
  • Content Ranking - Ranks video segments based on hooks and emotional peaks to isolate the most engaging moments.
  • Video Aspect Ratio Configurations - Modifies video dimensions to fit specific platform requirements like social media vertical formats.

Star history

Star history chart for samuraigpt/ai-youtube-shorts-generatorStar history chart for samuraigpt/ai-youtube-shorts-generator

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Projects sharing features with AI Youtube Shorts Generator

These projects share indexed features with AI Youtube Shorts Generator. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • linyqh/narratoailinyqh avatar

    linyqh/NarratoAI

    8,091View on GitHub↗

    NarratoAI is an automated video production pipeline that uses large language models to generate scripts, voiceovers, and edited video commentary. It functions as a combined scriptwriter, voiceover generator, and video editor to streamline the creation of movie and television commentary content. The system automates the production workflow by converting input data into structured narrative scripts, synthesizing artificial speech for narration, and programmatically assembling video clips based on script timestamps. It also converts spoken audio from video files into written text for subtitles a

    Pythonaiagentaiopsgemini-api
    View on GitHub↗8,091
  • modelscope/funclipmodelscope avatar

    modelscope/FunClip

    5,850View on GitHub↗

    FunClip is an open-source tool that transcribes speech from video files and clips segments based on text, speaker, or AI analysis. It combines speech recognition with speaker diarization, audio event detection, and visual content understanding to identify and extract relevant portions of a video. The tool distinguishes itself through several integrated capabilities. It supports hotword-weighted speech recognition, which improves transcription accuracy for specific terms like names or jargon by boosting their probability during decoding. A large language model can interpret the transcribed tex

    Pythonai-toolsai-video-editingasr
    View on GitHub↗5,850
  • cmusphinx/pocketsphinxcmusphinx avatar

    cmusphinx/pocketsphinx

    4,276View on GitHub↗

    PocketSphinx is an offline speech recognition engine that converts raw audio from files or live microphone streams into written text without requiring a network connection. It functions as a speech-to-text library, a real-time transcription engine, and a voice command processor, capable of detecting and transcribing spoken commands from continuous audio streams with configurable acoustic and language models. The engine uses weighted finite-state transducers to represent acoustic, phonetic, and language models as a single search graph for efficient decoding. It employs fixed-point acoustic mod

    Ccpythonspeech-recognition
    View on GitHub↗4,276
  • pluja/whishperpluja avatar

    pluja/whishper

    2,920View on GitHub↗

    Whishper is a graphical user interface for transcribing audio and video files into text using the Whisper model. It serves as a speech-to-text tool and subtitle file generator that converts spoken content into editable text and timed subtitle formats. The project features an integrated transcription and translation interface, allowing users to refine automated results and convert transcribed text into different languages. It includes a visual editor for correcting speech recognition errors, adjusting segment timecodes, and performing bilingual translation reviews. The system handles the full

    Svelteaiaudio-to-textgolang
    View on GitHub↗2,920
Compare all 30 related projects→

Frequently asked questions

What does samuraigpt/ai-youtube-shorts-generator do?

This project is an AI-driven suite of tools designed to repurpose long-form video content into short-form clips. It integrates a speech-to-text engine for automated transcription, a highlighting system that ranks engaging segments based on emotional hooks, and a video processor that converts horizontal footage into vertical formats.

What are the main features of samuraigpt/ai-youtube-shorts-generator?

The main features of samuraigpt/ai-youtube-shorts-generator are: Short-Form Video Generation, Video Clip Extraction, Audio and Video File Transcription, Face Tracking, Semantic Segment Scoring, Speech to Text Transcription, Automated Video Transcribers, Highlight Detection.

Which projects share features with samuraigpt/ai-youtube-shorts-generator?

Projects with overlapping indexed features include: linyqh/narratoai — NarratoAI is an automated video production pipeline that uses large language models to generate scripts, voiceovers,… modelscope/funclip — FunClip is an open-source tool that transcribes speech from video files and clips segments based on text, speaker, or… timerring/bilive — Bilive is a multimodal AI video pipeline and live stream recording tool designed to capture real-time broadcasts and… cmusphinx/pocketsphinx — PocketSphinx is an offline speech recognition engine that converts raw audio from files or live microphone streams… pluja/whishper — Whishper is a graphical user interface for transcribing audio and video files into text using the Whisper model. It… adithya-s-k/omniparse — Omniparse is a multimodal content parser and generative AI ingestion engine designed to convert documents, images, and…