awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
agermanidis avatar

agermanidis/autosub

0
View on GitHub↗
4,197 stars·1,635 forks·Python·MIT·7 views

Autosub

Autosub is a command-line media processor and automatic subtitle generator that converts audio streams from video and audio files into timed text overlays. It functions as an AI speech-to-text converter that uses OpenAI Whisper to generate synchronized subtitles.

The tool includes a language translation pipeline to convert transcribed speech into target languages, enabling multilingual video captioning. It manages the process from audio-stream extraction to the serialization of final subtitle files for local storage.

The system covers audio-to-text transcription, time-stamped text mapping, and terminal-based media processing.

Features

  • Automated Subtitle Generators - Provides an automated workflow that combines speech recognition and subtitle embedding for video files.
  • Language Translation Services - Passes transcribed text through a translation service to support multilingual subtitles.
  • Speech Recognition APIs - Integrates with speech recognition APIs to convert audio data into timed text segments.
  • Speech-to-Text Converters - Converts audio streams into text and supports translating transcriptions into different languages.
  • Whisper-Based Engines - Utilizes the OpenAI Whisper architecture to generate synchronized subtitles from audio and video files.
  • Speech to Text Transcription - Converts audio from media files into processed text with precise timestamps.
  • Automated Video Subtitling - Implements an AI-driven pipeline that combines transcription and translation to generate timed captions.
  • Timed Text Tracks - Synchronizes transcribed text with specific start and end timestamps for video playback.
  • Subtitle Translation - Translates transcribed speech into target languages specifically for video subtitles and captions.
  • Command-Line Media Processors - Provides a terminal-based application for automating the creation of subtitle files from various video formats.
  • Media Processing - Provides a terminal interface for manipulating and processing audio and video media files.
  • Multilingual Captioning - Generates and times subtitles in multiple languages for international audiences.
  • Machine Learning Projects - Automated speech recognition and subtitle generation for video files.
  • Media Automation - Command-line utility for auto-generating video subtitles.

Star history

Star history chart for agermanidis/autosubStar history chart for agermanidis/autosub

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does agermanidis/autosub do?

Autosub is a command-line media processor and automatic subtitle generator that converts audio streams from video and audio files into timed text overlays. It functions as an AI speech-to-text converter that uses OpenAI Whisper to generate synchronized subtitles.

What are the main features of agermanidis/autosub?

The main features of agermanidis/autosub are: Automated Subtitle Generators, Language Translation Services, Speech Recognition APIs, Speech-to-Text Converters, Whisper-Based Engines, Speech to Text Transcription, Automated Video Subtitling, Timed Text Tracks.

What are some open-source alternatives to agermanidis/autosub?

Open-source alternatives to agermanidis/autosub include: linyqh/narratoai — NarratoAI is an automated video production pipeline that uses large language models to generate scripts, voiceovers,… tmoroney/auto-subs — Auto-subs is an AI transcription and automatic captioning tool that converts spoken audio from video files into… umlx5h/llplayer — LLPlayer is a language learning media player and AI subtitle generator that integrates large language models for… wxbool/video-srt-windows — This is a Windows application for automatic speech recognition that transcribes spoken audio from video files into… abus-aikorea/voice-pro — Voice Pro is a comprehensive speech and audio processing toolkit that combines text-to-speech synthesis, voice… weifeng2333/videocaptioner — VideoCaptioner is an automated tool designed to generate and embed time-synchronized subtitles into video files. By…

Open-source alternatives to Autosub

Similar open-source projects, ranked by how many features they share with Autosub.
  • linyqh/narratoailinyqh avatar

    linyqh/NarratoAI

    8,091View on GitHub↗

    NarratoAI is an automated video production pipeline that uses large language models to generate scripts, voiceovers, and edited video commentary. It functions as a combined scriptwriter, voiceover generator, and video editor to streamline the creation of movie and television commentary content. The system automates the production workflow by converting input data into structured narrative scripts, synthesizing artificial speech for narration, and programmatically assembling video clips based on script timestamps. It also converts spoken audio from video files into written text for subtitles a

    Pythonaiagentaiopsgemini-api
    View on GitHub↗8,091
  • umlx5h/llplayerumlx5h avatar

    umlx5h/LLPlayer

    3,110View on GitHub↗

    LLPlayer is a language learning media player and AI subtitle generator that integrates large language models for real-time audio transcription and translation. It functions as an LLM-integrated video player and SRT transcription tool, utilizing local or remote AI models to generate text subtitles from audio and video streams. The project distinguishes itself through a contextual translation workflow that sends preceding subtitle lines to language models to maintain conversational flow and sentence structure. It also includes an optical character recognition system to convert bitmap-based subt

    C#asrcsharpflyleaf
    View on GitHub↗3,110
  • tmoroney/auto-substmoroney avatar

    tmoroney/auto-subs

    2,851View on GitHub↗

    Auto-subs is an AI transcription and automatic captioning tool that converts spoken audio from video files into synchronized subtitles. It functions as a subtitle generator and a transcription bridge, enabling the conversion of speech to text with automatic speaker identification and multi-language translation support. The software prioritizes data privacy by utilizing on-device AI inference to process audio and video files locally on the user's hardware. It distinguishes itself by offering deep integration with professional video editing workflows, allowing users to export timing and transcr

    TypeScriptaidavincidavinci-resolve
    View on GitHub↗2,851
  • wxbool/video-srt-windowswxbool avatar

    wxbool/video-srt-windows

    5,037View on GitHub↗

    This is a Windows application for automatic speech recognition that transcribes spoken audio from video files into timestamped SRT subtitle files. It serves as a subtitle generator and translation tool that converts media speech into synchronized text. The software functions as a batch media transcriber, allowing the simultaneous processing of multiple audio and video files to generate subtitles in bulk. It includes a translation workflow to convert transcriptions between different languages for the creation of bilingual or localized files. The system also provides text refinement capabiliti

    Goffmpeggogolang
    View on GitHub↗5,037
  • See all 30 alternatives to Autosub→