awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
tmoroney avatar

tmoroney/auto-subs

0
View on GitHub↗
2,851 stars·160 forks·TypeScript·mit·25 viewstom-moroney.com/auto-subs↗

Auto Subs

Auto-subs is an AI transcription and automatic captioning tool that converts spoken audio from video files into synchronized subtitles. It functions as a subtitle generator and a transcription bridge, enabling the conversion of speech to text with automatic speaker identification and multi-language translation support.

The software prioritizes data privacy by utilizing on-device AI inference to process audio and video files locally on the user's hardware. It distinguishes itself by offering deep integration with professional video editing workflows, allowing users to export timing and transcription data directly into external editing software for precise alignment.

The system provides comprehensive capabilities for speaker diarization, visual style customization for captions, and the ability to bake stylized overlays directly into video frames. It supports both the export of standardized SRT files and the generation of subtitled video exports.

Translation services are integrated for both raw audio content and generated transcription text across a wide range of supported languages.

Features

  • On-Device Inference - Runs transcription models locally on the user's hardware to process audio and video files privately.
  • Audio and Video File Transcription - Extracts speech from media files offline using on-device processing to produce subtitles and timestamps.
  • Subtitle Translation - Translates spoken audio or existing transcriptions into different languages to reach global audiences.
  • Audio Transcriptions - Converts spoken audio from video files into text with automatic speaker identification and translation support.
  • Local Speech-to-Text - Generates text from audio files using on-device processing to ensure sensitive data never leaves the machine.
  • Neural Machine Translation - Uses pre-trained language models to convert transcribed text or spoken audio between different natural languages.
  • Speaker Diarization - Implements speaker diarization to identify and separate different voices within a recording.
  • Automated Video Subtitling - Combines AI transcription and translation to generate accurate time-stamped subtitles from video files.
  • Automated Subtitle Generators - Uses Whisper AI to automate the end-to-end process of transcribing audio and embedding subtitles locally.
  • Automatic On-Device Captioning - Generates synchronized text overlays from audio streams using local processing for privacy and offline use.
  • Interactive Timing Adjustments - Allows for manual transcript modifications while automatically adjusting corresponding timestamps to maintain precise audio alignment.
  • Audio-Visual Translation - Translates spoken audio within video content into over 100 different target languages.
  • Burned-in Subtitle Exports - Embeds finalized transcription edits as professional subtitles directly into a video file for distribution.
  • Text Translation Services - Converts generated transcription text from one natural language to another using a wide library of supported languages.
  • Transcript-Based Editing - Enables alignment of captions with video cuts by exporting timing data into professional editing software.
  • Transcription Export Bridges - Exports AI-generated timing and transcription data into professional video editing software.
  • Subtitle Styling - Enables customization of caption colors and effects to meet professional video production requirements.
  • Video Compositing - Bakes stylized subtitle overlays directly into the video frames during the rendering process.
  • Transcription Data Exports - Provides integration to export transcription and timing data into professional video editing software for precise caption alignment.
  • Subtitle Visual Customization - Provides tools for customizing the visual appearance of subtitles and attributing text to different speakers.
  • Audio and Subtitle Tools - Tool to automatically transcribe editing timelines using OpenAI Whisper.

Star history

Star history chart for tmoroney/auto-subsStar history chart for tmoroney/auto-subs

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does tmoroney/auto-subs do?

Auto-subs is an AI transcription and automatic captioning tool that converts spoken audio from video files into synchronized subtitles. It functions as a subtitle generator and a transcription bridge, enabling the conversion of speech to text with automatic speaker identification and multi-language translation support.

What are the main features of tmoroney/auto-subs?

The main features of tmoroney/auto-subs are: On-Device Inference, Audio and Video File Transcription, Subtitle Translation, Audio Transcriptions, Local Speech-to-Text, Neural Machine Translation, Speaker Diarization, Automated Video Subtitling.

What are some open-source alternatives to tmoroney/auto-subs?

Open-source alternatives to tmoroney/auto-subs include: buxuku/smartsub — SmartSub is a cross-platform desktop application for AI-driven video transcription and subtitle generation. It… umlx5h/llplayer — LLPlayer is a language learning media player and AI subtitle generator that integrates large language models for… agermanidis/autosub — Autosub is a command-line media processor and automatic subtitle generator that converts audio streams from video and… huanshere/videolingo — VideoLingo is an automated video localization suite designed to transcribe, translate, and dub video content. It… wyattblue/auto-editor — Auto-editor is a command-line automated video editor that uses FFmpeg to remove silence and inactive footage from… k2-fsa/sherpa-onnx — Sherpa-ONNX is an ONNX-based speech processing toolkit that provides a local speech recognition engine, an on-device…

Open-source alternatives to Auto Subs

Similar open-source projects, ranked by how many features they share with Auto Subs.
  • buxuku/smartsubbuxuku avatar

    buxuku/SmartSub

    4,056View on GitHub↗

    SmartSub is a cross-platform desktop application for AI-driven video transcription and subtitle generation. It converts audio and video files into text subtitles using local AI models and incorporates hardware acceleration to increase processing speed. The tool features a subtitle translator that leverages large language models, such as OpenAI and DeepSeek, to convert subtitles between different languages. It includes a visual editor for proofreading and polishing transcribed text, paired with a video preview for frame-accurate synchronization. The software supports batch processing of multi

    TypeScriptdeepseekelectronnodejs
    View on GitHub↗4,056
  • umlx5h/llplayerumlx5h avatar

    umlx5h/LLPlayer

    3,110View on GitHub↗

    LLPlayer is a language learning media player and AI subtitle generator that integrates large language models for real-time audio transcription and translation. It functions as an LLM-integrated video player and SRT transcription tool, utilizing local or remote AI models to generate text subtitles from audio and video streams. The project distinguishes itself through a contextual translation workflow that sends preceding subtitle lines to language models to maintain conversational flow and sentence structure. It also includes an optical character recognition system to convert bitmap-based subt

    C#asrcsharpflyleaf
    View on GitHub↗3,110
  • agermanidis/autosubagermanidis avatar

    agermanidis/autosub

    4,197View on GitHub↗

    Autosub is a command-line media processor and automatic subtitle generator that converts audio streams from video and audio files into timed text overlays. It functions as an AI speech-to-text converter that uses OpenAI Whisper to generate synchronized subtitles. The tool includes a language translation pipeline to convert transcribed speech into target languages, enabling multilingual video captioning. It manages the process from audio-stream extraction to the serialization of final subtitle files for local storage. The system covers audio-to-text transcription, time-stamped text mapping, a

    Python
    View on GitHub↗4,197
  • huanshere/videolingoHuanshere avatar

    Huanshere/VideoLingo

    17,498View on GitHub↗

    VideoLingo is an automated video localization suite designed to transcribe, translate, and dub video content. It functions as a translation pipeline that utilizes large language models to convert spoken audio into precise text segments and translate them into multiple languages. The system differentiates itself through a multi-step translation refinement process and a specialized natural language processing utility that segments text into single-line captions meeting broadcast standards. It also integrates synthetic voiceover generation to replace or augment original audio tracks. The projec

    Pythonai-translationdubbinglocalization
    View on GitHub↗17,498
  • See all 30 alternatives to Auto Subs→