awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目关于排名机制媒体报道MCP 服务器
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
agermanidis avatar

agermanidis/autosub

0
View on GitHub↗
4,197 星标·1,635 分支·Python·MIT·4 次浏览

Autosub

Autosub 是一个命令行媒体处理器和自动字幕生成器,可将视频和音频文件中的音频流转换为带时间戳的文本覆盖层。它作为一个 AI 语音转文本转换器,使用 OpenAI Whisper 来生成同步字幕。

该工具包含一个语言翻译流水线,可将转录的语音转换为目标语言,从而实现多语言视频字幕制作。它管理从音频流提取到最终字幕文件序列化以进行本地存储的整个过程。

该系统涵盖了音频转文本转录、带时间戳的文本映射以及基于终端的媒体处理。

Features

  • Automated Subtitle Generators - Provides an automated workflow that combines speech recognition and subtitle embedding for video files.
  • Language Translation Services - Passes transcribed text through a translation service to support multilingual subtitles.
  • Speech Recognition APIs - Integrates with speech recognition APIs to convert audio data into timed text segments.
  • Speech-to-Text Converters - Converts audio streams into text and supports translating transcriptions into different languages.
  • Whisper-Based Engines - Utilizes the OpenAI Whisper architecture to generate synchronized subtitles from audio and video files.
  • Speech to Text Transcription - Converts audio from media files into processed text with precise timestamps.
  • Automated Video Subtitling - Implements an AI-driven pipeline that combines transcription and translation to generate timed captions.
  • Timed Text Tracks - Synchronizes transcribed text with specific start and end timestamps for video playback.
  • Subtitle Translation - Translates transcribed speech into target languages specifically for video subtitles and captions.
  • Command-Line Media Processors - Provides a terminal-based application for automating the creation of subtitle files from various video formats.
  • Media Processing - Provides a terminal interface for manipulating and processing audio and video media files.
  • Multilingual Captioning - Generates and times subtitles in multiple languages for international audiences.
  • Machine Learning Projects - Automated speech recognition and subtitle generation for video files.
  • Media Automation - Command-line utility for auto-generating video subtitles.

Star 历史

agermanidis/autosub 的 Star 历史图表agermanidis/autosub 的 Star 历史图表

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Start searching with AI

Autosub 的开源替代方案

相似的开源项目,按与 Autosub 的功能重合度排序。
  • linyqh/narratoailinyqh 的头像

    linyqh/NarratoAI

    8,091在 GitHub 上查看↗

    NarratoAI is an automated video production pipeline that uses large language models to generate scripts, voiceovers, and edited video commentary. It functions as a combined scriptwriter, voiceover generator, and video editor to streamline the creation of movie and television commentary content. The system automates the production workflow by converting input data into structured narrative scripts, synthesizing artificial speech for narration, and programmatically assembling video clips based on script timestamps. It also converts spoken audio from video files into written text for subtitles a

    Pythonaiagentaiopsgemini-api
    在 GitHub 上查看↗8,091
  • umlx5h/llplayerumlx5h 的头像

    umlx5h/LLPlayer

    3,110在 GitHub 上查看↗

    LLPlayer is a language learning media player and AI subtitle generator that integrates large language models for real-time audio transcription and translation. It functions as an LLM-integrated video player and SRT transcription tool, utilizing local or remote AI models to generate text subtitles from audio and video streams. The project distinguishes itself through a contextual translation workflow that sends preceding subtitle lines to language models to maintain conversational flow and sentence structure. It also includes an optical character recognition system to convert bitmap-based subt

    C#asrcsharpflyleaf
    在 GitHub 上查看↗3,110
  • tmoroney/auto-substmoroney 的头像

    tmoroney/auto-subs

    2,851在 GitHub 上查看↗

    Auto-subs is an AI transcription and automatic captioning tool that converts spoken audio from video files into synchronized subtitles. It functions as a subtitle generator and a transcription bridge, enabling the conversion of speech to text with automatic speaker identification and multi-language translation support. The software prioritizes data privacy by utilizing on-device AI inference to process audio and video files locally on the user's hardware. It distinguishes itself by offering deep integration with professional video editing workflows, allowing users to export timing and transcr

    TypeScriptaidavincidavinci-resolve
    在 GitHub 上查看↗2,851
  • wxbool/video-srt-windowswxbool 的头像

    wxbool/video-srt-windows

    5,037在 GitHub 上查看↗

    This is a Windows application for automatic speech recognition that transcribes spoken audio from video files into timestamped SRT subtitle files. It serves as a subtitle generator and translation tool that converts media speech into synchronized text. The software functions as a batch media transcriber, allowing the simultaneous processing of multiple audio and video files to generate subtitles in bulk. It includes a translation workflow to convert transcriptions between different languages for the creation of bilingual or localized files. The system also provides text refinement capabiliti

    Goffmpeggogolang
    在 GitHub 上查看↗5,037
查看 Autosub 的所有 30 个替代方案→

常见问题解答

agermanidis/autosub 是做什么的?

Autosub 是一个命令行媒体处理器和自动字幕生成器,可将视频和音频文件中的音频流转换为带时间戳的文本覆盖层。它作为一个 AI 语音转文本转换器,使用 OpenAI Whisper 来生成同步字幕。

agermanidis/autosub 的主要功能有哪些?

agermanidis/autosub 的主要功能包括:Automated Subtitle Generators, Language Translation Services, Speech Recognition APIs, Speech-to-Text Converters, Whisper-Based Engines, Speech to Text Transcription, Automated Video Subtitling, Timed Text Tracks。

agermanidis/autosub 有哪些开源替代品?

agermanidis/autosub 的开源替代品包括: linyqh/narratoai — NarratoAI is an automated video production pipeline that uses large language models to generate scripts, voiceovers,… tmoroney/auto-subs — Auto-subs is an AI transcription and automatic captioning tool that converts spoken audio from video files into… umlx5h/llplayer — LLPlayer is a language learning media player and AI subtitle generator that integrates large language models for… wxbool/video-srt-windows — This is a Windows application for automatic speech recognition that transcribes spoken audio from video files into… abus-aikorea/voice-pro — Voice Pro is a comprehensive speech and audio processing toolkit that combines text-to-speech synthesis, voice… weifeng2333/videocaptioner — VideoCaptioner is an automated tool designed to generate and embed time-synchronized subtitles into video files. By…