awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
buxuku avatar

buxuku/SmartSub

0
View on GitHub↗
4,056 stars·274 forks·TypeScript·MIT·54 viewssmartsub.linxiaodong.com↗

SmartSub

SmartSub is a cross-platform desktop application for AI-driven video transcription and subtitle generation. It converts audio and video files into text subtitles using local AI models and incorporates hardware acceleration to increase processing speed.

The tool features a subtitle translator that leverages large language models, such as OpenAI and DeepSeek, to convert subtitles between different languages. It includes a visual editor for proofreading and polishing transcribed text, paired with a video preview for frame-accurate synchronization.

The software supports batch processing of multiple media files and provides utilities for embedding subtitles as switchable soft tracks or permanently burning them into video frames. It also includes systems for managing transcription model files and configuring external AI service parameters.

Features

  • Automated Subtitle Generators - Combines speech recognition and subtitle generation into an automated workflow for video files.
  • Audio and Video File Transcription - A desktop application that extracts speech from media files to produce subtitles and timestamps offline.
  • Subtitle Translation - Converts generated subtitles into different languages using cloud APIs or local LLMs.
  • Local Model Execution - Runs speech recognition engines directly on the local device to ensure privacy and offline processing.
  • Hardware-Accelerated Transcribers - A processing tool that leverages graphics processors to speed up the transcription of media files into text.
  • Accelerated Transcriptions - Utilizes GPU compute engines to significantly speed up the conversion of audio signals into text.
  • Automated Video Transcribers - Converts spoken audio from video files into time-synced text subtitles using local AI.
  • Video and Subtitle Handling - Provides a specialized visual environment for processing video streams and refining subtitle text.
  • Subtitle Processing - Offers a visual editor for reviewing and correcting transcribed text with integrated video preview.
  • Burned-In Subtitle Rendering - Implements tools for permanently overlaying text captions directly onto video frames or creating soft tracks.
  • GPU Hardware Acceleration - Leverages graphics processing units to significantly increase the speed of audio transcription and media processing.
  • Subtitle Editors - Provides a visual editor with video preview for proofreading and polishing transcribed text and timestamps.
  • Subtitle Embedding - Burns subtitles permanently into video frames or embeds them as switchable soft tracks.
  • Burned-in Subtitle Exports - Supports rendering subtitles permanently into video frames or embedding them as switchable soft tracks.
  • Batch File Processing - Provides a workflow for processing multiple media files concurrently through transcription and translation pipelines.
  • Batch Video Processing - Automates the transcription and translation process across multiple video files simultaneously.
  • Timeline Synchronization - Pairs a video preview player with a text editor for frame-accurate modification of subtitle timestamps.

Star history

Star history chart for buxuku/smartsubStar history chart for buxuku/smartsub

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does buxuku/smartsub do?

SmartSub is a cross-platform desktop application for AI-driven video transcription and subtitle generation. It converts audio and video files into text subtitles using local AI models and incorporates hardware acceleration to increase processing speed.

What are the main features of buxuku/smartsub?

The main features of buxuku/smartsub are: Automated Subtitle Generators, Audio and Video File Transcription, Subtitle Translation, Local Model Execution, Hardware-Accelerated Transcribers, Accelerated Transcriptions, Automated Video Transcribers, Video and Subtitle Handling.

Which projects share features with buxuku/smartsub?

Projects with overlapping indexed features include: tmoroney/auto-subs — Auto-subs is an AI transcription and automatic captioning tool that converts spoken audio from video files into… wxbool/video-srt-windows — This is a Windows application for automatic speech recognition that transcribes spoken audio from video files into… browser-use/video-use — This project is an AI video post-production suite that uses large language models and programmatic tools to automate… weifeng2333/videocaptioner — VideoCaptioner is an automated tool designed to generate and embed time-synchronized subtitles into video files. By… umlx5h/llplayer — LLPlayer is a language learning media player and AI subtitle generator that integrates large language models for… yils-lin/short-video-factory — Short video factory is a local AI content generator and automated video editing tool. It provides a production…

Projects sharing features with SmartSub

These projects share indexed features with SmartSub. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • tmoroney/auto-substmoroney avatar

    tmoroney/auto-subs

    2,851View on GitHub↗

    Auto-subs is an AI transcription and automatic captioning tool that converts spoken audio from video files into synchronized subtitles. It functions as a subtitle generator and a transcription bridge, enabling the conversion of speech to text with automatic speaker identification and multi-language translation support. The software prioritizes data privacy by utilizing on-device AI inference to process audio and video files locally on the user's hardware. It distinguishes itself by offering deep integration with professional video editing workflows, allowing users to export timing and transcr

    TypeScriptaidavincidavinci-resolve
    View on GitHub↗2,851
  • wxbool/video-srt-windowswxbool avatar

    wxbool/video-srt-windows

    5,037View on GitHub↗

    This is a Windows application for automatic speech recognition that transcribes spoken audio from video files into timestamped SRT subtitle files. It serves as a subtitle generator and translation tool that converts media speech into synchronized text. The software functions as a batch media transcriber, allowing the simultaneous processing of multiple audio and video files to generate subtitles in bulk. It includes a translation workflow to convert transcriptions between different languages for the creation of bilingual or localized files. The system also provides text refinement capabiliti

    Goffmpeggogolang
    View on GitHub↗5,037
  • browser-use/video-usebrowser-use avatar

    browser-use/video-use

    9,743View on GitHub↗

    This project is an AI video post-production suite that uses large language models and programmatic tools to automate editing, transcription, and subtitle generation. It functions as an AI editing agent that translates natural language instructions into shell commands, providing a programmatic interface for manipulating media via FFmpeg. The toolkit includes a motion graphics engine that generates technical animations and visual overlays through code-driven rendering and mathematical definitions. It distinguishes itself by combining an AI-powered transcriber for word-level timestamps with an a

    Python
    View on GitHub↗9,743
  • weifeng2333/videocaptionerWEIFENG2333 avatar

    WEIFENG2333/VideoCaptioner

    13,278View on GitHub↗

    VideoCaptioner is an automated tool designed to generate and embed time-synchronized subtitles into video files. By leveraging speech recognition models, the software converts spoken audio into text and calculates precise timestamps to ensure captions align with the original media. The project operates as a local-first inference pipeline, performing all transcription tasks on the host machine to maintain data privacy. It utilizes a transformer-based neural network for speech recognition and integrates a multimedia framework to handle the technical aspects of video processing and subtitle stre

    Pythonaisubtitletranslate
    View on GitHub↗13,278
Compare all 30 related projects→