awesome-repositories.com
Blog
MCP
awesome-repositories.com

Découvrez les meilleurs dépôts open-source grâce à notre recherche par IA.

ExplorerRecherches sélectionnéesAlternatives open sourceLogiciels auto-hébergésBlogPlan du site
ProjetServeur MCPÀ proposNotre méthodologiePresse
Mentions légalesConfidentialitéConditions d'utilisation
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
modelscope avatar

modelscope/FunClip

0
View on GitHub↗
5,850 stars·704 forks·Python·MIT·16 vueswww.funasr.com↗

FunClip

FunClip is an open-source tool that transcribes speech from video files and clips segments based on text, speaker, or AI analysis. It combines speech recognition with speaker diarization, audio event detection, and visual content understanding to identify and extract relevant portions of a video.

The tool distinguishes itself through several integrated capabilities. It supports hotword-weighted speech recognition, which improves transcription accuracy for specific terms like names or jargon by boosting their probability during decoding. A large language model can interpret the transcribed text to automatically select video segments based on natural language prompts. Speaker diarization separates and labels audio segments by speaker identity, enabling clipping by a chosen speaker. Additionally, a visual-content understanding model analyzes video frames to select clips when the transcript alone is insufficient.

Beyond these core differentiators, FunClip generates SRT subtitle files for both the full video and each clipped segment. It provides a command-line interface for headless, scriptable execution of the entire recognition and clipping pipeline, as well as a web service interface accessible locally or over a network for browser-based use.

Features

  • Video Clip Extraction - An open-source tool that transcribes video speech and clips segments by text, speaker, or AI analysis.
  • Hotword Boosts - Improves speech recognition accuracy for specific terms by marking them as hotwords.
  • Speaker Diarizers - Separates and labels audio segments by speaker identity using clustering of voice embeddings.
  • Speaker-Based Video Clippers - Identify speakers in a video using speaker recognition and clip segments belonging to a chosen speaker.
  • Automated Video Transcribers - Transcribing speech from video files into text with accurate word-level timestamps for downstream processing.
  • Transcription-Based Video Clippers - Transcribe a video's speech into text and clip segments matching text phrases you specify.
  • Transcription Term Boosts - Improve transcription accuracy for specific terms like names or entities by marking them as hotwords.
  • Command-Line Video Clippers - Run the recognition and clipping workflow directly from a terminal for automated or scripted use.
  • AI-Assisted Clip Selectors - Using large language models to analyze transcripts and automatically select relevant video segments based on user prompts.
  • AI-Assisted Clips - Use a large language model to analyze a transcript and automatically select clip segments based on your prompt.
  • Text-Based Video Clippers - Extract video segments whose transcribed text matches words or phrases you specify.
  • Transcript-Based Video Clippers - Extract video segments that match text phrases you select from the transcription results.
  • Visual-Content Clips - Selecting video segments by analyzing both visual content and audio when transcript alone is insufficient.
  • LLM-Based Transcript Selectors - Uses a large language model to interpret transcribed text and select relevant video segments based on natural language prompts.
  • LLM-Based Video Clippers - Use a large language model to interpret the transcript and automatically pick relevant video segments.
  • Visual-Content Video Clippers - Select video segments by analyzing both visuals and audio using a video understanding model.
  • Hotword-Weighted Recognizers - Improving speech recognition accuracy for specific terms like names or jargon by providing a custom hotword list.
  • Frame-Level Video Analyzers - Analyzes video frames with a vision model to select clips when transcript alone is insufficient.
  • Subtitle Generators - Produce SRT subtitle files for the full video and for the clipped segments during the clipping process.
  • Video Subtitle Generators - Produces SRT subtitle files for both the full video and each clipped segment during the processing workflow.
  • Local Web Interfaces - Exposes the clipping functionality through a browser-based UI that can be accessed locally or over a network.

Historique des stars

Graphique de l'historique des stars pour modelscope/funclipGraphique de l'historique des stars pour modelscope/funclip

Recherche par IA

Explorez plus de dépôts awesome

Décrivez vos besoins en langage naturel — l'IA classe des milliers de projets open source sélectionnés par pertinence.

Start searching with AI

Alternatives open source à FunClip

Projets open source similaires, classés selon le nombre de fonctionnalités partagées avec FunClip.
  • samuraigpt/ai-youtube-shorts-generatorAvatar de SamurAIGPT

    SamurAIGPT/AI-Youtube-Shorts-Generator

    3,037Voir sur GitHub↗

    This project is an AI-driven suite of tools designed to repurpose long-form video content into short-form clips. It integrates a speech-to-text engine for automated transcription, a highlighting system that ranks engaging segments based on emotional hooks, and a video processor that converts horizontal footage into vertical formats. The system distinguishes itself through intelligent video cropping that utilizes face tracking and motion smoothing to keep subjects centered. It also employs an analysis system to extract viral highlights by scoring segments for engagement and practical value. T

    Pythonai-video-generatorartificial-intelligenceimage-to-video
    Voir sur GitHub↗3,037
  • wxbool/video-srt-windowsAvatar de wxbool

    wxbool/video-srt-windows

    5,037Voir sur GitHub↗

    This is a Windows application for automatic speech recognition that transcribes spoken audio from video files into timestamped SRT subtitle files. It serves as a subtitle generator and translation tool that converts media speech into synchronized text. The software functions as a batch media transcriber, allowing the simultaneous processing of multiple audio and video files to generate subtitles in bulk. It includes a translation workflow to convert transcriptions between different languages for the creation of bilingual or localized files. The system also provides text refinement capabiliti

    Goffmpeggogolang
    Voir sur GitHub↗5,037
  • timerring/biliveAvatar de timerring

    timerring/bilive

    3,125Voir sur GitHub↗

    Bilive is a multimodal AI video pipeline and live stream recording tool designed to capture real-time broadcasts and automate the creation of highlight clips. It functions as a multi-platform stream orchestrator capable of distributing looped pre-recorded content and managing the automated upload of processed video clips to various destinations. The system distinguishes itself through AI-driven content generation, using comment density to detect high-energy segments and multimodal models to automatically produce descriptive titles and synchronized subtitles. It further utilizes image-to-image

    Pythonassbilibilibili
    Voir sur GitHub↗3,125
  • breakthrough/pyscenedetectAvatar de Breakthrough

    Breakthrough/PySceneDetect

    4,556Voir sur GitHub↗

    PySceneDetect is a suite of tools for identifying cuts and transitions in video files using content, threshold, and histogram detection algorithms. It functions as a scene detector, frame extractor, statistics analyzer, metadata exporter, and video scene splitter. The project identifies scene boundaries and can divide video files into smaller clips using external processing tools. It allows for the extraction of representative image frames from detected changes and the export of scene lists into industry-standard formats such as EDL, FCP, HTML, OTIO, and CSV. The toolset includes capabilitie

    Pythonanalysisimage-processingopencv
    Voir sur GitHub↗4,556
Voir les 30 alternatives à FunClip→

Questions fréquentes

Que fait modelscope/funclip ?

FunClip is an open-source tool that transcribes speech from video files and clips segments based on text, speaker, or AI analysis. It combines speech recognition with speaker diarization, audio event detection, and visual content understanding to identify and extract relevant portions of a video.

Quelles sont les fonctionnalités principales de modelscope/funclip ?

Les fonctionnalités principales de modelscope/funclip sont : Video Clip Extraction, Hotword Boosts, Speaker Diarizers, Speaker-Based Video Clippers, Automated Video Transcribers, Transcription-Based Video Clippers, Transcription Term Boosts, Command-Line Video Clippers.

Quelles sont les alternatives open-source à modelscope/funclip ?

Les alternatives open-source à modelscope/funclip incluent : samuraigpt/ai-youtube-shorts-generator — This project is an AI-driven suite of tools designed to repurpose long-form video content into short-form clips. It… wxbool/video-srt-windows — This is a Windows application for automatic speech recognition that transcribes spoken audio from video files into… vvo/gifify — Gifify is a tool for converting video files into optimized animated GIFs. It functions as a video to GIF converter and… breakthrough/pyscenedetect — PySceneDetect is a suite of tools for identifying cuts and transitions in video files using content, threshold, and… timerring/bilive — Bilive is a multimodal AI video pipeline and live stream recording tool designed to capture real-time broadcasts and… buxuku/smartsub — SmartSub is a cross-platform desktop application for AI-driven video transcription and subtitle generation. It…