awesome-repositories.com
博客
MCP
awesome-repositories.com

通过 AI 驱动的搜索,发现最优秀的开源仓库。

探索精选搜索开源替代品自托管软件博客网站地图
项目MCP 服务器关于排名机制媒体报道
法律隐私政策服务条款
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
mvdctop avatar

mvdctop/Movie_Data_Capture

0
View on GitHub↗
7,405 星标·1,381 分支·Python·GPL-3.0·5 次浏览

Movie Data Capture

Movie Data Capture is a media library organizer and movie metadata scraper designed to automatically categorize and name files in a local media collection. It functions as an automated content tagger that identifies movie files and applies descriptive tags by extracting film details from web databases.

The system utilizes an HTTP web scraper to fetch information from external APIs and remote HTML content. It employs a filename pattern parser to extract movie titles and release years from local files using regular expressions, which are then used to automate search queries.

The tool maps scraped metadata to folders on a local file system and persists movie details and organization mappings using a JSON data store. These capabilities support home media server management by ensuring local titles are matched with correct descriptions and technical details.

Features

  • Movie and Show Metadata - Retrieves detailed cinematic information such as titles and plot summaries from external databases.
  • Media Metadata Scrapers - Provides a specialized web scraper to extract movie metadata for organizing local media libraries.
  • Filename-to-Tag Parsing - Extracts metadata from filenames to automatically populate descriptive tags for movie files.
  • Movie Information - Extracts detailed film information from web databases to build a structured digital movie catalog.
  • Automated Taggers - Identifies movie files and applies descriptive tags by matching filenames against external data sources.
  • Local File Organizers - Maps scraped digital metadata to physical directory structures to organize the local media library.
  • Content Tagging Systems - Assigns descriptive tags to movie files by matching scraped attributes against naming conventions.
  • Library Organization Automation - Automatically organizes movie files using naming and folder structures based on scraped metadata.
  • Web Scrapers - Fetches and parses live internet content from external APIs to collect movie data.
  • Web Scraping - Implements a web scraper to retrieve movie metadata from external APIs and HTML content.
  • Local Media Library Management - Coordinates the organization and naming of movie files on a local filesystem using scraped metadata.
  • Filename Parsers - Extracts movie titles and release years from filenames using custom pattern parsing.
  • Regular Expression-Based Parsing - Employs regular-expression based parsing to extract structured movie data from raw file paths.
  • Media Filename Parsers - Uses regular expression engines specifically designed to categorize media files based on filename structures.
  • JSON Data Stores - Stores all movie details and organization mappings in structured JSON files for easy editing.
  • JSON-Based Persistence - Persists movie details and organization mappings using structured JSON files on the local filesystem.
  • Flat-File Storage - Uses a flat-file storage architecture to manage metadata without the overhead of a full database.
  • Home Media Utility Setup - Provides utility setup for maintaining a curated collection of movie files on home network hardware.

Star 历史

mvdctop/movie_data_capture 的 Star 历史图表mvdctop/movie_data_capture 的 Star 历史图表

AI 搜索

探索更多 awesome 仓库

用简单的语言描述您的需求 —— AI 将根据相关性为您从数千个精选开源项目中进行排序。

Start searching with AI

常见问题解答

mvdctop/movie_data_capture 是做什么的?

Movie Data Capture is a media library organizer and movie metadata scraper designed to automatically categorize and name files in a local media collection. It functions as an automated content tagger that identifies movie files and applies descriptive tags by extracting film details from web databases.

mvdctop/movie_data_capture 的主要功能有哪些?

mvdctop/movie_data_capture 的主要功能包括:Movie and Show Metadata, Media Metadata Scrapers, Filename-to-Tag Parsing, Movie Information, Automated Taggers, Local File Organizers, Content Tagging Systems, Library Organization Automation。

mvdctop/movie_data_capture 有哪些开源替代品?

mvdctop/movie_data_capture 的开源替代品包括: jxxghp/moviepilot — MoviePilot is a self-hosted media orchestrator and NAS media library automator. It coordinates workflows between… kaina404/flutterdouban — FlutterDouBan is a cross-platform social media client and media community application built with Flutter. It serves as… javscraper/emby.plugins.javscraper — This project is a metadata plugin for Emby that functions as a specialized media scraper. Its primary purpose is to… public-apis/public-apis — This project is a community-curated directory of REST and GraphQL service endpoints designed to assist developers in… rootphantomer/blasting_dictionary — Blasting Dictionary provides curated datasets of common usernames and passwords designed for auditing authentication… freeok/so-novel — so-novel is a web novel downloader and scraping engine designed to extract structured text from websites and convert…

Movie Data Capture 的开源替代方案

相似的开源项目,按与 Movie Data Capture 的功能重合度排序。
  • jxxghp/moviepilotjxxghp 的头像

    jxxghp/MoviePilot

    11,254在 GitHub 上查看↗

    MoviePilot is a self-hosted media orchestrator and NAS media library automator. It coordinates workflows between downloaders, metadata scrapers, and file systems to automate the discovery, downloading, renaming, and organization of movie and television content. The system functions as an LLM media management agent, allowing users to control subscriptions, searches, and file organization through conversational text commands. It also acts as a Model Context Protocol server, exposing internal media management tools via a standardized interface for external AI clients and agents. The project inc

    Python
    在 GitHub 上查看↗11,254
  • kaina404/flutterdoubankaina404 的头像

    kaina404/FlutterDouBan

    9,103在 GitHub 上查看↗

    FlutterDouBan is a cross-platform social media client and media community application built with Flutter. It serves as a mobile interface for discovering and tracking books, movies, and music while providing access to community feeds and user profiles. The project functions as an integration sample that demonstrates how to fetch and display live platform data from external APIs. It includes a simulation layer to interchange live network calls with local mock data for development and testing. The application covers a broad capability surface, including media catalog interfaces, community foru

    Dartandroiddartflutter
    在 GitHub 上查看↗9,103
  • javscraper/emby.plugins.javscraperJavScraper 的头像

    JavScraper/Emby.Plugins.JavScraper

    3,743在 GitHub 上查看↗

    This project is a metadata plugin for Emby that functions as a specialized media scraper. Its primary purpose is to automatically fetch movie details and images from external databases to populate adult media libraries. The tool identifies video content by extracting unique product identification codes from filenames, which it then uses to match media files to the correct database entries. This process automates the retrieval of detailed metadata and cover art, removing the need for manual data entry. The plugin integrates into the media server's metadata provider pipeline, using asynchronou

    C#adultembyfanart-poster
    在 GitHub 上查看↗3,743
  • public-apis/public-apispublic-apis 的头像

    public-apis/public-apis

    441,986在 GitHub 上查看↗

    This project is a community-curated directory of REST and GraphQL service endpoints designed to assist developers in discovering and integrating third-party data sources. It functions as a centralized registry where external services are organized by domain to facilitate rapid software prototyping and application development. The registry relies on a peer-reviewed contribution model, utilizing distributed version control to manage updates and ensure the accuracy of listed endpoints. To maintain high data quality, the project employs schema-based validation for all incoming submissions and com

    Pythonapiapisdataset
    在 GitHub 上查看↗441,986
查看 Movie Data Capture 的所有 30 个替代方案→