awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
mvdctop avatar

mvdctop/Movie_Data_Capture

0
View on GitHub↗
7,405 stars·1,381 forks·Python·GPL-3.0·20 views

Movie Data Capture

Movie Data Capture is a media library organizer and movie metadata scraper designed to automatically categorize and name files in a local media collection. It functions as an automated content tagger that identifies movie files and applies descriptive tags by extracting film details from web databases.

The system utilizes an HTTP web scraper to fetch information from external APIs and remote HTML content. It employs a filename pattern parser to extract movie titles and release years from local files using regular expressions, which are then used to automate search queries.

The tool maps scraped metadata to folders on a local file system and persists movie details and organization mappings using a JSON data store. These capabilities support home media server management by ensuring local titles are matched with correct descriptions and technical details.

Features

  • Movie and Show Metadata - Retrieves detailed cinematic information such as titles and plot summaries from external databases.
  • Media Metadata Scrapers - Provides a specialized web scraper to extract movie metadata for organizing local media libraries.
  • Filename-to-Tag Parsing - Extracts metadata from filenames to automatically populate descriptive tags for movie files.
  • Movie Information - Extracts detailed film information from web databases to build a structured digital movie catalog.
  • Automated Taggers - Identifies movie files and applies descriptive tags by matching filenames against external data sources.
  • Local File Organizers - Maps scraped digital metadata to physical directory structures to organize the local media library.
  • Content Tagging Systems - Assigns descriptive tags to movie files by matching scraped attributes against naming conventions.
  • Library Organization Automation - Automatically organizes movie files using naming and folder structures based on scraped metadata.
  • Web Scrapers - Fetches and parses live internet content from external APIs to collect movie data.
  • Web Scraping - Implements a web scraper to retrieve movie metadata from external APIs and HTML content.
  • Local Media Library Management - Coordinates the organization and naming of movie files on a local filesystem using scraped metadata.
  • Filename Parsers - Extracts movie titles and release years from filenames using custom pattern parsing.
  • Regular Expression-Based Parsing - Employs regular-expression based parsing to extract structured movie data from raw file paths.
  • Media Filename Parsers - Uses regular expression engines specifically designed to categorize media files based on filename structures.
  • JSON Data Stores - Stores all movie details and organization mappings in structured JSON files for easy editing.
  • JSON-Based Persistence - Persists movie details and organization mappings using structured JSON files on the local filesystem.
  • Flat-File Storage - Uses a flat-file storage architecture to manage metadata without the overhead of a full database.
  • Home Media Utility Setup - Provides utility setup for maintaining a curated collection of movie files on home network hardware.

Star history

Star history chart for mvdctop/movie_data_captureStar history chart for mvdctop/movie_data_capture

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Frequently asked questions

What does mvdctop/movie_data_capture do?

Movie Data Capture is a media library organizer and movie metadata scraper designed to automatically categorize and name files in a local media collection. It functions as an automated content tagger that identifies movie files and applies descriptive tags by extracting film details from web databases.

What are the main features of mvdctop/movie_data_capture?

The main features of mvdctop/movie_data_capture are: Movie and Show Metadata, Media Metadata Scrapers, Filename-to-Tag Parsing, Movie Information, Automated Taggers, Local File Organizers, Content Tagging Systems, Library Organization Automation.

Which projects share features with mvdctop/movie_data_capture?

Projects with overlapping indexed features include: jxxghp/moviepilot — MoviePilot is a self-hosted media orchestrator and NAS media library automator. It coordinates workflows between… kaina404/flutterdouban — FlutterDouBan is a cross-platform social media client and media community application built with Flutter. It serves as… javscraper/emby.plugins.javscraper — This project is a metadata plugin for Emby that functions as a specialized media scraper. Its primary purpose is to… public-apis/public-apis — This project is a community-curated directory of REST and GraphQL service endpoints designed to assist developers in… rootphantomer/blasting_dictionary — Blasting Dictionary provides curated datasets of common usernames and passwords designed for auditing authentication… freeok/so-novel — so-novel is a web novel downloader and scraping engine designed to extract structured text from websites and convert…

Projects sharing features with Movie Data Capture

These projects share indexed features with Movie Data Capture. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • jxxghp/moviepilotjxxghp avatar

    jxxghp/MoviePilot

    11,254View on GitHub↗

    MoviePilot is a self-hosted media orchestrator and NAS media library automator. It coordinates workflows between downloaders, metadata scrapers, and file systems to automate the discovery, downloading, renaming, and organization of movie and television content. The system functions as an LLM media management agent, allowing users to control subscriptions, searches, and file organization through conversational text commands. It also acts as a Model Context Protocol server, exposing internal media management tools via a standardized interface for external AI clients and agents. The project inc

    Python
    View on GitHub↗11,254
  • kaina404/flutterdoubankaina404 avatar

    kaina404/FlutterDouBan

    9,103View on GitHub↗

    FlutterDouBan is a cross-platform social media client and media community application built with Flutter. It serves as a mobile interface for discovering and tracking books, movies, and music while providing access to community feeds and user profiles. The project functions as an integration sample that demonstrates how to fetch and display live platform data from external APIs. It includes a simulation layer to interchange live network calls with local mock data for development and testing. The application covers a broad capability surface, including media catalog interfaces, community foru

    Dartandroiddartflutter
    View on GitHub↗9,103
  • javscraper/emby.plugins.javscraperJavScraper avatar

    JavScraper/Emby.Plugins.JavScraper

    3,743View on GitHub↗

    This project is a metadata plugin for Emby that functions as a specialized media scraper. Its primary purpose is to automatically fetch movie details and images from external databases to populate adult media libraries. The tool identifies video content by extracting unique product identification codes from filenames, which it then uses to match media files to the correct database entries. This process automates the retrieval of detailed metadata and cover art, removing the need for manual data entry. The plugin integrates into the media server's metadata provider pipeline, using asynchronou

    C#adultembyfanart-poster
    View on GitHub↗3,743
  • public-apis/public-apispublic-apis avatar

    public-apis/public-apis

    441,986View on GitHub↗

    This project is a community-curated directory of REST and GraphQL service endpoints designed to assist developers in discovering and integrating third-party data sources. It functions as a centralized registry where external services are organized by domain to facilitate rapid software prototyping and application development. The registry relies on a peer-reviewed contribution model, utilizing distributed version control to manage updates and ensure the accuracy of listed endpoints. To maintain high data quality, the project employs schema-based validation for all incoming submissions and com

    Pythonapiapisdataset
    View on GitHub↗441,986
  • Compare all 30 related projects→