awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
tangshimin avatar

tangshimin/MuJing

0
View on GitHub↗
4,238 stars·301 forks·Kotlin·GPL-3.0·29 viewsmujingx.com↗

MuJing

MuJing is a contextual English vocabulary learner and interactive media player designed for language study. It extracts words from videos and documents to provide real-world examples and media clips for memorization, functioning as a subtitle-based language tool and a lemma-based word list generator.

The system differentiates itself by linking vocabulary lists to specific video timestamps and subtitles for auditory and visual reinforcement. It includes a video player with bilingual subtitles and keyboard-based transcription and spelling exercises to build muscle memory through movie and television contexts.

The project covers vocabulary extraction from documents, subtitles, and video tracks, paired with word list refinement through lemmatization, frequency filtering, and dictionary-based exclusion. It also manages multimedia learning sources and streams specific video segments associated with target words to reinforce memory.

Features

  • Contextual Language Acquisition - Facilitates English vocabulary acquisition through immersive examples from movies, television shows, and documents.
  • Timestamped Vocabulary Linking - Extracts words from subtitles and links them to timestamps for auditory and visual reinforcement.
  • Video-Tracked Vocabulary - Pulls subtitles from English video tracks to create vocabulary lists linked to specific media timestamps.
  • Intensive Listening Transcription - Uses video playback and transcription exercises to identify vocabulary gaps and improve listening skills.
  • Document-Based Generators - Provides the capability to extract study words from uploaded documents using lemmatization and frequency filters.
  • Subtitle-Based Language Tools - Links vocabulary lists to specific video timestamps and subtitles for reinforcement.
  • Transcript-to-Timestamp Mapping - Maps vocabulary terms to precise video playback offsets for immediate retrieval of audiovisual examples.
  • High-Frequency Word Targeting - Implements pedagogical prioritization of high-frequency word lists to maximize learning efficiency.
  • General Vocabulary Acquisition - Builds personalized word lists from media using frequency filtering and lemmatization.
  • Contextual Vocabulary Learners - Extracts words from videos and documents to provide real-world examples and media clips for memorization.
  • Subtitle-to-Video Mappings - Synchronizes external subtitle files with video streams to enable bilingual display and playback control.
  • Subtitle-Based Extraction - Extracts words from subtitle files and links them to video timestamps for contextual language learning.
  • Vocabulary Learning Tools - Extracts words from media and documents to help users memorize and master specific word lists.
  • Known-Word Exclusion Lists - Filters out known words by comparing extracted lists against core dictionaries and user mastery databases.
  • Language Learning Video Players - Provides a specialized video player with bilingual subtitles and transcription exercises for language study.
  • Lemma-Based Word List Generators - Normalizes derived word forms into root lemmas and filters frequency to create targeted study lists.
  • Lemma Resolution Systems - Normalizes inflected word forms into their base lemmas to consolidate vocabulary lists.
  • Frequency-Based Vocabularies - Uses global corpus frequency data to remove overly common or rare words from study lists.
  • Bilingual Subtitle Renderers - Provides a video player with bilingual switching and automatic pausing for vocabulary extraction.
  • Custom Study Material Management - Converts documents and video files into structured, personalized vocabulary lists with authentic context.
  • Known-Word Exclusions - Excludes words found in core dictionaries or personal known lists to focus on unfamiliar terms.
  • Multimedia Learning Source Management - Allows importing vocabulary and context from video files, subtitle files, and text documents.
  • Transcription and Spelling Exercises - Uses keyboard-based exercises to type words and transcribe subtitles to build muscle memory.
  • Contextual Clip Retrieval - Provides a system to retrieve and play short video clips associated with target words to reinforce memory.
  • Video Playback Components - Streams specific video segments and subtitles associated with a word to reinforce memory.
  • Lemmatization Utilities - Replaces derived word versions with base lemmas to simplify and consolidate vocabulary lists.

Star history

Star history chart for tangshimin/mujingStar history chart for tangshimin/mujing

How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Projects sharing features with MuJing

These projects share indexed features with MuJing. Shared tags can include platform or build tooling; verify the primary use case before treating a result as a replacement.
  • skywind3000/ecdictskywind3000 avatar

    skywind3000/ECDICT

    7,840View on GitHub↗

    ECDICT is a collection of structured linguistic datasets and an English-Chinese dictionary database. It provides bilingual word definitions, phonetic symbols, and parts of speech, alongside a bilingual geographic gazetteer that maps English place names to Chinese equivalents. These resources are available as a multi-format lexicon export in CSV, SQL, StarDict, and MDX formats. The project distinguishes itself by integrating a linguistic corpus dataset that includes word frequency rankings and academic syllabus markers derived from national corpora. It functions as an educational vocabulary re

    Python
    View on GitHub↗7,840
  • umlx5h/llplayerumlx5h avatar

    umlx5h/LLPlayer

    3,110View on GitHub↗

    LLPlayer is a language learning media player and AI subtitle generator that integrates large language models for real-time audio transcription and translation. It functions as an LLM-integrated video player and SRT transcription tool, utilizing local or remote AI models to generate text subtitles from audio and video streams. The project distinguishes itself through a contextual translation workflow that sends preceding subtitle lines to language models to maintain conversational flow and sentence structure. It also includes an optical character recognition system to convert bitmap-based subt

    C#asrcsharpflyleaf
    View on GitHub↗3,110
  • yujiangshui/a-programmers-guide-to-englishyujiangshui avatar

    yujiangshui/A-Programmers-Guide-to-English

    16,428View on GitHub↗

    This project is a systematic framework for English language acquisition that applies structured workflows and cognitive strategies to build linguistic proficiency. It focuses on the construction of a linguistic knowledge base, enabling learners to master vocabulary and grammar through methodical training. The methodology is distinguished by its use of computer science concepts, such as mental-model-based learning and memory buffers, to organize progression. It emphasizes a cognitive-translation bypass to develop target language thinking, reducing mental latency by processing information direc

    englishenglish-learning
    View on GitHub↗16,428
  • kaiyiwing/qwerty-learnerKaiyiwing avatar

    Kaiyiwing/qwerty-learner

    22,429View on GitHub↗

    qwerty-learner is an English typing tutor and vocabulary learning tool designed to combine word memorization with muscle memory training. It functions as a study system for mastering academic and professional word lists through active typing and repetition to improve spelling speed and accuracy. The software utilizes a chapter-driven workflow that progresses from vocabulary introduction to active typing and final dictation. It enables the study of specialized terminology for professional fields or academic exams using curated word lists and targeted typing exercises. The system provides real

    TypeScript
    View on GitHub↗22,429
Compare all 30 related projects→

Frequently asked questions

What does tangshimin/mujing do?

MuJing is a contextual English vocabulary learner and interactive media player designed for language study. It extracts words from videos and documents to provide real-world examples and media clips for memorization, functioning as a subtitle-based language tool and a lemma-based word list generator.

What are the main features of tangshimin/mujing?

The main features of tangshimin/mujing are: Contextual Language Acquisition, Timestamped Vocabulary Linking, Video-Tracked Vocabulary, Intensive Listening Transcription, Document-Based Generators, Subtitle-Based Language Tools, Transcript-to-Timestamp Mapping, High-Frequency Word Targeting.

Which projects share features with tangshimin/mujing?

Projects with overlapping indexed features include: skywind3000/ecdict — ECDICT is a collection of structured linguistic datasets and an English-Chinese dictionary database. It provides… umlx5h/llplayer — LLPlayer is a language learning media player and AI subtitle generator that integrates large language models for… yujiangshui/a-programmers-guide-to-english — This project is a systematic framework for English language acquisition that applies structured workflows and… kaiyiwing/qwerty-learner — qwerty-learner is an English typing tutor and vocabulary learning tool designed to combine word memorization with… solidspoon/dashplayer — DashPlayer is a language learning video player designed for vocabulary and grammar study. It integrates an AI subtitle… 1c7/crash-course-computer-science-chinese — This project is a structured computer science educational course consisting of video lessons, curated playlists, and…