awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descoperă cele mai bune repository-uri open source cu căutare AI.

ExploreazăCăutări recomandateAlternative open-sourceSoftware self-hostedBlogHartă site
ProiectServer MCPDespreCum realizăm clasamentulPresă
LegalConfidențialitateTermeni
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

12 repository-uri

Awesome GitHub RepositoriesAudio Visualization Tools

Resources for generating visual representations from audio input data.

Explore 12 awesome GitHub repositories matching graphics & multimedia · Audio Visualization Tools. Refine with filters or upvote what's useful.

Awesome Audio Visualization Tools GitHub Repositories

Găsește cele mai bune repo-uri cu AI.Vom căuta cele mai potrivite repository-uri folosind AI.
  • rigellute/spotify-tuiAvatar Rigellute

    Rigellute/spotify-tui

    19,019Vezi pe GitHub↗

    This project is a terminal-based music controller that provides a text-based interface for managing audio streaming, library navigation, and playback device selection. It functions as a client for remote music services, allowing users to browse catalogs, control playback states, and manage their streaming accounts directly from the command line. The application distinguishes itself through a highly customizable interface and automation capabilities. Users can modify the visual layout, adjust themes, and define custom keyboard shortcuts to create a personalized control workflow. Beyond interac

    Renders real-time visual animations of audio pitch and track analysis data directly within the terminal.

    Rustclirustspotify
    Vezi pe GitHub↗19,019
  • pipecat-ai/pipecatAvatar pipecat-ai

    pipecat-ai/pipecat

    12,846Vezi pe GitHub↗

    Pipecat is a framework and software development kit for building real-time multimodal AI agents and speech-to-speech systems. It utilizes a frame-based data pipeline to route audio, video, and text through a modular sequence of processors, enabling the orchestration of low-latency conversational AI. The project is distinguished by its ability to coordinate complex multimodal services, including speech-to-text, language models, and text-to-speech, within a single pipeline. It features semantic voice activity detection for natural turn-taking, state-machine conversation flows for dialogue manag

    Renders a dynamic visual bar graph representing real-time audio input levels.

    Pythonaichatbot-frameworkchatbots
    Vezi pe GitHub↗12,846
  • katspaugh/wavesurfer.jsAvatar katspaugh

    katspaugh/wavesurfer.js

    10,114Vezi pe GitHub↗

    wavesurfer.js is a WebAudio playback library and interactive waveform visualizer that renders audio data onto an HTML5 canvas. It enables users to see and navigate sound files through a visual representation of audio peaks, allowing for direct seeking and playback control within a web browser. The project is distinguished by its flexible rendering model, which can use precomputed peak data to display waveforms without downloading or decoding the full audio file. It utilizes a plugin-based extension model to integrate advanced tools such as spectrograms, interactive audio timelines, and real-t

    Renders interactive audio waveforms and spectrograms on a web canvas for audio navigation and analysis.

    TypeScriptaudiojavascriptmusic
    Vezi pe GitHub↗10,114
  • netease-youdao/emotivoiceAvatar netease-youdao

    netease-youdao/EmotiVoice

    8,446Vezi pe GitHub↗

    EmotiVoice is an emotional text-to-speech engine and bilingual speech synthesizer designed to generate synthetic audio in English and Chinese. It utilizes a deep learning architecture to produce high-fidelity speech with controllable emotional states and timbres. The project includes a voice cloning framework for replicating specific speaker identities by training custom acoustic models on personal audio datasets. It employs a jointly-trained acoustic-vocoder pipeline and style-embedding-based synthesis to manage expression and reduce audio artifacts. The system covers a broad range of speec

    Generates mel spectrogram plots to visualize and compare predicted audio quality against target speech signals.

    Pythonaideep-learningemotion
    Vezi pe GitHub↗8,446
  • worldveil/dejavuAvatar worldveil

    worldveil/dejavu

    6,764Vezi pe GitHub↗

    Dejavu is a Python audio fingerprinting library and recognition engine. It functions as a digital audio signature tool used to analyze sound waves and create unique identifiers for the purposes of audio search and retrieval. The project enables automatic music identification by matching live audio feeds or recorded clips against a database of fingerprints. It covers audio content matching and digital audio archiving to identify original source recordings from a stored collection. The system incorporates capabilities for generating audio fingerprints, identifying audio tracks, and recognizing

    Analyzes spectrograms to identify peak energy points, creating unique digital signatures for audio tracks.

    Python
    Vezi pe GitHub↗6,764
  • tyiannak/pyaudioanalysisAvatar tyiannak

    tyiannak/pyAudioAnalysis

    6,242Vezi pe GitHub↗

    pyAudioAnalysis este o bibliotecă și un framework Python pentru procesarea și analiza semnalelor audio. Acesta oferă instrumente pentru extragerea reprezentărilor matematice ale sunetului, cum ar fi spectrogramele, și implementează un sistem pentru antrenarea și evaluarea modelelor de machine learning pentru a clasifica segmentele audio pe baza tiparelor acustice. Proiectul include utilitare dedicate pentru segmentarea audio, care permit eliminarea tăcerii și detectarea unor evenimente audio specifice pentru a împărți înregistrările în secțiuni semnificative. De asemenea, oferă capabilități de vizualizare a datelor care utilizează reducerea dimensionalității pentru a mapa similitudinile de conținut și a identifica clustere în datele sonore. Biblioteca acoperă o gamă largă de capabilități de procesare a semnalelor, inclusiv extragerea caracteristicilor în domeniul spectral, analiza temporală și regresia audio pentru estimarea valorilor continue. Aceste funcții sunt accesibile atât ca bibliotecă programabilă, cât și printr-o interfață de linie de comandă pentru procesarea în lot (batch) a fișierelor audio.

    Provides data visualization capabilities that use dimensionality reduction to map content similarities and identify clusters within sound data.

    Python
    Vezi pe GitHub↗6,242
  • syedhali/ezaudioAvatar syedhali

    syedhali/EZAudio

    4,991Vezi pe GitHub↗

    EZAudio este o bibliotecă audio pentru platformele Apple care oferă interfețe standardizate pentru captarea microfonului, redarea fișierelor și output-ul hardware. Acesta funcționează ca un procesor audio cu latență scăzută și un framework de vizualizare conceput pentru a manipula bufferele audio și a ruta semnalele cu o întârziere minimă. Proiectul dispune de un renderer de forme de undă accelerat hardware pentru desenarea amplitudinilor audio în timp real și a graficelor dinamice. Include, de asemenea, un analizor Fast Fourier Transform care convertește mostrele audio din domeniul timp în date din domeniul frecvență pentru analiză spectrală. Biblioteca acoperă o gamă largă de capabilități, inclusiv înregistrarea audio digitală pe disc și gestionarea redării fișierelor audio cu controlul căutării și al volumului. Suportă procesarea audio în timp real prin înlănțuirea efectelor audio și rutarea input-ului de la microfon direct către output-ul hardware.

    Supplies a framework for real-time audio processing and spectral visualization using Core Audio.

    Objective-C
    Vezi pe GitHub↗4,991
  • makcedward/nlpaugAvatar makcedward

    makcedward/nlpaug

    4,658Vezi pe GitHub↗

    nlpaug este o bibliotecă de augmentare a datelor concepută pentru a genera text sintetic, audio și date de spectrogramă, cu scopul de a îmbunătăți robustețea modelelor de machine learning. Funcționează ca un sintetizator de date textuale și un augmentator de semnal audio, oferind instrumente specializate pentru extinderea seturilor de date prin diverse metode de transformare. Proiectul se distinge prin capacitatea de a orchestra fluxuri de lucru complexe folosind un orchestrator de pipeline, care permite înlănțuirea mai multor funcții de augmentare secvențial sau aleatoriu. Suportă sinteza sofisticată de text prin back-translation, contextual word embeddings și integrarea modelelor de limbaj pre-antrenate, oferind în același timp augmentarea imaginilor de spectrogramă prin mascarea timpului și a frecvenței. Biblioteca acoperă o gamă largă de capabilități, inclusiv modificarea semnalului audio cu injectare de zgomot și pitch shifting, alterări de text bazate pe reguli pentru simularea greșelilor de scriere și extinderea seturilor de date prin generarea de propoziții și substituție semantică. Oferă, de asemenea, controale pentru volumul de augmentare și filtrarea țintei folosind expresii regulate pentru a proteja anumite token-uri de modificare.

    Transforms audio spectrograms using time and frequency masking to improve speech recognition robustness.

    Jupyter Notebook
    Vezi pe GitHub↗4,658
  • serversideup/amplitudejsAvatar serversideup

    serversideup/amplitudejs

    4,313Vezi pe GitHub↗

    AmplitudeJS is a JavaScript library and framework for building custom HTML5 audio players. It serves as a client-side playlist manager and media controller that bridges the gap between HTML elements and the Web Audio API, allowing developers to create branded media interfaces without relying on default browser styles. The project is distinguished by its use of CSS-class-based DOM binding and data-attribute state mapping, which links HTML elements directly to playback controls and track metadata. It includes a dedicated visualization system that uses the Web Audio API to render real-time SVG w

    Renders real-time SVG waveforms and frequency-based visual effects using audio signal data.

    JavaScriptcsshtmlhtml5
    Vezi pe GitHub↗4,313
  • alexkay/spekAvatar alexkay

    alexkay/spek

    3,185Vezi pe GitHub↗

    Spek is an acoustic spectrum tool and audio frequency visualizer designed to decode audio streams and analyze their spectral density. It functions as an audio spectrogram analyzer that displays frequency distributions to help identify the sonic characteristics of audio files. The tool specifically includes capabilities as a lossy compression detector, allowing for the identification of encoding artifacts and frequency cut-offs caused by lossy transcoding. The software covers audio file inspection and spectral analysis, providing the ability to select individual audio streams and channels. Us

    Generates frequency-based heat maps to analyze audio content over time as spectrograms.

    C++
    Vezi pe GitHub↗3,185
  • dominicbreuker/stego-toolkitAvatar DominicBreuker

    DominicBreuker/stego-toolkit

    2,636Vezi pe GitHub↗

    This project is a steganography analysis toolkit and digital forensics suite designed to detect, extract, and embed hidden data within image and audio files. It provides a dockerized security environment that bundles various analysis tools into a containerized workspace, including a media spectrogram visualizer for revealing visually hidden patterns. The toolkit features a dedicated brute force system for recovering password-protected messages using automated wordlists and candidate password testing. It distinguishes itself by providing rule-based wordlist generation that uses expansion patte

    Includes a graphical interface to render audio spectrograms for revealing visually hidden patterns.

    Shellctf-toolsdocker-imagesteganography
    Vezi pe GitHub↗2,636
  • miek/inspectrumAvatar miek

    miek/inspectrum

    2,466Vezi pe GitHub↗

    Divides the time-frequency display into cached image tiles recomputed only on zoom or pan changes.

    C++dspsdr
    Vezi pe GitHub↗2,466
  1. Home
  2. Graphics & Multimedia
  3. Media Processing and Analysis
  4. Media Manipulation
  5. Media Processing
  6. Audio Visualization Tools

Explorează sub-etichetele

  • Dimensionality Reduction VisualizationTools that project high-dimensional audio features into lower dimensions to visualize clusters and similarities. **Distinct from Audio Visualization Tools:** Specifically handles dimensionality reduction for audio feature sets rather than general visual representations of audio input
  • Metadata OverlaysVisual representations of audio properties such as frequency spectrums, time labels, and region markers. **Distinct from Audio Visualization Tools:** Adds technical context (markers, labels) over the waveform, whereas the parent is general visualization.
  • Spectrogram Renderers3 sub-tag-uriTools that generate frequency-based heat maps to analyze audio content over time. **Distinct from Audio Visualization Tools:** Specifically focuses on the generation of spectrograms rather than general audio visualization or feature extraction.