Showing 140 open source projects for "waveform"

View related business solutions
  • 99.99% Uptime for MySQL and PostgreSQL Databases Icon
    99.99% Uptime for MySQL and PostgreSQL Databases

    Sub-second maintenance. 2x read/write performance. Built-in vector search for AI apps.

    Cloud SQL Enterprise Plus delivers near-zero downtime with 35 days of point-in-time recovery. Supports MySQL, PostgreSQL, and SQL Server.
    Try Free
  • Train ML Models With SQL You Already Know Icon
    Train ML Models With SQL You Already Know

    BigQuery automates data prep, analysis, and predictions with built-in AI assistance.

    Build and deploy ML models using familiar SQL. Automate data prep with built-in Gemini. Query 1 TB and store 10 GB free monthly.
    Try Free
  • 1
    Color to Waveform

    Color to Waveform

    Convert colors to synth presets

    The purpose of the program is to convert a color to a waveform you can use as a synthesizer oscillator inside a DAW such as FL Studio from Image Line. Many synths are provided with an option to load your own waveform, to replace the basic saw, square and sine waveforms commonly used to create synth sounds. The waveform generated by the program will correspond to the subliminal synesthetic sensation of the selected color.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 2
    FDWaveformView

    FDWaveformView

    Reads an audio file and displays the waveform

    FDWaveformView is an easy way to display an audio waveform in your app. It is a nice visualization to show a playing audio file or to select a position in a file. To use it, add an FDWaveformView using Interface Builder or programmatically and then just load your audio as per this example.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 3
    wavesurfer.js

    wavesurfer.js

    Navigable waveform built on Web Audio and Canvas

    ...The audio will start playing as you press play. A thin line will be displayed until the whole audio file is downloaded and decoded to draw the waveform. Web Audio needs the whole file to decode it in the browser. You can however load pre-decoded waveform data to draw the waveform immediately. wavesurfer.js runs on modern browsers supporting Web Audio, including Firefox, Chrome, Safari (desktop and mobile) and Opera.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 4
    Downloads: 0 This Week
    Last Update:
    See Project
  • Save Up to 91% on Cloud Compute With Spot VMs Icon
    Save Up to 91% on Cloud Compute With Spot VMs

    Automatic sustained-use discounts. One free VM per month. No negotiation needed.

    Run batch jobs at 60-91% off with Spot VMs. Long-running workloads get automatic discounts with sustained use.
    Try Free
  • 5
    PSLab Android App

    PSLab Android App

    PSLab Android App

    ...This repository holds the Android App for performing experiments with PSLab. PSLab is a tiny pocket science lab that provides an array of equipment for doing science and engineering experiments. It can function like an oscilloscope, waveform generator, frequency counter, programmable voltage and current source and also as a data logger. PSLab is a tiny pocket science lab that provides an array of equipment for doing science and engineering experiments.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 6
    Shutter Encoder

    Shutter Encoder

    A professional video compression tool accessible to all

    ...The software supports numerous codecs and formats, enabling users to prepare media for broadcasting, streaming, or archiving. It includes advanced features such as subtitle creation, waveform visualization, and timecode overlays, making it useful for post-production workflows. Shutter Encoder also allows batch processing and automation of repetitive tasks, improving efficiency when handling large media libraries. Its interface balances simplicity with advanced control, exposing detailed encoding settings while remaining approachable. ...
    Downloads: 19 This Week
    Last Update:
    See Project
  • 7
    JUDI.jl

    JUDI.jl

    Julia Devito inversion

    JUDI is a framework for large-scale seismic modeling and inversion and is designed to enable rapid translations of algorithms to fast and efficient code that scales to industry-size 3D problems. The focus of the package lies on seismic modeling as well as PDE-constrained optimization such as full-waveform inversion (FWI) and imaging (LS-RTM). Wave equations in JUDI are solved with Devito, a Python domain-specific language for automated finite-difference (FD) computations. JUDI's modeling operators can also be used as layers in (convolutional) neural networks to implement physics-augmented deep learning algorithms thanks to its implementation of ChainRules's rrule for the linear operators representing the discre wave equation.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8
    Must Reading on ISAC

    Must Reading on ISAC

    Must Reading Papers, Research Library, Open-Source Code

    A design paradigm and corresponding enabling technologies, in which sensing and comms systems are integrated to efficiently utilize congested wireless/hardware resources, and even to pursue mutual benefits. Waveform Design and Signal Processing Aspects for Fusion of Wireless Communications and Radar Sensing. Supported by IEEE ComSoc ISAC Emerging Technology Initiative (ETI). Integrating Sensing and Communications for Ubiquitous IoT. Include reproducible codes, good papers and a research libary.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 9
    MLX-Audio

    MLX-Audio

    A text-to-speech, speech-to-text and speech-to-speech library

    ...It includes examples such as audiobook generation to demonstrate long-form synthesis and joined audio segments. On top of that, MLX-Audio offers a modern web interface powered by FastAPI, with real-time waveform and 3D visualizations, file upload, and audio management.
    Downloads: 9 This Week
    Last Update:
    See Project
  • Build Agents and Models on One Platform Icon
    Build Agents and Models on One Platform

    Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

    Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
    Try It Free
  • 10
    voice-input-src

    voice-input-src

    Documents and exposes generated source for menu-bar workflows

    ...The application supports English, Simplified Chinese, Traditional Chinese, Japanese, and Korean, with the chosen language stored locally. A floating capsule displays a live waveform and expanding transcription text while recording. Clipboard restoration and temporary input-method switching make text injection safer for CJK users. An optional OpenAI-compatible model can conservatively correct obvious recognition errors before the final text is inserted.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 11
    Praat

    Praat

    Doing Phonetics By Computer

    Praat is a speech analysis application and source repository for doing phonetics by computer. It was created by Paul Boersma and David Weenink at the University of Amsterdam. The software lets researchers, linguists, clinicians, teachers, and students analyze, synthesize, manipulate, and annotate speech recordings. It supports acoustic inspection through waveforms, spectrograms, pitch tracks, formants, intensity, and related phonetic measurements. Praat also includes scripting tools for...
    Downloads: 4 This Week
    Last Update:
    See Project
  • 12
    AERIS-10

    AERIS-10

    Open-source, low-cost 10.5 GHz PLFM phased array RADAR system

    ...The repository structure suggests an emphasis on simulation rather than hardware integration, allowing users to test radar concepts in a controlled software environment. It likely includes tools for waveform synthesis, matched filtering, and spectral analysis, which are critical for interpreting radar returns.
    Downloads: 11 This Week
    Last Update:
    See Project
  • 13
    ThinkDSP

    ThinkDSP

    Digital Signal Processing in Python, by Allen B. Downey

    Think DSP is an educational Python project that teaches digital signal processing through executable examples rather than starting with heavy mathematical formalism. It accompanies Allen B. Downey’s book and organizes most lessons as Jupyter notebooks. Readers work directly with waves, spectra, harmonics, filtering, convolution, and other signal-processing concepts. Early exercises show how to decompose sounds, modify frequency components, and synthesize new audio. The repository includes...
    Downloads: 5 This Week
    Last Update:
    See Project
  • 14
    Furnace

    Furnace

    A multi-system chiptune tracker compatible with DefleMask modules

    Furnace is a powerful multi-system chiptune tracker that enables users to compose music using the sound chips of classic computers, consoles, and arcade hardware. It supports an extensive range of audio chips, including FM synthesis, wavetable synthesis, and sample-based systems, making it one of the most versatile trackers available. The software is compatible with multiple operating systems and can be used both as a standalone application and as a development tool for retro-style audio...
    Downloads: 9 This Week
    Last Update:
    See Project
  • 15
    TorchAudio

    TorchAudio

    Data manipulation and transformation for audio signal processing

    The aim of torchaudio is to apply PyTorch to the audio domain. By supporting PyTorch, torchaudio follows the same philosophy of providing strong GPU acceleration, having a focus on trainable features through the autograd system, and having consistent style (tensor names and dimension names). Therefore, it is primarily a machine learning library and not a general signal processing library. The benefits of PyTorch can be seen in torchaudio through having all the computations be through PyTorch...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 16
    AudioKit

    AudioKit

    Swift audio synthesis, processing, & analysis platform

    AudioKit is an entire audio development ecosystem of code repositories, packages, libraries, algorithms, applications, playgorunds, tests, and scripts, built and used by a community of audio programmers, app developers, engineers, researchers, scientists, musicians, gamers, and people new to programming. An important goal for AudioKit is to allow it to grow and be maintainable by a handful of volunteers. For this reason we have extensive tests that are run whenever changes are made to any...
    Downloads: 4 This Week
    Last Update:
    See Project
  • 17
    uPlot

    uPlot

    Chart for time series, lines, areas, ohlc and bars

    ...However, if you need 60fps performance with massive streaming datasets, uPlot can only get you so far. WebGL should still be the tool of choice for applications like realtime signal or waveform visualizations: See danchitnis/webgl-plot, huww98/TimeChart, epezent/implot, or commercial products like LightningChart®.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 18
    II ElevenLabs UI

    II ElevenLabs UI

    Component library and custom registry built on top of shadcn/ui

    ...Built on top of modern frontend tooling such as React, Tailwind CSS, and shadcn/ui, it provides a collection of pre-built, customizable components that developers can easily integrate into their applications. The library includes specialized UI elements such as audio players, waveform visualizers, conversational interfaces, and interactive voice components, all tailored for building agent-driven experiences. It is tightly aligned with ElevenLabs’ ecosystem, allowing seamless integration with their voice synthesis and conversational AI tools. The project also includes a CLI that simplifies the process of adding components directly into existing projects, streamlining development workflows.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 19
    pyAudioAnalysis

    pyAudioAnalysis

    Python Audio Analysis Library: Feature Extraction, Classification

    pyAudioAnalysis is an open-source Python library designed for audio signal analysis, machine learning, and music information retrieval tasks. The project provides a collection of tools that allow developers to extract meaningful features from audio files and use those features for classification, segmentation, and analysis. The library supports multiple audio processing workflows, including feature extraction from raw audio signals, training of machine learning models, and automatic audio...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 20
    XiangShan

    XiangShan

    Open-source high-performance RISC-V processor

    ...The project invests heavily in verification: differential testing against reference models, extensive random instruction tests, and full software stacks (bootloaders, Linux) to validate correctness under realistic workloads. Tooling around the core (build scripts, simulators, waveform/debug support) lowers the barrier for academics and industry contributors to experiment and extend.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 21
    WavTokenizer

    WavTokenizer

    SOTA discrete acoustic codec models with 40/75 tokens per second

    ...Its architecture incorporates a broader vector-quantization space, extended contextual windows, and improved attention networks, combined with multi-scale discriminators and inverse Fourier transform blocks to enhance waveform reconstruction. Extensive experiments show that WavTokenizer matches or surpasses previous neural codecs across speech, music, and general audio on both objective metrics and subjective listening tests.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22
    WhisperSpeech

    WhisperSpeech

    An Open Source text-to-speech system built by inverting Whisper

    ...Its architecture follows a token-based, multi-stage pipeline inspired by AudioLM and SPEAR-TTS: Whisper is used to produce semantic tokens, EnCodec compresses the waveform into acoustic tokens, and Vocos reconstructs high-fidelity audio from those tokens. The repository includes notebooks and scripts for inference, long-form synthesis, and finetuning, as well as pre-trained models and converted datasets hosted on Hugging Face. Performance optimizations like torch.compile, KV-caching, and architectural tweaks allow the main model to reach up to 12× real-time speed on a consumer RTX 4090.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 23
    Step-Audio-EditX

    Step-Audio-EditX

    LLM-based Reinforcement Learning audio edit model

    Step-Audio-EditX is an open-source, 3 billion-parameter audio model from StepFun AI designed to make expressive and precise editing of speech and audio as easy as text editing. Rather than treating audio editing as low-level waveform manipulation, this model converts speech into a sequence of discrete “audio tokens” (via a dual-codebook tokenizer) — combining a linguistic token stream and a semantic (prosody/emotion/style) token stream — thereby abstracting audio editing into high-level token operations. This allows users to modify not only what is said (the text) but also how it's said: emotion, tone, speaking style, prosody, accent, even paralinguistic cues. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 24
    VisualSubSync 2.x.x.x

    VisualSubSync 2.x.x.x

    Professional subtitle editor with waveform sync

    VisualSubSync 2.x.x.x - First release from fansubbers.gr This is a complete modernization of VisualSubSync: - Ported to Delphi 12.1 for Windows 10/11 optimization - Modern user interface with improved waveform rendering - Full Unicode support for international subtitles - Enhanced JavaScript QC plugins system - Greek language support and Greek spell-check rules - Improved compatibility with all modern Windows versions - Performance enhancements and bug fixes - All original features preserved from the classic VisualSubSync Based on the original VisualSubSync by Christophe Paris and VisualSubSync Enhanced by Red5goahead, now maintained by fansubbers.gr Website: https://www.fansubbers.gr Privacy Policy: https://fansubbers.gr/help/vss-privacy/
    Downloads: 49 This Week
    Last Update:
    See Project
  • 25

    Subtitle-Workshop-Classic-v6.3.4

    Subtitle Editor derived from 6.0c, but with VLC and Hunspell checker

    Audio waveform, VLC Video Renderer, UTF8 coding, Audio stream detection and Selection, Resizeable screens, Hunspell spellcheck, Easy shortcut editing, user profiles and more than 70 filetypes supported.
    Leader badge
    Downloads: 98 This Week
    Last Update:
    See Project