Showing 114 open source projects for "ffdshow audio decoder"

View related business solutions
  • Ship Agents Faster Icon
    Ship Agents Faster

    Transform your applications and workflows into powerful agentic systems at global scale.

    Gemini Enterprise Agent Platform lets you rapidly build, scale, govern and optimize production-ready agents grounded in your organization's data. The platform enables developers to build custom or pre-built agents for virtually any use case. New customers get $300 in free credits.
    Start Free
  • Veeam Data Platform v13.1 - Get Your Free Trial Icon
    Veeam Data Platform v13.1 - Get Your Free Trial

    Secure by design, portable by default. Recover clean, fast, anywhere. Start a free trial.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try it Free
  • 1
    Super Audio CD Decoder
    Super Audio CD Decoder input plugin for foobar2000. Decoder is capable of playing back Super Audio CD ISO images, DSDIFF, DSF and DSD WavPack files. DSD(DoP) and PCM output modes. Separate DSD Processor/DSD Converter plugins for track extraction into DSD/DST encoded files.
    Leader badge
    Downloads: 4,469 This Week
    Last Update:
    See Project
  • 2
    DVD-Audio Decoder and Watermark Detector
    DVD-Audio Decoder input plugin and Watermark Detector/Neutralizer DSP plugins for foobar2000. Decoder is capable of playing back DVD-Audio discs, ISO images, AOB, MLP and Dolby TrueHD files in full resolution. Dedicated plugin for DTS-HD playback. APT-x100 plugin for *.AUD and *.AUE files from DTS Movie/Trailer Discs.
    Leader badge
    Downloads: 207 This Week
    Last Update:
    See Project
  • 3
    Miso TTS

    Miso TTS

    Miso TTS is an 8 billion, highly emotive text-to-speech model

    Miso TTS is an advanced 8-billion-parameter text-to-speech model developed by Miso Labs for generating highly expressive and natural-sounding conversational speech. Built on an RVQ Transformer architecture inspired by Sesame CSM, it combines a powerful Llama-based backbone with an autoregressive audio decoder to produce high-quality audio from text. The model supports both standard speech synthesis and voice-conditioned generation using optional audio prompts for voice cloning. Miso TTS generates Mimi audio codes and can leverage conversation history to create more contextually aware and realistic dialogue. Designed for local deployment, it offers watermarking by default to help promote responsible use of generated audio. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 4
    h3-metal

    h3-metal

    MiniMax H3 inference engine for Mac computers

    h3-metal is a native MiniMax-H3 inference engine for Apple Silicon focused on generating video and audio locally with Metal. It supports prompt-to-video and prompt-to-audio workflows as well as first-frame, last-frame, image, video, and audio references. An interactive terminal session keeps prompt conditioning, the diffusion transformer, and the video decoder in memory for faster repeated generations. Users can trade speed, quality, and memory through denoising steps, layer counts, reuse modes, token reduction, internal render size, and SSD streaming. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • Host LLMs in Production With On-Demand GPUs Icon
    Host LLMs in Production With On-Demand GPUs

    NVIDIA L4 GPUs. 5-second cold starts. Scale to zero when idle.

    Deploy your model, get an endpoint, pay only for compute time. No GPU provisioning or infrastructure management required.
    Start Free
  • 5
    Whisper

    Whisper

    Robust Speech Recognition via Large-Scale Weak Supervision

    ...These tasks are jointly represented as a sequence of tokens to be predicted by the decoder, allowing a single model to replace many stages of a traditional speech-processing pipeline. The multitask training format uses a set of special tokens that serve as task specifiers or classification targets.
    Downloads: 58 This Week
    Last Update:
    See Project
  • 6
    Levyra

    Levyra

    Open-source music player for Android & Windows

    Levyra is an open-source music player for Android and Windows designed around native playback, privacy, and user-controlled music libraries. It combines streaming, local music, downloads, discovery, live radio, synchronized lyrics, and music recognition within a single application. Levyra uses Media3/ExoPlayer on Android and libvlc on Windows, providing native playback rather than wrapping a web player inside the app. The platform supports advanced audio capabilities including gapless...
    Downloads: 30 This Week
    Last Update:
    See Project
  • 7
    TorchAudio

    TorchAudio

    Data manipulation and transformation for audio signal processing

    The aim of torchaudio is to apply PyTorch to the audio domain. By supporting PyTorch, torchaudio follows the same philosophy of providing strong GPU acceleration, having a focus on trainable features through the autograd system, and having consistent style (tensor names and dimension names). Therefore, it is primarily a machine learning library and not a general signal processing library. The benefits of PyTorch can be seen in torchaudio through having all the computations be through PyTorch...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8
    XLD

    XLD

    A tool for transcoding lossless audio files

    X Lossless Decoder(XLD) is a tool for Mac OS X that is able to decode/convert/play various 'lossless' audio files. The supported audio files can be split into some tracks with cue sheet when decoding. It works on Mac OS X 10.4 and later.
    Leader badge
    Downloads: 4,081 This Week
    Last Update:
    See Project
  • 9
    IndexTTS2

    IndexTTS2

    Industrial-level controllable zero-shot text-to-speech system

    ...It builds on state-of-the-art models such as XTTS and other modern neural TTS backbones, improving them with a conformer-based speech conditional encoder and upgrading the decoder to a high-quality vocoder (BigVGAN2), leading to clearer and more natural audio output. The system supports zero-shot voice cloning — meaning it can mimic a target speaker’s voice from a short reference sample — making it versatile for multi-voice uses. Compared to many open-source TTS tools, IndexTTS emphasizes efficiency and controllability: it offers faster inference, simpler training pipelines, and controllable speech parameters (like duration, pitch, and prosody), which is critical for production use.
    Downloads: 10 This Week
    Last Update:
    See Project
  • Save Up to 91% on Cloud Compute With Spot VMs Icon
    Save Up to 91% on Cloud Compute With Spot VMs

    Automatic sustained-use discounts. One free VM per month. No negotiation needed.

    Run batch jobs at 60-91% off with Spot VMs. Long-running workloads get automatic discounts with sustained use.
    Start Free
  • 10
    BlackBelt CodecPack

    BlackBelt CodecPack

    A clean, lean CoDec Pack. FFDShow and LAV Combined.

    Contains support for popular formats. Works especially well with MediaPortal. LAV, ffdshow - why choose between when you can have both in one pack ! WMV/WMA, DivX, AVI, ASF, FLV, Ogg FLAC, HEV1, x264, x265 etc. NO SPYWARE, NO ADWARE, NO TOOLBARS, NO PLAYER - JUST PURE CODECS Windows XP / Vista / 7 / 8 / 10 - 32/64 bit.
    Downloads: 4 This Week
    Last Update:
    See Project
  • 11

    opencore-amr

    Audio codecs extracted from Android Open Source Project

    Library of OpenCORE Framework implementation of Adaptive Multi Rate Narrowband and Wideband (AMR-NB and AMR-WB) speech codec. Library of VisualOn implementation of Adaptive Multi Rate Wideband (AMR-WB) encoder and Advanced Audio Coding (AAC) encoder. Modified library of Fraunhofer AAC decoder and encoder.
    Leader badge
    Downloads: 10,897 This Week
    Last Update:
    See Project
  • 12
    fleck

    fleck

    mp3 decoder with multi-thread support

    MP3 decoder for Linux using libmad with multi-thread support, ideia based on ffmpeg Developed by DeepSeek
    Downloads: 0 This Week
    Last Update:
    See Project
  • 13

    pmaudio

    Precise MPEG Audio

    Precise MPEG Audio Decoder - Open source (GPL) - Small - Fast - Very Precise and Very Accurate - Floating-point and Fixed-point varieties - Works with Linux and Windows - Examples for using the library - Sample Input DLL for WinAmp - Sample command-line player - Decoding library derived from mpg123
    Downloads: 3 This Week
    Last Update:
    See Project
  • 14
    Multimodal

    Multimodal

    TorchMultimodal is a PyTorch library

    ...The library provides modular building blocks such as encoders, fusion modules, loss functions, and transformations that support combining modalities (vision, text, audio, etc.) in unified architectures. It includes a collection of ready model classes—like ALBEF, CLIP, BLIP-2, COCA, FLAVA, MDETR, and Omnivore—that serve as reference implementations you can adopt or adapt. The design emphasizes composability: you can mix and match encoder, fusion, and decoder components rather than starting from monolithic models. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 15
    Tabuleiro

    Tabuleiro

    mp3 decoder with multi-thread support

    MP3 decoder for Linux using libmad with multi-thread support, ideia based on ffmpeg Developed by DeepSeek After 23/may/2026 the latest project files are in the section Files and in the Code section only older versions will be available
    Downloads: 0 This Week
    Last Update:
    See Project
  • 16

    jflac-kse

    FLAC decoder for Java 8

    Java decoder for FLAC sound files. This project arrives to update the jFLAC package 1.5.2 to relieve nasty bugs and where possible improve program code appearance. Currently it has only practical intentions of usability and doesn't target to review the codec algorithm. The target format is Java 8 for maximum compatibility.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 17
    CSM (Conversational Speech Model)

    CSM (Conversational Speech Model)

    A Conversational Speech Generation Model

    The CSM (Conversational Speech Model) is a speech generation model developed by Sesame AI that creates RVQ audio codes from text and audio inputs. It uses a Llama backbone and a smaller audio decoder to produce audio codes for realistic speech synthesis. The model has been fine-tuned for interactive voice demos and is hosted on platforms like Hugging Face for testing. CSM offers a flexible setup and is compatible with CUDA-enabled GPUs for efficient execution.
    Downloads: 5 This Week
    Last Update:
    See Project
  • 18
    MMC is a commander-style media player for Windows, with native, hw accelerated video playing and translucent gui. Mpxplay is a console audio player for DOS and Win32 operating systems. x264vfw, x265vfw and xAV1vfw are video for windows encoder and decoder codecs, useful with VirtualDub.
    Leader badge
    Downloads: 190 This Week
    Last Update:
    See Project
  • 19

    MPD

    Music Player Daemon SACD/DVD-A ISO decoder plugins

    Downloads: 5 This Week
    Last Update:
    See Project
  • 20
    jamailmar

    jamailmar

    Ogg Vorbis decoder with multi-thread support

    Linux code to decode Ogg Vorbis files with multi-thread support Developed by DeepSeek
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21
    libTiMidity is a MIDI to WAVE converter library that uses Gravis Ultrasound-compatible patch files to generate digital audio data from General MIDI files. This library based on the TiMidity decoder from SDL_sound library.
    Leader badge
    Downloads: 14 This Week
    Last Update:
    See Project
  • 22
    This project aims to create a DVD player for Linux and the Creative DXR3 (aka Sigma Designs Hollywood+) MPEG2 decoder board
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23
    SonicTree

    SonicTree

    Folder-based music player

    SonicTree is a Linux music player for local music playback and filesystem-based navigation. It allows you to browse and play music directly from your existing folder structure without requiring a separate music library database. * Version - 1.4.0 * Platform - Linux * Architecture - x86_64 (64-bit) Whats new in 1.4.0 SonicTree 1.4.0 introduces 2 new ways to browse your music while keeping the familiar folder-based approach at its core. - Default File Mode and optional Metadata...
    Leader badge
    Downloads: 29 This Week
    Last Update:
    See Project
  • 24
    Nexus Ham Radio

    Nexus Ham Radio

    All-mode ham radio operations center: 9 modes TX, APRS, awards

    ...Speaks WSJT-X's UDP protocol byte-for-byte, so GridTracker and JTAlert keep working. Its FT8/FT4 decode floor measures -21.3 dB against stock WSJT-X's -20.7 on identical audio. Windows, Linux, Raspberry Pi. GPL-3.0, by KD9TAW.
    Leader badge
    Downloads: 270 This Week
    Last Update:
    See Project
  • 25
    Pakö

    Pakö

    Pakö 2 - FT8, FT4 and JTTY for amateur radio. Free software (GPL v3).

    Pakö ("to communicate" in the Bribri language of Costa Rica) is a free, open-source program for amateur radio digital modes: FT8, FT4 and the new JTTY mode. It is a simple, friendly alternative for operators who want to get on the air quickly: pick a station, double-click, and Pakö completes the QSO and logs it. Version 2 adds JTTY (keyboard-to-keyboard, RTTY-style operation without time slots, using the engine from WSJT-X 3.2), editable macros, rig control through Hamlib, new/worked...
    Downloads: 36 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • 2
  • 3
  • 4
  • 5
  • Next