high free download - SourceForge

whisper.cpp

Port of OpenAI's Whisper model in C/C++

whisper.cpp is a lightweight, C/C++ reimplementation of OpenAI’s Whisper automatic speech recognition (ASR) model—designed for efficient, standalone transcription without external dependencies. The entire high-level implementation of the model is contained in whisper.h and whisper.cpp. The rest of the code is part of the ggml machine learning library. The command downloads the base.en model converted to custom ggml format and runs the inference on all .wav samples in the folder samples. whisper.cpp supports integer quantization of the Whisper ggml models. ...

Downloads: 446 This Week

Last Update: 2026-06-19

See Project

FireRedASR

Open-source industrial-grade ASR models

FireRedASR is an industrial-grade family of open-source automatic speech recognition models designed to provide high-precision speech-to-text performance across languages including Mandarin, English, and various Chinese dialects, achieving new state-of-the-art benchmarks on public test sets. The project includes multiple model variants to meet different application needs, such as high-accuracy end-to-end interaction using an encoder-adapter-LLM framework and efficient real-time recognition using attention-based encoder-decoder architectures, giving developers flexibility in balancing performance and resource constraints. ...

Downloads: 0 This Week

Last Update: 2026-02-25

See Project

Omnilingual ASR

Omnilingual ASR Open-Source Multilingual SpeechRecognition

Omnilingual-ASR is a research codebase exploring automatic speech recognition that generalizes across a very large number of languages using shared modeling and training recipes. It focuses on leveraging self-supervised audio pretraining and scalable fine-tuning so low-resource languages can benefit from high-resource data. The project provides data preparation pipelines, training scripts, decoding utilities, and evaluation tools so researchers can reproduce results and extend to new language sets. It emphasizes modularity: acoustic modeling, language modeling, tokenization, and decoding are separable pieces you can swap or ablate. The repo is aimed at pushing practical multilingual ASR—robust to accents, code-switching, and domain shifts—rather than language-by-language systems. ...

Downloads: 0 This Week

Last Update: 2025-12-12

See Project

WhisperJAV

A subtitle generator for Japanese Adult Videos.

A subtitle generator for Japanese Adult Videos. Transformer-based ASR architectures like Whisper suffer significant performance degradation when applied to the spontaneous and noisy domain of JAV. This degradation is driven by specific acoustic and temporal characteristics that defy the statistical distributions of standard training data.

1 Review

Downloads: 46 This Week

Last Update: 3 days ago

See Project

Scribe

Free, open-source, and offline speech-to-text & voice control app.

... > Designed with privacy as a top priority, Scribe works completely offline. Your voice data never leaves your computer. Powered by the Vosk engine, it supports multiple languages and provides high-quality recognition without an internet connection. > Scribe is the perfect tool for anyone looking to boost productivity, improve accessibility, or simply interact with their computer in a new, hands-free way.

Downloads: 65 This Week

Last Update: 2025-12-13

See Project

Flow Teleprompter

A Windows first teleprompter with voice tracking and AI drafting.

Flow is an ultra-lightweight, high-performance desktop teleprompter built with Tauri and Rust. It is designed for creators & presenters who need a clean reading surface without sacrificing advanced features like native voice tracking via local Vosk models and app-wide voice commands. It features five playback styles (highlight, scroll, line, arrow, and voice tracking), local-first privacy, and a built-in script editor.

Downloads: 25 This Week

Last Update: 2026-06-05

See Project

Flashlight library

A C++ standalone library for machine learning

Flashlight is a fast, flexible machine learning library written entirely in C++ by Facebook AI Research and the creators of Torch, TensorFlow, Eigen, and Deep Speech. Native support in C++ and simple extensibility make Flashlight a powerful research framework that's hackable to its core and enables fast iteration on new experimental setups and algorithms with little unopinionated and without sacrificing performance. In a single repository, Flashlight provides apps for research across...

Downloads: 3 This Week

Last Update: 2022-05-27

See Project

VideoSrt

Windows-GUI

...Recognize video/audio speech to generate subtitle files (support Chinese-English translation, bilingual subtitles) Extract speech text from video/audio. Batch translation, filter processing/encoding SRT subtitle files. Using the Alibaba Cloud speech recognition interface, the accuracy is high, and the standard Mandarin/English recognition rate is over 95%. Video recognition does not need to upload the original video, which is convenient, fast and time-saving.

Downloads: 12 This Week

Last Update: 2023-01-13

See Project

High-order HMM in Matlab

Implementation of duration high-order hidden Markov model in Matlab.

Implementation of duration high-order hidden Markov model (DHO-HMM) in Matlab with application in speech recognition.

2 Reviews

Downloads: 0 This Week

Last Update: 2015-02-15

See Project

Speech Recognition System

Speech Recognition System - Matlab source code

Speech recognition technology is used more and more for telephone applications like travel booking and information, financial account information, customer service call routing, and directory assistance. Using constrained grammar recognition, such applications can achieve remarkably high accuracy. Research and development in speech recognition technology has continued to grow as the cost for implementing such voice-activated systems has dropped and the usefulness and efficacy of these systems has improved. For example, recognition systems optimized for telephone applications can often supply information about the confidence of a particular recognition, and if the confidence is low, it can trigger the application to prompt callers to confirm or repeat their request. ...

Downloads: 4 This Week

Last Update: 2015-03-18

See Project

High-order HMM in Java

A duration high-order hidden Markov model (DHO-HMM) in Java.

This project provides an implementation of duration high-order hidden Markov model (DHO-HMM) in Java. It is compactible with JDK 5 & 6. It was used in the author's research on speech recognition of Mandarin digits. There are some Chinese words in this project and I am afraid that I don't have enough time to translate to English recently.

Downloads: 0 This Week

Last Update: 2013-09-16

See Project

Search Results for "high"

Showing 11 open source projects for "high"

whisper.cpp

FireRedASR

Omnilingual ASR

WhisperJAV

Scribe

Flow Teleprompter

Flashlight library

VideoSrt

High-order HMM in Matlab

Speech Recognition System

High-order HMM in Java

Search Results for "high"

Showing 11 open source projects for "high"

whisper.cpp

FireRedASR

Omnilingual ASR

WhisperJAV

Scribe

Flow Teleprompter

Flashlight library

VideoSrt

High-order HMM in Matlab

Speech Recognition System

High-order HMM in Java

Related Searches

Related Categories