Showing 31 open source projects for "tamil audio transcript"

View related business solutions
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • Veeam Data Platform v13.1 - Get Your Free Trial Icon
    Veeam Data Platform v13.1 - Get Your Free Trial

    Secure by design, portable by default. Recover clean, fast, anywhere. Start a free trial.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try it Free
  • 1
    quill macOS

    quill macOS

    Ultraminimalist macOS recording + transcription

    Quill is a minimalist macOS meeting recorder and transcription utility that runs from the menu bar. It captures microphone input and system audio as separate tracks, which naturally distinguishes the user from other speakers. When recording stops, both tracks are transcribed locally and merged into a timestamped, speaker-labeled transcript. The app uses an on-device Parakeet model, so recordings and text never need to leave the Mac. Sessions are stored as audio, metadata, transcript, and log files in organized folders. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 2
    Podcastfy.ai

    Podcastfy.ai

    Transforming Multimodal Content into Captivating Multilingual Audio

    Podcastfy is an open-source Python package that transforms multi-modal content (text, images) into engaging, multi-lingual audio conversations using GenAI. Input content includes websites, PDFs, youtube videos as well as images. Unlike UI-based tools focused primarily on note-taking or research synthesis (e.g. NotebookLM), Podcastfy focuses on the programmatic and bespoke generation of engaging, conversational transcripts and audio from a multitude of multi-modal sources enabling...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 3
    Buzz

    Buzz

    Buzz transcribes and translates audio offline

    Buzz is a desktop application for transcribing and translating audio locally with speech recognition models based on Whisper. It can process audio files, video files, and YouTube links without requiring cloud transcription. Live microphone transcription supports real-time captions and a presentation view for accessible events. Speech separation can improve results on noisy recordings, while speaker identification distinguishes voices within transcribed media. Multiple Whisper backends,...
    Downloads: 662 This Week
    Last Update:
    See Project
  • 4
    GladiaFlow

    GladiaFlow

    A desktop app for real-time voice dictation

    GladiaFlow is an open-source desktop app for turning spoken language into text in virtually any text field. It captures microphone audio through a global hotkey and streams it to Gladia’s Live Transcription API. Partial and final results arrive in real time, then the app cleans punctuation, capitalization, and duplicate text before pasting the transcript into the active application. Users can choose push-to-talk or toggle activation and configure languages, code switching, vocabulary, and pronunciations. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • Earn up to 16% annual interest with Nexo. Icon
    Earn up to 16% annual interest with Nexo.

    Let your crypto work for you

    Put idle assets to work with competitive interest rates, borrow without selling, and trade with precision. All in one platform. Geographic restrictions, eligibility, and terms apply.
    Get started with Nexo.
  • 5
    claude-video

    claude-video

    Give Claude the ability to watch any video

    Claude Video is an agent skill that gives Claude and compatible coding assistants the ability to analyze video content. It accepts public video URLs or local video files, then extracts the information needed to answer user questions about what happened on screen and in the audio. The workflow checks captions first, downloads only what is necessary, extracts timestamped frames, and produces a transcript through native captions or Whisper fallback. It supports different detail levels so users can trade speed, token cost, and visual coverage depending on the task. The skill is useful for summarizing videos, reviewing screen recordings, analyzing content structure, diagnosing visual bugs, and turning course material into notes. ...
    Downloads: 12 This Week
    Last Update:
    See Project
  • 6
    claude-real-video

    claude-real-video

    Let Claude (or any LLM) actually watch a video

    claude-real-video is a local video processing tool that prepares videos for analysis by Claude or other LLMs. It accepts public video URLs or local files and extracts the visual and audio evidence an AI model needs. Instead of sampling at a fixed interval, it detects scene changes and removes near-duplicate frames to reduce unnecessary token usage. It can transcribe audio, generate a clean output folder, and create a local viewer with the video, keyframe grid, and transcript. The tool can also be installed as a Claude Code skill so an agent can process videos more directly. ...
    Downloads: 9 This Week
    Last Update:
    See Project
  • 7
    video-use

    video-use

    Edit videos with Claude Code

    Video Use is an open-source AI-powered video editing tool that allows users to transform raw footage into polished videos using natural language commands. Designed to work with Claude Code, it automates the entire editing process—from cutting clips to rendering the final output—without requiring manual timelines or complex software interfaces. The system intelligently analyzes audio transcripts and visual cues to make precise, context-aware editing decisions. It supports a wide range of...
    Downloads: 19 This Week
    Last Update:
    See Project
  • 8
    Meetily

    Meetily

    Privacy first, AI meeting assistant with 4x faster Parakeet/Whisper

    This project is a privacy-first AI meeting assistant that captures meeting audio, produces real-time transcripts, and generates summaries while keeping processing entirely on your own machine or infrastructure. It’s built for organizations that want meeting intelligence without sending recordings or transcripts to third-party cloud services, which helps address compliance and data sovereignty requirements. The app supports live transcription with local model options (including Whisper- and Parakeet-based workflows) and presents the transcript as the meeting happens, making it useful both for note-taking and accessibility. ...
    Downloads: 16 This Week
    Last Update:
    See Project
  • 9
    LARA is software for musical analysis using (new) scientific methods for analysis and visualization. LARA is part of the core research: “Interpretation and performance” of the HSLU – Musik (University of Applied Sciences Luzern – Music depart
    Downloads: 5 This Week
    Last Update:
    See Project
  • AI-generated apps that pass security review Icon
    AI-generated apps that pass security review

    Stop waiting on engineering. Build production-ready internal tools with AI—on your company data, in your cloud.

    Retool lets you generate dashboards, admin panels, and workflows directly on your data. Type something like “Build me a revenue dashboard on my Stripe data” and get a working app with security, permissions, and compliance built in from day one. Whether on our cloud or self-hosted, create the internal software your team needs without compromising enterprise standards or control.
    Try Retool free
  • 10
    HearWrite PDF

    HearWrite PDF

    Local audio-to-PDF transcription for Windows and Debian.

    HearWrite PDF is a privacy-focused desktop transcription application created by Don Fritz. It converts a single audio recording or an entire folder of recordings into polished PDF and editable TXT transcripts. For folders containing multiple recordings, users can arrange files by date and time or filename, create one combined transcript, individual transcripts, or both, and optionally organize the combined document into calendar-date sections.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 11
    TranscribeGeek

    TranscribeGeek

    Free offline Whisper transcription for Windows - no upload, no limit

    TranscribeGeek turns audio and video recordings into text on your own machine. It runs OpenAI Whisper models locally, writes a transcript and an optional SubRip .srt subtitle file, and has no account, no upload and no per-minute limit. Windows 10 1809 or later, 64-bit. GPL-3.0.
    Downloads: 8 This Week
    Last Update:
    See Project
  • 12
    PyTube Downloader

    PyTube Downloader

    Let's quickly download YouTube videos & playlists with one click.

    PyTube Downloader lets you quickly download YouTube videos and playlists with one click. Choose from 144p to 8K quality and download multiple videos simultaneously. PyTube Downloader 让你轻松一键下载 YouTube 视频和播放列表。支持选择 144p 到 8K 的质量,并支持同时下载多个视频。
    Leader badge
    Downloads: 191 This Week
    Last Update:
    See Project
  • 13
    CC2.TV / CC2 - Audio- und TV-Datenbank

    CC2.TV / CC2 - Audio- und TV-Datenbank

    Meta-Datenbank-Anwendung für die Audio- und TV-Sendungen des CC2.TV

    Dieses Programm stellt eine Meta-Datenbank-Anwendung für die Audio- und Video-Sendungen des CC2.TV für GNU/Linux Systeme zur Verfügung. Es ermöglicht das Durchsuchen, Verwalten und Abspielen der umfangreichen Inhalte des CC2.TV-Audiocasts und -Videocasts. Ziel ist es, die über 3000 Audiocast-Themen und über 1000 Videocast-Themen, die sich auf Computerthemen, Technik und gesellschaftliche Aspekte konzentrieren, komfortabel zugänglich zu machen. Für die volle Funktionalität,...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 14
    AudioEnhancerMAX

    AudioEnhancerMAX

    Local-first AI audio processing, transcription and mastering

    AudioEnhancerMAX 3.7 is an open-source, local-first audio production suite for podcasters, creators, journalists, educators, researchers, and developers. It combines deterministic Smart Enhance diagnosis, non-destructive transcript-linked speech editing, audio cleanup and mastering, local Faster-Whisper transcription, delivery profiles, review tools, TTS, system monitoring, and optional trusted-LAN Android workers.
    Leader badge
    Downloads: 26 This Week
    Last Update:
    See Project
  • 15
    Meeting AI Analyser

    Meeting AI Analyser

    Live AI co-pilot for Windows meetings: local Whisper + Claude AI.

    Meeting AI Analyser is a Windows desktop app that acts as a live AI co-pilot during any meeting. It captures system audio and microphone on your PC and transcribes everything in real time using OpenAI Whisper locally, so your audio never leaves your machine. Only the resulting text transcript is sent to Claude AI via the Claude Code CLI to generate structured summaries every 60 seconds with decisions, action items, participants and next steps. Claude can also explain unfamiliar jargon and translate in real time. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 16
    MARS5

    MARS5

    MARS5 speech model (TTS) from CAMB.AI

    ...The model is built to handle prosodically challenging content such as sports commentary, anime dialogue, and other high-energy or highly varied speech patterns with realistic rhythm and intonation. To control speaker identity, MARS5 uses a short reference audio clip, typically between 2 and 12 seconds, from which it learns the voice characteristics. It supports two main inference modes: shallow clone, which is faster and only needs the reference audio, and deep clone, which additionally uses the transcript of the reference audio to increase similarity and naturalness at the cost of more computation.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 17
    Ainee

    Ainee

    Ainee - AI Notetaking and Learning Companion

    Ainee is your ultimate AI-powered notetaking and learning companion. Capture lecture notes in real-time and effortlessly transform audio, text, files, and YouTube videos into formatted notes, mindmaps, quizzes, flashcards, podcasts, and more. Explore our AI meeting note taker, AI notes, video transcript generator, PDF to AI converter, and AI flashcard maker. Enhance your learning with our AI voice recorder, article summarizer AI, and AI quiz generator. Additionally, share your knowledge base with others to foster the flow of information and help new users benefit from collective insights. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 18
    footswitch2basic

    footswitch2basic

    Audio Transcription software for Linux (Vlc) with a foot pedal

    Footswitch 2 (Basic) is a media player for transcribers on Linux. This version is a stripped down version of Footswitch2, containing only the absolute essentials for transcription. Written in python and using the python bindings for VLC it allows a transcriber to control the audio or video with a footpedal, and includes a set of macros that integrate into LibreOffice. This allows the transcriber to control the media player from within Libreoffice as well, making it useful for those who do...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 19
    Carnatic Music Guru / JRaaga

    Carnatic Music Guru / JRaaga

    Carnatic Music Guru - JRaaga

    VISIT THIS PAGE AS MORE FEATURES ARE BEING ADDED. If you have downloaded 2.03 or above - Use Help>Check Updates to download latest version. If that does not update: Download CMGUpdater from here: update/CMGUpdater.jar Copy it to the update/ folder of your JRaaga installation path. Try again Help>Check for updates. If it does not work. Delete existing installation. Download the latest version and try. Carnatic Music Guru is a tutor/player/lesson generator. YOU NEED Java Runtime...
    Downloads: 7 This Week
    Last Update:
    See Project
  • 20
    Defox text to speech and downloader

    Defox text to speech and downloader

    Written or imported text offline read or online download.

    This software design to convert text to speech and download the converted speech. Description : • Installation setup with two languages (English, French) • Two areas called text reading and speech downloading • Many languages supported to download center Note 1: I'm a student yet and I'm not in the software designing industry. Therefore maybe I haven't software making skills. I'm worried about that. ! Note 2 : When you double click on the software maybe it will get some seconds...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21
    Naam Tamilar HTML 5 Shoutcast Player

    Naam Tamilar HTML 5 Shoutcast Player

    Dedicated to Naam Tamilar Political Party

    Experienced Many service blocks by other company owned players and decided to have our own coded html5 audio player with the saying "தற்சார்பு " - "tharsaarpu" fulfill your service delivery with what you have ? Dedicated to Naam Tamil Political Party
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22
    Tamil Radio

    Tamil Radio

    tamil radio, stream tamil songs from online,tamil mp3 player

    A simple java based mp3 player which streams and plays tamil songs from online.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 23
    Transcription Aid

    Transcription Aid

    Transcription Aid helps you type text from recordings.

    This software is to help type in text from speech recordings. It has several functions proven to help this type of work. However it is fully manual (aside from auto-completion), so no speech recognition if you are looking for that, but it is a great tool to do the job.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 24

    Rhema STH

    Free Open Source Software for the Speech & Hearing Impaired

    RHEMA - Speak to Hear Software Application RHEMA is a software designed to help people with speech disability. Thiruvalluvar, the Tamil Sage of the 1st Century CE had said: “Wealth of wealth is wealth acquired be ear attent; Wealth mid all wealth supremely excellent. “ Kural No : 411 This software is the first version, with limited words in Tamil for them to practice. We have tested it with the help of a school and atleast two children were able to pick up some...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 25
    Dhvani is Text-to-Speech System for Indic Languages. Current C- GNU/Linux implementation supports Hindi, Kannada, Marathi, Malayalam, Gujarati, Bengali, Telugu, Panjabi, Tamil and Oriya.
    Downloads: 1 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • 2
  • Next