Showing 2 open source projects for "tasks"

View related business solutions
  • $300 Free Credits to Build on Google Cloud Icon
    $300 Free Credits to Build on Google Cloud

    New customers can spin up VMs, build with AI, and query data at no cost.

    Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
    Start Free
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • 1
    Whisper

    Whisper

    Robust Speech Recognition via Large-Scale Weak Supervision

    ...It is trained on a large dataset of diverse audio and is also a multitasking model that can perform multilingual speech recognition, speech translation, and language identification. A Transformer sequence-to-sequence model is trained on various speech processing tasks, including multilingual speech recognition, speech translation, spoken language identification, and voice activity detection. These tasks are jointly represented as a sequence of tokens to be predicted by the decoder, allowing a single model to replace many stages of a traditional speech-processing pipeline. The multitask training format uses a set of special tokens that serve as task specifiers or classification targets.
    Downloads: 68 This Week
    Last Update:
    See Project
  • 2
    VoiceInk

    VoiceInk

    The best open-source alternative to Superwhisper & Wispr Flow

    ...It can process audio entirely on the device with local Whisper or Parakeet models, giving users a private offline workflow. Optional AI enhancement can clean up, rewrite, format, or adapt dictated text for different tasks. Modes apply different models, prompts, context settings, and shortcuts according to the active app or website. A personal dictionary improves recognition of names, technical terms, and recurring phrases, while configurable global shortcuts support toggle and push-to-talk recording. Context-aware tools can use selected text, clipboard content, or visible screen text to improve the final output. ...
    Downloads: 4 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • Next