1348 projects for "can" with 2 filters applied:

  • $300 Free Credits for Your Google Cloud Projects Icon
    $300 Free Credits for Your Google Cloud Projects

    Start building on Google Cloud with $300 in free credits. No commitment, no credit card required until you're ready to scale.

    Launch your next project with $300 in free Google Cloud credits—no strings attached. Test, build, and deploy without risk. Use your credits across the entire Google Cloud platform to find what works best for your needs. After your credits are used, continue with always-free tier services. Only pay when you're ready to scale. Sign up in minutes and start exploring.
    Start Free Trial
  • Build Agents and Models on One Platform Icon
    Build Agents and Models on One Platform

    Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

    Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
    Try It Free
  • 1
    kokoro-onnx

    kokoro-onnx

    TTS with kokoro and onnx runtime

    kokoro-onnx is a text-to-speech toolkit that wraps the Kokoro neural TTS model in an easy-to-use ONNX Runtime interface, so you can generate speech from Python with minimal setup. It focuses on running efficiently on commodity hardware, including macOS with Apple Silicon, while still delivering near real-time performance for many use cases. The project ships prebuilt model files and a simple example script, so you can go from installation to producing an audio.wav file in just a few steps. ...
    Downloads: 623 This Week
    Last Update:
    See Project
  • 2
    Humanizer Skill

    Humanizer Skill

    Claude Code skill that removes signs of AI-generated writing from text

    ...It provides a suite of algorithms that convert timestamps, identifiers, file sizes, code tokens, and structured data into phrases that resemble typical human phrasing rather than compact machine output. For example, date and time values can be expressed as relative terms (“two hours ago”), and file sizes can be shown in intuitive units instead of raw bytes. It also includes functions for transforming camelCase, snake_case, or PascalCase identifiers into spaced and capitalized representations suitable for user interfaces, reports, or documentation. Beyond text formatting, the library can handle pluralization, enumeration formatting (“A, B, and C”), and token expansion so that program-generated content feels more conversational.
    Downloads: 223 This Week
    Last Update:
    See Project
  • 3
    whisper.cpp

    whisper.cpp

    Port of OpenAI's Whisper model in C/C++

    ...The command downloads the base.en model converted to custom ggml format and runs the inference on all .wav samples in the folder samples. whisper.cpp supports integer quantization of the Whisper ggml models. Quantized models require less memory and disk space and depending on the hardware can be processed more efficiently.
    Downloads: 1,434 This Week
    Last Update:
    See Project
  • 4
    Multica

    Multica

    The open-source managed agents platform

    ...Multica emphasizes collaboration between humans and AI by allowing agents to operate alongside developers in shared workspaces. It also supports reusable skill accumulation, meaning that solutions generated by agents can be reused across projects to improve efficiency over time.
    Downloads: 392 This Week
    Last Update:
    See Project
  • Our Free Plans just got better! | Auth0 Icon
    Our Free Plans just got better! | Auth0

    With up to 25k MAUs and unlimited Okta connections, our Free Plan lets you focus on what you do best—building great apps.

    You asked, we delivered! Auth0 is excited to expand our Free and Paid plans to include more options so you can focus on building, deploying, and scaling applications without having to worry about your security. Auth0 now, thank yourself later.
    Try free now
  • 5
    NeuralNote

    NeuralNote

    Audio Plugin for Audio to MIDI transcription using deep learning

    ...The system relies on neural network models to analyze audio signals and infer pitch, timing, and other musical attributes that can be represented as MIDI data. The resulting MIDI output can be edited, quantized, or exported to other instruments within a music production workflow.
    Downloads: 136 This Week
    Last Update:
    See Project
  • 6
    Claude Skills

    Claude Skills

    Public repository for Agent Skills

    ...Each Skill lives in its own directory with a SKILL.md file containing metadata and instructions, and can include supplemental scripts or assets that the agent uses to perform complex operations when relevant.
    Downloads: 98 This Week
    Last Update:
    See Project
  • 7
    ProtoMotions

    ProtoMotions

    ProtoMotions is a GPU-accelerated simulation and learning framework

    ...It provides modular environments and learning components for animation, robotics, and reinforcement learning research. Large motion datasets such as AMASS and BONES can be divided across multiple GPUs for scalable policy training. A built-in PyRoki workflow retargets human motion data to different robot bodies with a single command. Policies can be tested across Isaac Gym, Isaac Lab, Newton, Genesis, MuJoCo, and high-fidelity Isaac Sim environments. Its deployment pipeline exports ONNX policies with observation processing included, simplifying transfers from simulation to real Unitree G1 hardware. ...
    Downloads: 129 This Week
    Last Update:
    See Project
  • 8
    LLPlayer

    LLPlayer

    The media player for language learning, with dual subtitles

    ...Unlike traditional media players, the application focuses on advanced subtitle-related features that help learners understand and interact with foreign language media more effectively. The player supports dual subtitles so users can simultaneously view text in both the original language and their native language while watching videos. It can also automatically generate subtitles in real time using speech-to-text systems such as Whisper, allowing subtitles to be created even when none are available. Real-time translation capabilities enable subtitles to be translated using multiple translation engines and language models. ...
    Downloads: 85 This Week
    Last Update:
    See Project
  • 9
    MiroFish

    MiroFish

    A Simple and Universal Swarm Intelligence Engine

    ...The system extracts “seed” information from sources such as breaking news, policy documents, and market signals to construct a high-fidelity digital parallel world populated by thousands of virtual agents with independent memory and behavior rules. Users can inject variables or conditions into this simulated environment from a “god’s eye view,” enabling iterative prediction of future trends under different assumptions, which can be useful for decision support, scenario planning, or creative exploration. The engine includes both backend and frontend components, with configuration and deployment instructions for local and containerized setups, and is designed to produce detailed predictive reports based on interactions and emergent patterns within the simulated world.
    Downloads: 66 This Week
    Last Update:
    See Project
  • Custom VMs From 1 to 96 vCPUs With 99.95% Uptime Icon
    Custom VMs From 1 to 96 vCPUs With 99.95% Uptime

    General-purpose, compute-optimized, or GPU/TPU-accelerated. Built to your exact specs.

    Live migration and automatic failover keep workloads online through maintenance. One free e2-micro VM every month.
    Try Free
  • 10
    Remotion

    Remotion

    Make videos programmatically with React

    ...The framework supports exporting to standard video formats, audio synchronization, frame callbacks, and powerful tooling for previewing and debugging, so teams can iterate quickly and reliably.
    Downloads: 32 This Week
    Last Update:
    See Project
  • 11
    WhisperLive

    WhisperLive

    A nearly-live implementation of OpenAI's Whisper

    ...On the client side, you can set the language, whether to translate into English, model size, voice activity detection, and output recording behavior.
    Downloads: 39 This Week
    Last Update:
    See Project
  • 12
    Llama Coder

    Llama Coder

    Open source Claude Artifacts – built with Llama 3.1 405B

    ...Developers can fork the repo to plug in different models, change the UI, or integrate it into their own IDE-adjacent workflows.
    Downloads: 38 This Week
    Last Update:
    See Project
  • 13
    edge-tts

    edge-tts

    Use Microsoft Edge's online text-to-speech service from Python

    ...From the CLI you can adjust parameters such as speaking rate, volume, and pitch, giving you some control over prosody without diving into SSML. The library is asynchronous under the hood, which makes it efficient for batch jobs or web services that need to synthesize many utterances concurrently.
    Downloads: 41 This Week
    Last Update:
    See Project
  • 14
    Open-LLM-VTuber

    Open-LLM-VTuber

    Open source AI VTuber platform with voice chat and Live2D avatars

    Open-LLM-VTuber is an open source platform designed to create AI-powered VTuber characters that can interact with users through voice and animated avatars. It enables hands-free conversations with large language models by combining speech recognition, language processing, and text-to-speech synthesis into a single system. Users can speak directly to the AI character, and the system can respond with a generated voice while animating a Live2D avatar to simulate a talking virtual personality. ...
    Downloads: 18 This Week
    Last Update:
    See Project
  • 15
    MiMo Code

    MiMo Code

    Where Models and Agents Co-Evolve

    MiMo-Code is a terminal-native AI coding assistant built for real software projects. It can read and edit code, run commands, manage Git, and preserve project knowledge across sessions. The tool includes multiple agent modes, including a build mode for development, a plan mode for read-only analysis, and a compose mode for structured workflows. Its persistent memory system stores project notes, checkpoints, scratch notes, and task progress so the assistant can resume work with context. ...
    Downloads: 44 This Week
    Last Update:
    See Project
  • 16
    OpenWork

    OpenWork

    An open-source alternative to Claude Cowork, powered by opencode

    ...Rather than a single monolithic workflow engine, it emphasizes openness — providing APIs and interfaces so communities can build custom dashboards, integrate specialized agents, or add bespoke evaluation criteria.
    Downloads: 22 This Week
    Last Update:
    See Project
  • 17
    No AI slop

    No AI slop

    Removes 20+ patterns of AI slop from any piece of writing

    ...It targets formulaic contrasts, throat-clearing openings, false insight setups, dramatic fragments, vague attribution, inflated importance, and other recognizable habits. The skill improves clarity, grammar, sentence structure, active voice, and the use of concrete details without flattening the writer’s personal voice. It can also analyze a passage for these patterns without claiming to determine whether AI created it. When editing, it makes the minimum effective changes and returns the revised draft with a brief explanation. A built-in evaluation file provides pass-or-fail checks that the skill applies to its own work. The project can be installed in Claude Code, Codex, other AI harnesses, and packaged ChatGPT or Codex plugin environments.
    Downloads: 34 This Week
    Last Update:
    See Project
  • 18
    PenEcho

    PenEcho

    Think with AI beyond the chat box

    PenEcho is an open-source shared canvas that makes handwriting, equations, diagrams, and spatial context part of an AI conversation. Users can draw with pressure-sensitive ink, erase, pan, zoom, type formatted text, and arrange ideas across a large virtual workspace. The application sends focused visual regions to an API, Codex CLI, or Claude CLI and converts model responses into editable canvas commands. AI output can include text, formulas, plots, mixed drawings, erasures, and declarative animations. ...
    Downloads: 36 This Week
    Last Update:
    See Project
  • 19
    Whisper

    Whisper

    Robust Speech Recognition via Large-Scale Weak Supervision

    OpenAI Whisper is a general-purpose speech recognition model. It is trained on a large dataset of diverse audio and is also a multitasking model that can perform multilingual speech recognition, speech translation, and language identification. A Transformer sequence-to-sequence model is trained on various speech processing tasks, including multilingual speech recognition, speech translation, spoken language identification, and voice activity detection. These tasks are jointly represented as a sequence of tokens to be predicted by the decoder, allowing a single model to replace many stages of a traditional speech-processing pipeline. ...
    Downloads: 78 This Week
    Last Update:
    See Project
  • 20
    AutoSubs

    AutoSubs

    Instantly generate AI-powered subtitles on your device

    ...It supports both standalone usage and integration with professional video editing software such as DaVinci Resolve, allowing creators to generate and edit subtitles within their existing workflows. The tool leverages speech-to-text models, including OpenAI Whisper, to produce high-quality transcriptions and can differentiate between speakers using diarization techniques. Users can customize subtitle styling, adjust timing, and export results in multiple formats, making it suitable for content creators, filmmakers, and editors. AutoSubs is designed with performance in mind, offering efficient processing through a Rust-based backend and supporting multiple operating systems including Windows, macOS, and Linux.
    Downloads: 31 This Week
    Last Update:
    See Project
  • 21
    img2threejs

    img2threejs

    Rebuild the object in a reference image as a code-only, procedural

    ...The system can classify objects, characters, and hybrid subjects, with specialized reconstruction paths for each. It works with Claude Code, Codex, or OpenCode and can refine its specification or code when a review fails.
    Downloads: 15 This Week
    Last Update:
    See Project
  • 22
    OpenAI.fm

    OpenAI.fm

    Code for openai.fm, a demo for the OpenAI Speech API

    OpenAI.fm is an official interactive demo application built to showcase the OpenAI Speech API and its advanced text-to-speech capabilities, providing developers and creators with a hands-on web interface to convert text into high-quality, customizable audio using state-of-the-art TTS models. Developed using Next.js and the OpenAI Speech API, this demo illustrates how the latest neural voice models can produce natural, expressive speech with adjustable styles and voices, highlighting features like emotional range, tone, and real-time playback. Users can experiment with different input text and voice options directly in their browser, gaining a sense of how high-fidelity AI audio can be integrated into applications ranging from podcasts and narration to accessibility tools and interactive agents. ...
    Downloads: 15 This Week
    Last Update:
    See Project
  • 23
    AirLLM

    AirLLM

    AirLLM 70B inference with single 4GB GPU

    ...This layer-wise inference approach allows models with tens of billions of parameters to run on devices with only a few gigabytes of VRAM. AirLLM preprocesses model weights so that each transformer layer can be loaded independently during computation, reducing the memory footprint while still performing full inference. As a result, developers can experiment with models that previously required specialized high-end GPUs.
    Downloads: 29 This Week
    Last Update:
    See Project
  • 24
    GPT4All

    GPT4All

    Run Local LLMs on Any Device. Open-source

    GPT4All is an open-source project that allows users to run large language models (LLMs) locally on their desktops or laptops, eliminating the need for API calls or GPUs. The software provides a simple, user-friendly application that can be downloaded and run on various platforms, including Windows, macOS, and Ubuntu, without requiring specialized hardware. It integrates with the llama.cpp implementation and supports multiple LLMs, allowing users to interact with AI models privately. This project also supports Python integrations for easy automation and customization. ...
    Downloads: 94 This Week
    Last Update:
    See Project
  • 25
    FLUX.1

    FLUX.1

    Official inference repo for FLUX.1 models

    ...This repo focuses on running the open-source model variants efficiently, providing scripts, model loading logic, and examples for local installations, and supports integration with Python toolchains like PyTorch and popular generative pipelines. Users can launch CLI tools to generate images, experiment with different FLUX variants, and extend the base code for research-oriented applications.
    Downloads: 89 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • 2
  • 3
  • 4
  • 5
  • Next