Showing 38 open source projects for "unit-api"

View related business solutions
  • $300 Free Credits to Build on Google Cloud Icon
    $300 Free Credits to Build on Google Cloud

    New customers can spin up VMs, build with AI, and query data at no cost.

    Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
    Start Free
  • Veeam Data Platform v13.1 - Get Your Free Trial Icon
    Veeam Data Platform v13.1 - Get Your Free Trial

    Secure by design, portable by default. Recover clean, fast, anywhere. Start a free trial.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try it Free
  • 1
    Qdrant

    Qdrant

    Vector Database for the next generation of AI applications

    Qdrant is a vector similarity engine & vector database. It deploys as an API service providing search for the nearest high-dimensional vectors. With Qdrant, embeddings or neural network encoders can be turned into full-fledged applications for matching, searching, recommending, and much more! Provides the OpenAPI v3 specification to generate a client library in almost any programming language. Alternatively, utilize ready-made client for Python or other programming languages with additional functionality. ...
    Downloads: 34 This Week
    Last Update:
    See Project
  • 2
    Codex-X

    Codex-X

    Codex Switch & Instruct desktop manager

    ...It replaces repeated manual file editing with a visual interface for prompts, providers, sessions, skills, MCP servers, and TOML settings. Users can categorize, import, edit, enable, disable, cache, and synchronize Markdown instruction templates. Provider tools store multiple API configurations, test connections, retrieve models, and switch between official and third-party services. Session management can search, group, inspect, synchronize, and permanently delete local Codex histories. The application also displays authentication and configuration files while creating backups before important changes. Packages are offered for macOS, Windows, and Linux through a Tauri-based desktop application.
    Downloads: 15 This Week
    Last Update:
    See Project
  • 3
    MagicAPI AI Gateway

    MagicAPI AI Gateway

    Built for demanding AI workflows

    The world's fastest AI Gateway proxy, written in Rust and optimized for maximum performance. This high-performance API gateway routes requests to various AI providers (OpenAI, GROQ) with streaming support, making it perfect for developers who need reliable and blazing-fast AI API access.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 4
    DeepReasoning

    DeepReasoning

    High-performance API combining reasoning and creative AI models

    DeepReasoning is a high-performance large language model inference API designed to unify advanced reasoning and creative generation capabilities into a single system. It combines DeepSeek R1’s chain-of-thought reasoning with Claude’s strengths in code generation and conversational output, enabling more capable and balanced responses. DeepReasoning provides both an API and a chat interface, allowing developers and users to interact with the combined models in a streamlined way. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • Host LLMs in Production With On-Demand GPUs Icon
    Host LLMs in Production With On-Demand GPUs

    NVIDIA L4 GPUs. 5-second cold starts. Scale to zero when idle.

    Deploy your model, get an endpoint, pay only for compute time. No GPU provisioning or infrastructure management required.
    Try Free
  • 5
    shimmy

    shimmy

    Python-free Rust inference server

    The shimmy project is a lightweight local inference server designed to run large language models with minimal overhead. Written primarily in Rust, the tool provides a small standalone binary that exposes an API compatible with the OpenAI interface, allowing existing applications to interact with local models without significant code changes. This compatibility enables developers to replace remote AI services with locally hosted models while keeping their existing software architecture intact. Shimmy focuses on performance and simplicity, using efficient runtime components to minimize memory usage and startup time compared to heavier inference frameworks. ...
    Downloads: 9 This Week
    Last Update:
    See Project
  • 6
    Win-CodexBar

    Win-CodexBar

    Show usage stats for OpenAI Codex and Claude Code

    ...It is designed as a lightweight desktop utility that aggregates usage data from various providers and displays it in a compact, always-accessible interface. The app supports dozens of AI services, allowing developers to monitor API usage, quotas, and costs without logging into each platform individually. It includes a dynamic tray icon that visually represents usage levels, along with a detailed panel for deeper insights. The system supports importing credentials and browser cookies to access provider data securely. It also includes a command-line interface for scripting and automation, making it useful in development and CI environments. ...
    Downloads: 30 This Week
    Last Update:
    See Project
  • 7
    herdr

    herdr

    Agent multiplexer that lives in your terminal

    ...Sessions continue running after detach, which allows users to reconnect later from another terminal or over SSH. The project is distributed as a lightweight Rust binary for Linux, macOS, and beta Windows support. It also includes a local socket API and CLI so agents can create panes, read output, and orchestrate terminal workflows themselves.
    Downloads: 5 This Week
    Last Update:
    See Project
  • 8
    aaif-goose

    aaif-goose

    An open source, extensible AI agent that goes beyond code suggestions

    ...It is built for more than coding, supporting research, writing, automation, data analysis, workflows, and general task execution. The project provides a desktop app, command-line interface, and API, which gives users multiple ways to work with the agent. goose is designed to connect with tools and local context so it can help complete practical tasks instead of only answering chat prompts. It is useful for developers, analysts, writers, and technical users who want an extensible assistant that can operate inside their environment. ...
    Downloads: 3 This Week
    Last Update:
    See Project
  • 9
    LangDB AI Gateway

    LangDB AI Gateway

    Govern, secure, and optimize your AI traffic

    AI Gateway is a high-performance, open-source API gateway optimized for managing and monitoring LLM traffic at scale. Developed by the LangDB team, AI Gateway acts as an intermediary between clients and backend LLMs, providing advanced features like caching, rate limiting, prompt management, and observability. It helps teams secure and optimize their LLM deployments, whether using local models or external APIs like OpenAI or Anthropic.
    Downloads: 2 This Week
    Last Update:
    See Project
  • Build Agents and Models on One Platform Icon
    Build Agents and Models on One Platform

    Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

    Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
    Try It Free
  • 10
    Google Workspace CLI

    Google Workspace CLI

    Command-line tool for Drive, Gmail, Calendar, Sheets, Docs, Chat, etc.

    Google Workspace CLI (gws) is a command-line tool designed to interact with Google Workspace services such as Drive, Gmail, Calendar, Sheets, and more from a single interface. It dynamically generates its command structure using Google’s Discovery Service, allowing it to automatically support new API endpoints as they become available. The tool eliminates the need for manual REST API calls by providing structured commands and built-in help for each resource and method. It outputs structured JSON responses, making it easy for developers, scripts, and AI agents to process results programmatically. The CLI supports multiple authentication methods, including OAuth login, service accounts, and environment-based credentials for automated environments. ...
    Downloads: 2 This Week
    Last Update:
    See Project
  • 11
    Text Embeddings Inference

    Text Embeddings Inference

    High-performance inference server for text embeddings models API layer

    ...It is built to support transformer-based embedding models, making it suitable for tasks such as semantic search, clustering, and retrieval-augmented systems. It provides an API interface that allows developers to integrate embedding capabilities into applications without managing model internals directly. Text Embeddings Inference is optimized for throughput and low latency, enabling it to handle large volumes of requests reliably. It also emphasizes ease of deployment, often using containerization and configurable runtime options to adapt to different infrastructure setups.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 12
    AOS Community Edition

    AOS Community Edition

    AOS Community Edition: the open agent operating system

    AOS Community Edition is an open agent operating system for building inspectable and composable agent environments. It provides the aos command-line interface, an HTTP API, community distributions, first-party capsules, provider and model controls, and Unicity Audit. The supported installer provisions a pinned runtime and a versioned collection of 21 Community Edition capsules. Local assets allow both standard and offline initialization, while coordinated upgrades preserve product compatibility. Signed release metadata, checksums, provenance attestations, and runtime compatibility files strengthen supply-chain verification. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 13
    webclaw

    webclaw

    Fast, local-first web content extraction for LLMs

    ...The tool addresses a major inefficiency in AI workflows by removing irrelevant elements like navigation menus, ads, and scripts, significantly reducing token usage when feeding data into language models. It supports multiple modes of operation, including CLI usage, REST API access, and an MCP server for direct integration with agent-based systems. Webclaw also provides advanced capabilities such as recursive crawling, structured JSON extraction, summarization, and content comparison, making it suitable for research and data pipelines. Its local-first architecture ensures privacy and eliminates the need for API keys.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 14
    uzu

    uzu

    A high-performance inference engine for AI models

    ...Written primarily in Rust and leveraging Apple’s Metal framework, the project focuses on maximizing performance when executing large language models and other AI workloads on devices such as Mac computers with M-series chips. The engine implements a hybrid architecture in which model layers can be executed either as custom GPU kernels or through Apple’s MPSGraph API, allowing it to balance performance and compatibility depending on the workload. By utilizing Apple’s unified memory architecture, uzu reduces memory copying overhead and improves inference throughput for local AI workloads. The system includes a simple high-level API that enables developers to run models, create inference sessions, and generate outputs with minimal configuration.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 15
    AIChat

    AIChat

    All-in-one LLM CLI tool featuring Shell Assistant

    AIChat is a lightweight terminal-based chatbot powered by GPT models, enabling AI-driven conversations directly from the command line.
    Downloads: 4 This Week
    Last Update:
    See Project
  • 16
    ort

    ort

    Fast ML inference & training for ONNX models in Rust

    ort is a high-performance Rust library that provides bindings to ONNX Runtime, enabling developers to run machine learning inference and training workflows directly within Rust applications using the standardized ONNX model format. It is designed to bridge the gap between modern machine learning frameworks and systems programming by offering a safe, ergonomic API for executing models originally built in ecosystems like PyTorch, TensorFlow, or scikit-learn. The library emphasizes speed and efficiency, leveraging hardware acceleration across CPUs, GPUs, and specialized accelerators to deliver low-latency inference both on-device and in server environments. One of its key strengths is its flexibility, as it supports multiple backends and allows developers to configure execution providers depending on available hardware. ort also includes advanced capabilities such as model compilation and optimization, reducing startup time and improving runtime performance in production systems.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 17
    Daft

    Daft

    Distributed DataFrame for Python designed for the cloud

    ...Underneath its Python API, Daft is built in blazing fast Rust code. Rust powers Daft’s vectorized execution and async I/O, allowing Daft to outperform frameworks such as Spark.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 18
    Moltis

    Moltis

    A Rust-native claw you can trust

    ...The platform also includes long-term memory powered by hybrid vector and full-text search, allowing the assistant to retain context across sessions. With multi-channel access such as web UI, Telegram, and API endpoints, Moltis functions as a unified automation hub intended for developers and advanced users who want full control.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 19
    OpenAI Harmony

    OpenAI Harmony

    Renderer for the harmony response format to be used with gpt-oss

    ...It defines a structured way for language models to produce outputs, including regular text, reasoning traces, tool calls, and structured data. By mimicking the OpenAI Responses API, Harmony provides developers with a familiar interface while enabling more advanced capabilities such as multiple output channels, instruction hierarchies, and tool namespaces. The format is essential for ensuring gpt-oss models operate correctly, as they are trained to rely on this structure for generating and organizing their responses. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 20
    AgentENV

    AgentENV

    Distributed platform for running agent environments at scale

    ...The platform includes an aenv command-line interface for creating templates, starting shells, executing commands, and managing sandbox lifecycles. Its E2B-compatible HTTP API lets existing Python and TypeScript integrations target a self-hosted deployment with minimal changes. AgentENV is designed for high-density agentic reinforcement learning and currently requires trusted-network deployment because built-in authorization is not yet available.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21
    Cog

    Cog

    Package and deploy machine learning models using Docker containers

    ...Cog also resolves compatibility issues between frameworks and GPU libraries by automatically selecting compatible combinations of CUDA, cuDNN, and machine learning frameworks such as PyTorch or TensorFlow. Cog automatically generates a RESTful HTTP API for running predictions, enabling models to be accessed programmatically through a built-in prediction server.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22
    BaseRT

    BaseRT

    Fastest LLM inference runtime for Apple Silicon

    BaseRT is a local large language model inference runtime optimized for Apple Silicon computers. It accelerates model execution through hand-written Metal kernels and requires an M1 or newer Mac running macOS 14 or later. A unified command-line interface can download models from Hugging Face, convert checkpoints, launch chats, benchmark performance, and inspect model packages. Its server implements OpenAI-compatible chat, completion, embedding, transcription, tool-calling, and multimodal...
    Downloads: 4 This Week
    Last Update:
    See Project
  • 23
    Forge Code

    Forge Code

    AI enabled pair programmer for Claude, GPT, O Series, Grok, Deepseek

    ...Rather than requiring a separate UI or web-based IDE, Forge respects the developer’s existing habits and setups, and keeps all operations local, ensuring your code doesn’t get sent to unknown external services — a strong point for privacy and security. It supports many model providers (e.g. GPT, Claude, Grok, and others) via API keys.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 24
    CocoIndex

    CocoIndex

    ETL framework to index data for AI, such as RAG

    CocoIndex is an open-source framework designed for building powerful, local-first semantic search systems. It lets users index and retrieve content based on meaning rather than keywords, making it ideal for modern AI-based search applications. CocoIndex leverages vector embeddings and integrates with various models and frameworks, including OpenAI and Hugging Face, to provide high-quality semantic understanding. It’s built for transparency, ease of use, and local control over your search...
    Downloads: 3 This Week
    Last Update:
    See Project
  • 25
    Lingua-RS

    Lingua-RS

    The most accurate natural language detection library for Rust

    Lingua-RS is a language detection library implemented in Rust, designed to accurately identify the language of given text samples. It tells you which language some text is written in. This is very useful as a preprocessing step for linguistic data in natural language processing applications such as text classification and spell checking. Other use cases, for instance, might include routing e-mails to the right geographically located customer service department, based on the e-mails' languages.
    Downloads: 2 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • 2
  • Next