Showing 2064 open source projects for "system linux"

View related business solutions
  • Host LLMs in Production With On-Demand GPUs Icon
    Host LLMs in Production With On-Demand GPUs

    NVIDIA L4 GPUs. 5-second cold starts. Scale to zero when idle.

    Deploy your model, get an endpoint, pay only for compute time. No GPU provisioning or infrastructure management required.
    Start Free
  • Fully Managed MySQL, PostgreSQL, and SQL Server Icon
    Fully Managed MySQL, PostgreSQL, and SQL Server

    Automatic backups, patching, replication, and failover. Focus on your app, not your database.

    Cloud SQL handles your database ops end to end, so you can focus on your app.
    Start Free
  • 1
    Mistral Medium 3.5

    Mistral Medium 3.5

    Dense 128B multimodal model for reasoning, coding, vision, and agents

    Mistral Medium 3.5 128B is Mistral AI’s first flagship merged model, unifying instruction following, reasoning, coding, vision, and agentic capabilities within a single set of weights. It uses a dense 128B-parameter architecture and supports a 256K-token context window for large documents, codebases, and extended workflows. The model accepts both text and images while generating text, using a vision encoder trained from scratch to accommodate variable image sizes and aspect ratios. Reasoning...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 2
    Rampart

    Rampart

    Lightweight on-device model for private AI text redaction

    Rampart is a lightweight, on-device privacy protection model developed by the National Design Studio to detect and redact personally identifiable information (PII) before text leaves a user's device. Rather than relying on server-side filtering, Rampart performs token-level PII detection locally, enabling privacy-preserving AI interactions with minimal latency and without exposing sensitive information to external services. The released model is a 14.7 MB ONNX artifact based on a fine-tuned...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 3
    Gemopus

    Gemopus

    Stable fine-tuned Gemma model for structured, clear responses

    Gemopus is a supervised fine-tuned version of the Gemma 4 26B instruction model, designed with a “stability first” philosophy that prioritizes reliable reasoning structure over aggressive chain-of-thought imitation. Instead of relying on distilled reasoning traces from external models, it focuses on preserving Gemma’s native reasoning style while improving answer clarity, structure, and consistency. The model enhances response organization through better use of formatting, improves...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 4
    Llama-3.2-1B

    Llama-3.2-1B

    Llama 3.2–1B: Multilingual, instruction-tuned model for mobile AI

    meta-llama/Llama-3.2-1B is a lightweight, instruction-tuned generative language model developed by Meta, optimized for multilingual dialogue, summarization, and retrieval tasks. With 1.23 billion parameters, it offers strong performance in constrained environments like mobile devices, without sacrificing versatility or multilingual support. It is part of the Llama 3.2 family, trained on up to 9 trillion tokens and aligned using supervised fine-tuning, preference optimization, and safety...
    Downloads: 0 This Week
    Last Update:
    See Project
  • Build Agents and Models on One Platform Icon
    Build Agents and Models on One Platform

    Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

    Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
    Start Free
  • 5
    Qwen2.5-14B-Instruct

    Qwen2.5-14B-Instruct

    Powerful 14B LLM with strong instruction and long-text handling

    Qwen2.5-14B-Instruct is a powerful instruction-tuned language model developed by the Qwen team, based on the Qwen2.5 architecture. It features 14.7 billion parameters and is optimized for tasks like dialogue, long-form generation, and structured output. The model supports context lengths up to 128K tokens and can generate up to 8K tokens, making it suitable for long-context applications. It demonstrates improved performance in coding, mathematics, and multilingual understanding across over...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 6
    LocalLLM Studio

    LocalLLM Studio

    Chat with local GGUF LLMs on your own machine

    Chat with local GGUF LLMs on your own machine — private, offline, model-agnostic. LLM Apps & Agents Apache-2.0 v1.0.5 100% AI-built 0 0 Get LocalLLM Studio View source Website Run open-weight language models (GGUF) fully offline on your own hardware via llama.cpp: a chat interface with system prompts, conversation history, adjustable sampling, and simple retrieval over your own files. You supply the model file — nothing is sent to any server. At a glance Downloads 0 Views 5 Rating No...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 7
    Gemma 4 12B

    Gemma 4 12B

    Unified multimodal Gemma model for local coding and reasoning

    Gemma 4 12B is Google DeepMind’s unified open-weight multimodal model designed for efficient local reasoning, coding, and multimodal understanding. Unlike other Gemma 4 models that rely on separate encoders, the 12B Unified model uses an encoder-free architecture that projects raw image patches and audio waveforms directly into the language model’s embedding space, reducing multimodal latency and simplifying fine-tuning. It supports text, image, audio, and video inputs with text output,...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8
    MiMo-V2.5-Pro

    MiMo-V2.5-Pro

    Flagship MoE model for long-context agents and complex coding

    MiMo-V2.5-Pro is Xiaomi’s flagship Mixture-of-Experts (MoE) model built for the most demanding agentic, software engineering, and long-horizon reasoning tasks. It features approximately 1.02 trillion total parameters with 42B activated per inference, balancing extreme capability with efficient execution. The model supports a 1 million token context window, enabling it to maintain coherence across long workflows involving thousands of tool calls and multi-step reasoning chains....
    Downloads: 0 This Week
    Last Update:
    See Project
  • 9
    MiMo-V2.5

    MiMo-V2.5

    Omnimodal AI model for agents, coding, and long-context tasks

    MiMo-V2.5 is a native omnimodal large language model developed by Xiaomi, designed for advanced agentic workflows, multimodal reasoning, and long-context processing. Built on a Mixture-of-Experts architecture with approximately 309B total parameters and around 15B activated per inference, it balances high capability with efficient execution. The model natively processes text, images, video, and audio within a unified system, enabling cross-modal understanding and complex task execution in a...
    Downloads: 0 This Week
    Last Update:
    See Project
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • 10
    Smaug Flash

    Smaug Flash

    304B MoE model optimized for agentic coding and long-running workflows

    Smaug Flash is an open-weight agentic coding model from Abacus.AI, fine-tuned from DeepSeek-V4-Flash-0731 to improve autonomous software engineering, tool use, and long-running agent workflows. It uses a 304B-parameter Mixture-of-Experts architecture with 43 layers, 256 routed experts, six selected experts per token, and one shared expert. Its attention system combines Multi-Head Latent Attention with a sparse token indexer, while DSpark multi-token prediction provides speculative decoding...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 11
    Qwable-v1

    Qwable-v1

    Agentic coding model combining Opus reasoning and Fable tools

    Qwable-v1 is an open-weight agentic coding model created through a chained distillation process based on Qwen3.6-35B-A3B. The model combines two distinct training stages: first, it was fine-tuned on reasoning traces derived from Claude Opus 4.7 to improve structured reasoning, and then further trained on Claude Fable-5 agentic tool-use traces to develop autonomous coding and tool-calling behavior. The result is a 35B Mixture-of-Experts model with only 3B active parameters that can switch...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 12
    Laguna M.1

    Laguna M.1

    Flagship Poolside model for agentic coding and software engineering

    Laguna M.1 is Poolside’s flagship Mixture-of-Experts model built specifically for agentic coding, software engineering, and long-horizon autonomous workflows. It contains approximately 225.8B total parameters with 23.4B activated per token, making it substantially larger and more capable than Laguna XS.2 while maintaining efficient inference through sparse activation. Trained from scratch on roughly 30 trillion tokens using Poolside’s in-house “Model Factory” pipeline, the model focuses on...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 13
    Mistral Large 3 675B Instruct 2512 NVFP4

    Mistral Large 3 675B Instruct 2512 NVFP4

    Quantized 675B multimodal instruct model optimized for NVFP4

    Mistral Large 3 675B Instruct 2512 NVFP4 is a frontier-scale multimodal Mixture-of-Experts model featuring 675B total parameters and 41B active parameters, trained from scratch on 3,000 H200 GPUs. This NVFP4 checkpoint is a post-training-activation quantized version of the original instruct model, created through a collaboration between Mistral AI, vLLM, and Red Hat using llm-compressor. It retains the same instruction-tuned behavior as the FP8 model, making it ideal for production...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 14
    Reputation-Sentinel-OS-v3

    Reputation-Sentinel-OS-v3

    Autonomous n8n framework for real-time brand protection & AI analysis.

    Reputation Sentinel OS v3.0: Autonomous Brand Defense Infrastructure Sentinel OS v3.0 is a sovereign AI framework for brand protection and automated response. Built on n8n with a massive 1,500+ logic nodes architecture, it offers depth and privacy far beyond traditional SaaS. Total Data Sovereignty: Unlike expensive monthly platforms, Sentinel OS runs on your own infrastructure. You own your data and your brand's defense. Key Capabilities: Omni-Channel: 24/7 autonomous tracking...
    Downloads: 0 This Week
    Last Update:
    See Project