Showing 55 open source projects for "ai models"

View related business solutions
  • Host LLMs in Production With On-Demand GPUs Icon
    Host LLMs in Production With On-Demand GPUs

    NVIDIA L4 GPUs. 5-second cold starts. Scale to zero when idle.

    Deploy your model, get an endpoint, pay only for compute time. No GPU provisioning or infrastructure management required.
    Start Free
  • Ship Agents Faster Icon
    Ship Agents Faster

    Transform your applications and workflows into powerful agentic systems at global scale.

    Gemini Enterprise Agent Platform lets you rapidly build, scale, govern and optimize production-ready agents grounded in your organization's data. The platform enables developers to build custom or pre-built agents for virtually any use case. New customers get $300 in free credits.
    Start Free
  • 1
    Clarity AI Upscaler

    Clarity AI Upscaler

    AI Image Upscaler & Enhancer

    Clarity AI Upscaler is an open-source AI image enhancement tool designed to increase the resolution and visual quality of images using modern generative techniques. The system uses deep learning models based on diffusion and other image generation methods to reconstruct high-resolution versions of low-resolution images while preserving important visual details.
    Downloads: 15 This Week
    Last Update:
    See Project
  • 2
    OpenVINO AI Plugins for Audacity

    OpenVINO AI Plugins for Audacity

    A set of AI-enabled effects, generators, and analyzers for Audacity

    A set of AI-enabled effects, generators, and analyzers for Audacity. These AI features run 100% locally on your PC, no internet connection is necessary. OpenVINO™ is used to run AI models on supported accelerators found on the user's system such as CPU, GPU, and NPU.
    Downloads: 212 This Week
    Last Update:
    See Project
  • 3
    Open Design

    Open Design

    Local-first, open-source alternative to Anthropic's Claude Design

    Open Design is a local-first, open-source AI design platform that enables coding agents to generate complete design systems and visual artifacts from prompts. It functions as an alternative to proprietary AI design tools by allowing users to connect their own models and run everything locally or deploy it as a web application. The system includes a library of design skills and brand-grade design systems that guide the generation process, ensuring consistency and quality. ...
    Downloads: 22 This Week
    Last Update:
    See Project
  • 4
    Lama Cleaner

    Lama Cleaner

    Image inpainting tool powered by SOTA AI Model

    Image inpainting tool powered by SOTA AI Model. Remove any unwanted object, defect, or people from your pictures or erase and replace(powered by stable diffusion) anything on your pictures. Lama Cleaner is a free, open-source and fully self-hostable inpainting tool powered by state-of-the-art AI models. You can use it to remove any unwanted object, defect, or people from your pictures or erase and replace anything on your pictures.
    Downloads: 23 This Week
    Last Update:
    See Project
  • Veeam Data Platform v13.1 - Get Your Free Trial Icon
    Veeam Data Platform v13.1 - Get Your Free Trial

    Secure by design, portable by default. Recover clean, fast, anywhere. Start a free trial.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try it Free
  • 5
    Upscayl

    Upscayl

    Free and Open Source AI Image Upscaler for Linux, MacOS and Windows

    ...You can also download the flatpak version and double-click the flatpak file to install via Store but wait for the full release, we'll be pushing it to Flathub for easy access. Upscayl uses AI models to enhance your images by guessing what the details could be. It uses Real-ESRGAN (and more in the future) model to achieve this. The CLI tool is called real-esrgan-ncnn-vulkan and it's available on the Real-ESRGAN repository.
    Downloads: 140 This Week
    Last Update:
    See Project
  • 6
    IOPaint

    IOPaint

    Image inpainting tool powered by SOTA AI Model

    IOPaint is a powerful open-source image editing tool focused on inpainting, outpainting, object removal, and general image manipulation driven by state-of-the-art AI models, delivering these capabilities through both local and hosted workflows. Designed to be fully self-hosted and flexible, IOPaint supports a variety of underlying generators and inpaint models — from LaMa erase networks to Stable Diffusion-based replace/object generation — giving users multiple ways to refine or reconstruct images by removing unwanted elements or expanding artwork beyond its original boundaries. ...
    Downloads: 29 This Week
    Last Update:
    See Project
  • 7
    Modly

    Modly

    Desktop app to generate 3D models from images using local AI

    Modly is a lightweight platform designed to simplify the creation and management of modular AI-driven workflows and tools. It focuses on breaking down complex processes into reusable modules that can be combined and orchestrated to achieve specific goals. The system emphasizes flexibility, allowing users to define custom modules and integrate them into larger pipelines. It supports experimentation, enabling users to test and refine workflows quickly. Modly is designed to be accessible, with...
    Downloads: 54 This Week
    Last Update:
    See Project
  • 8
    HyperFrames

    HyperFrames

    Write HTML. Render video. Built for agents

    ...It supports integration with AI models to generate or modify content within these frames, allowing real-time adaptation. The framework emphasizes composability, enabling developers to build complex experiences by combining smaller units. It is also designed to be extensible, allowing integration with different platforms and tools. Overall, Hyperframes provides a flexible infrastructure for building dynamic, AI-driven content systems that go beyond static outputs.
    Downloads: 30 This Week
    Last Update:
    See Project
  • 9
    OpenShorts

    OpenShorts

    Free & open source AI video platform

    ...It also supports generating marketing videos using AI actors, voiceovers, and scripted narratives without requiring cameras or production resources. The platform integrates publishing capabilities, allowing users to distribute content directly to TikTok, Instagram, and YouTube. Its architecture uses modern technologies such as FastAPI, FFmpeg, and AI models for transcription, analysis, and rendering.
    Downloads: 14 This Week
    Last Update:
    See Project
  • Fully Managed MySQL, PostgreSQL, and SQL Server Icon
    Fully Managed MySQL, PostgreSQL, and SQL Server

    Automatic backups, patching, replication, and failover. Focus on your app, not your database.

    Cloud SQL handles your database ops end to end, so you can focus on your app.
    Start Free
  • 10
    TRELLIS.2

    TRELLIS.2

    Native and Compact Structured Latents for 3D Generation

    TRELLIS.2 is a cutting-edge open-source model and codebase for high-fidelity 3D asset generation from 2D images, developed to push forward the state of the art in image-to-3D generation. At its core is a novel sparse voxel structure called O-Voxel that jointly encodes both geometry and surface appearance, enabling reconstruction and generation of complex 3D shapes with arbitrary topology, open surfaces, and physically based rendering (PBR) textures. The system leverages a large...
    Downloads: 30 This Week
    Last Update:
    See Project
  • 11
    Palmier Pro

    Palmier Pro

    macOS video editor built for AI

    Palmier Pro is an open-source video editor for Mac built around AI-assisted video creation. It lets users and coding agents work together directly inside a timeline, blending traditional editing with generative workflows. The app is written from scratch in Swift and takes inspiration from professional editors like Premiere Pro while rethinking the workflow around AI. Users can generate videos and images inside the editor with models such as Seedance, Kling, and Nano Banana Pro. ...
    Downloads: 13 This Week
    Last Update:
    See Project
  • 12
    nunif

    nunif

    Misc; latest version of waifu2x; 2D video to stereo 3D video

    nunif is a deep learning–based image processing framework focused on image upscaling, restoration, denoising, and enhancement tasks using neural network models. The project provides a collection of AI-powered utilities designed primarily for anime-style artwork, illustrations, and high-quality image restoration workflows. It includes command-line tools and graphical interfaces for applying trained neural models to improve image resolution and visual clarity while minimizing artifacts. nunif supports GPU acceleration and batch processing, making it suitable for creators, archivists, and enthusiasts handling large image collections. ...
    Downloads: 4 This Week
    Last Update:
    See Project
  • 13
    OpenBrand

    OpenBrand

    Extract brand assets (logos, colors, backdrops) from any website

    OpenBrand is an open-source platform aimed at helping users generate, manage, and experiment with branding assets using modern AI-driven workflows. It focuses on simplifying the creation of brand identities by integrating tools for generating logos, visual assets, and design systems in a cohesive environment. The project is built with extensibility in mind, allowing developers to integrate additional AI models or design pipelines to expand its capabilities. It provides a structured approach to branding by combining automation with user input, enabling rapid prototyping of brand concepts without requiring deep design expertise. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 14
    VGGSfM

    VGGSfM

    VGGSfM: Visual Geometry Grounded Deep Structure From Motion

    ...With minimal configuration, users can process single scenes or full video sequences, apply motion masks to exclude moving objects, and train neural radiance or splatting models directly from reconstructed outputs.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 15
    VMZ (Video Model Zoo)

    VMZ (Video Model Zoo)

    VMZ: Model Zoo for Video Modeling

    The codebase was designed to help researchers and practitioners quickly reproduce FAIR’s results and leverage robust pre-trained backbones for downstream tasks. It also integrates Gradient Blending, an audio-visual modeling method that fuses modalities effectively (available in the Caffe2 implementation). Although VMZ is now archived and no longer actively maintained, it remains a valuable reference for understanding early large-scale video model training, transfer learning, and multimodal...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 16
    3DCellForge

    3DCellForge

    AI-powered interactive 3D cell generation and exploration studio

    3DCellForge is an AI-powered 3D cell generation and exploration studio built as a polished browser prototype. It uses React, Vite, Three.js, React Three Fiber, Drei, and Framer Motion to create an interactive WebGL environment for exploring biological cell models. Users can rotate, zoom, inspect organelles, compare views, take notes, capture screenshots, and export or import GLB models.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 17
    Bloom

    Bloom

    An open source, agentic Loom alternative

    ...Bloom supports the concept of “agentic video,” where recordings are not just passive media files but interactive assets that can be searched, summarized, and transformed into workflows or insights. This makes it particularly useful for use cases such as debugging AI agents, documenting processes, training models, or creating automated video-based reports.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 18
    Transcoder

    Transcoder

    Hardware-accelerated video transcoding using Android MediaCodec APIs

    Transcoder by DeepMedia is an AI-powered video-to-video speech translation engine that enables fully automated multilingual dubbing. Unlike traditional speech translation systems that rely on multi-stage pipelines, Transcoder directly translates one speaker’s video into another language while preserving facial expressions, lip-sync, and vocal identity. Designed for real-time use and production-grade pipelines, Transcoder combines advanced deep learning models with GPU acceleration to deliver high-quality translations across languages. ...
    Downloads: 4 This Week
    Last Update:
    See Project
  • 19
    LingBot-Map

    LingBot-Map

    A feed-forward 3D foundation model for reconstructing scenes

    ...It can be particularly useful in complex chatbot systems where multiple branches and conditions need to be managed effectively. The project supports extensibility, allowing developers to adapt the mapping system to different chatbot frameworks or AI models. Its design encourages clarity and transparency in conversational design, reducing ambiguity in dialogue flows. Overall, lingbot-map serves as a tool for improving the structure and reliability of conversational AI systems.
    Downloads: 3 This Week
    Last Update:
    See Project
  • 20
    Perfect Pixel

    Perfect Pixel

    Refine and quantize messy AI pixel art into clean, perfect pixels

    perfectPixel is a workflow tool for turning messy “pixel-style” images, especially those produced by generative models, into truly grid-aligned pixel art that reads cleanly at any scale. It tackles a common problem with AI pixel art: edges that look pixelated at first glance but are not actually aligned to a coherent pixel grid, which causes shimmer, blur, and uneven block sizes when you zoom in. The tool analyzes an image to infer the intended grid size, then refines and quantizes the artwork so pixels snap into consistent cells and the final result looks crisp and intentional. ...
    Downloads: 2 This Week
    Last Update:
    See Project
  • 21
    Viral-Clips-Crew

    Viral-Clips-Crew

    Your CrewAI Powered Video Editing Assistant

    Viral-Clips-Crew is an AI-driven video processing pipeline designed to generate short-form, engaging clips from long-form video content automatically. It analyzes transcripts and video data to identify the most engaging or “viral” moments, reducing the need for manual editing. The system integrates tools like FFmpeg and AI models to handle segmentation, cropping, and formatting for vertical video platforms.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22
    Mesh R-CNN

    Mesh R-CNN

    code for Mesh R-CNN, ICCV 2019

    Mesh R-CNN is a 3D reconstruction and object understanding framework developed by Facebook Research that extends Mask R-CNN into the 3D domain. Built on top of Detectron2 and PyTorch3D, Mesh R-CNN enables end-to-end 3D mesh prediction directly from single RGB images. The model learns to detect, segment, and reconstruct detailed 3D mesh representations of objects in natural images, bridging the gap between 2D perception and 3D understanding. Unlike voxel-based or point-based approaches, Mesh...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23
    Kornia

    Kornia

    Open Source Differentiable Computer Vision Library

    ...At its core, the package uses PyTorch as its main backend both for efficiency and to take advantage of the reverse-mode auto-differentiation to define and compute the gradient of complex functions. Inspired by existing packages, this library is composed by a subset of packages containing operators that can be inserted within neural networks to train models to perform image transformations, epipolar geometry, depth estimation, and low-level image processing such as filtering and edge detection that operate directly on tensors. With Kornia we fill the gap between classical and deep computer vision that implements standard and advanced vision algorithms for AI. Our libraries and initiatives are always according to the community needs.
    Downloads: 8 This Week
    Last Update:
    See Project
  • 24
    pyTheory

    pyTheory

    Music Theory for Humans

    ...A browser playground provides access without installation, while optional integrations support Ableton Link and external MIDI input. The project also offers Claude Code skills for AI-assisted composition and analysis.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 25
    AutoCrop-Vertical

    AutoCrop-Vertical

    Smart video converter using YOLOv8 and FFmpeg

    AutoCrop-Vertical is a Python-based video processing tool that automatically converts horizontal videos into vertical formats optimized for social media platforms. It uses computer vision techniques and AI models such as YOLOv8 to analyze each frame, detect subjects, and dynamically adjust cropping decisions. Instead of applying a static center crop, the system intelligently tracks people or key objects to preserve visual focus and composition. When cropping would degrade the scene, it can switch to alternative layouts such as letterboxing to maintain context. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • 2
  • 3
  • Next