Search Results for "image processing framework" - Page 52

Showing 1300 open source projects for "image processing framework"

View related business solutions
  • Build Agents and Models on One Platform Icon
    Build Agents and Models on One Platform

    Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

    Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
    Try It Free
  • Cut Data Warehouse Costs by 54% Icon
    Cut Data Warehouse Costs by 54%

    Easily migrate from Snowflake, Redshift, or Databricks with free tools.

    BigQuery delivers 54% lower TCO with exabyte scale and flexible pricing. Free migration tools handle the SQL translation automatically.
    Try Free
  • 1
    A numerical library designed for n-dimensional image-processing written in Java.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 2
    A set of Matlab toolboxes to perform signal and image processing.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 3
    Various projects from Digital Image Processing and Computer Graphics club of Warsaw University of Technology.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 4
    The Data Fusion Peer is a multitier computer vision internet application. The system provides image processing, motion tracking, and visualization information. Application will convert data into 3-Deminsional and other digital environments.
    Downloads: 0 This Week
    Last Update:
    See Project
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • 5

    geoscipy

    Python-accessible toolkit for Geoscience & Remote Sensing Applications

    ...There is also a Python API that provides a procedural as well as an object-oriented interface to these functions. A Python GUI for interacting with such datasets is also part of the project. While there are other open-source GIS, and image-processing packages available, this one is designed to be comprehensive, work on 3 major platforms, user-extensible, fast, and able to handle huge datasets. Click the Blog tab for more info.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 6
    A javascript templating toolkit that allows on-demand asynchronous dependencies loading (= html/css/js/other files ) for a given screen element, and automates template processing and display. It acts like a multitasking browser environment.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 7
    CLIP-ViT-bigG-14-laion2B-39B-b160k

    CLIP-ViT-bigG-14-laion2B-39B-b160k

    CLIP ViT-bigG/14: Zero-shot image-text model trained on LAION-2B

    CLIP-ViT-bigG-14-laion2B-39B-b160k is a powerful vision-language model trained on the English subset of the LAION-5B dataset using the OpenCLIP framework. Developed by LAION and trained by Mitchell Wortsman on Stability AI’s compute infrastructure, it pairs a ViT-bigG/14 vision transformer with a text encoder to perform contrastive learning on image-text pairs. This model excels at zero-shot image classification, image-to-text and text-to-image retrieval, and can be adapted for tasks such as image captioning or generation guidance. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8
    LSM4J is a lightweight Java based framework to model a state machine. The developer implements a few generic interfaces and let the framework take care about processing the graph.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 9
    SeeHawk is a cross-platform, cross-device video/image capturing, viewing, and editing application.
    Downloads: 0 This Week
    Last Update:
    See Project
  • Custom VMs From 1 to 96 vCPUs With 99.95% Uptime Icon
    Custom VMs From 1 to 96 vCPUs With 99.95% Uptime

    General-purpose, compute-optimized, or GPU/TPU-accelerated. Built to your exact specs.

    Live migration and automatic failover keep workloads online through maintenance. One free e2-micro VM every month.
    Try Free
  • 10

    Anapgen

    Another Application Generator

    Anapgen is a java based framework meant to generate java applications from a custom language. Anapgen is primarily designed to generate data processing applications (generating both the model and the view logic).
    Downloads: 0 This Week
    Last Update:
    See Project
  • 11
    RapidWS is a web service framework, targeted at faster processing of SOAP messages.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 12
    Nex-N2-mini

    Nex-N2-mini

    Compact agentic model for coding, tools, and productivity tasks

    Nex-N2-mini is an open-source agentic model from Nex AGI designed for real-world productivity, coding, tool use, deep research, and terminal-based execution. Built on Qwen3.5-35B-A3B-Base, it offers a lighter latency and deployment profile than Nex-N2-Pro while preserving the core Nex-N2 “Agentic Thinking” framework. This framework unifies requirement understanding, planning, code implementation, environmental feedback, debugging, evaluation, and iteration into a closed loop. It uses adaptive thinking to decide when deeper reasoning is needed and coherent thinking to keep reasoning consistent across tasks and modalities. Nex-N2-mini supports image-text-to-text workflows, explicit reasoning traces, robust function calling, and deployment through Transformers, vLLM, SGLang, Docker, and quantized local apps. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 13
    PNG.Net is a free, open source PNG library for the .NET Framework. It supports both high-level access of image pixels and low-level manipulation of PNG chunks and attributes.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 14

    KinectStreamer

    Streams data from 3D cameras over a network.

    This is an application that streams data from the Microsoft Kinect or cameras like it over a network. The program is Intended to be used in robotics applications where the controller cannot use such cameras directly due to hardware/software limitations--such as lacking usb ports or appropriate drivers--or in situations where the camera is not in close proximity to the device that needs to access it. Given that the controller can accept data from over the network, another embedded controller...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 15
    Gemma 4

    Gemma 4

    Google’s flagship dense multimodal model for coding and reasoning

    Gemma 4 is Google DeepMind’s flagship dense open-weight multimodal model, designed for high-end reasoning, coding, agentic workflows, and multimodal understanding. The model contains approximately 30.7B parameters and supports text and image inputs with text generation output, while also processing video as image-frame sequences. Built as the most capable model in the Gemma 4 family, it combines strong reasoning performance with a large 256K-token context window and configurable thinking modes. Gemma 4 31B supports native function calling, structured outputs, and more than 140 languages, making it suitable for enterprise assistants, coding agents, document analysis, and multilingual applications. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 16
    Ministral 3 3B Base 2512

    Ministral 3 3B Base 2512

    Small 3B-base multimodal model ideal for custom AI on edge hardware

    ...It supports dozens of languages, making it practical for multilingual, global, or distributed environments. With a large 256k token context window, it can handle long documents, extended inputs, or multi-step processing workflows even at its small size.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 17
    A framework for image analysis and data analysis, aimed towards microscopy and the needs of current research.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 18
    Nex-N2-Pro

    Nex-N2-Pro

    Large agentic model for coding, tools, research, and execution

    ...It supports image-text-to-text workflows, explicit reasoning traces, robust function calling, and deployment through Transformers, vLLM, SGLang, Docker, and quantized local apps. Nex-N2-Pro performs strongly across agentic, coding, search, and reasoning benchmarks, including Terminal-Bench, SWE-Bench Pro, BrowseComp, Toolathlon, WideSearch, GPQA Diamond, and GDPval.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 19
    Reputation-Sentinel-OS-v3

    Reputation-Sentinel-OS-v3

    Autonomous n8n framework for real-time brand protection & AI analysis.

    ...Deep AI Analysis: Surgical sentiment detection (sarcasm, threats, urgency) via local or cloud AI models. Incident Response: Automated rule-based workflows to mitigate damage instantly. Enterprise Power: Optimized for high-load processing without monthly fees. The ultimate self-hosted operating system for PR agencies and security-focused enterprises. Buy once, own forever.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 20
    Inkling-Small

    Inkling-Small

    Efficient multimodal MoE model for coding, tools, and reasoning

    Inkling-Small is an open-weight general-purpose multimodal model from Thinking Machines Lab, designed for agentic systems, coding assistants, chatbots, retrieval workflows, and natural-language applications. It accepts text, images, and audio as input and produces text output, with multilingual and multi-programming-language capabilities. The model uses a sparse Mixture-of-Experts architecture with 276B total parameters and 12B active per token, enabling strong performance with lower...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21
    MiMo-V2.5

    MiMo-V2.5

    Omnimodal AI model for agents, coding, and long-context tasks

    MiMo-V2.5 is a native omnimodal large language model developed by Xiaomi, designed for advanced agentic workflows, multimodal reasoning, and long-context processing. Built on a Mixture-of-Experts architecture with approximately 309B total parameters and around 15B activated per inference, it balances high capability with efficient execution. The model natively processes text, images, video, and audio within a unified system, enabling cross-modal understanding and complex task execution in a...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22
    Ministral 3 8B Instruct 2512

    Ministral 3 8B Instruct 2512

    Compact 8B multimodal instruct model optimized for edge deployment

    Ministral 3 8B Instruct 2512 is a balanced, efficient model in the Ministral 3 family, offering strong multimodal capabilities within a compact footprint. It combines an 8.4B-parameter language model with a 0.4B vision encoder, enabling both text reasoning and image understanding. This FP8 instruct-fine-tuned variant is optimized for chat, instruction following, and structured outputs, making it ideal for daily assistant tasks and lightweight agentic workflows. Designed for edge deployment,...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23
    Ministral 3 14B Instruct 2512

    Ministral 3 14B Instruct 2512

    Efficient 14B multimodal instruct model with edge deployment and FP8

    Ministral 3 14B Instruct 2512 is the largest model in the Ministral 3 family, delivering frontier performance comparable to much larger systems while remaining optimized for edge-level deployment. It combines a 13.5B-parameter language model with a 0.4B-parameter vision encoder, enabling strong multimodal understanding in both text and image tasks. This FP8 instruct-tuned variant is designed specifically for chat, instruction following, and agentic workflows with robust system-prompt...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 24
    Monk Computer Vision

    Monk Computer Vision

    A low code unified framework for computer vision and deep learning

    Monk is an open source low code programming environment to reduce the cognitive load faced by entry level programmers while catering to the needs of Expert Deep Learning engineers. There are three libraries in this opensource set. - Monk Classiciation- https://monkai.org. A Unified wrapper over major deep learning frameworks. Our core focus area is at the intersection of Computer Vision and Deep Learning algorithms. - Monk Object Detection -...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 25
    Fraunhofer SimWebApp
    Multi-user online service implementation for web-based factory modeling. Rich user interface and scalable backend system to build online simulation services on or do other processing of factory data. Demo at https://fabriksimulation.ipa.fraunhofer.de
    Downloads: 0 This Week
    Last Update:
    See Project