Showing 440 open source projects for "benchmarks"

View related business solutions
  • $300 Free Credits to Build on Google Cloud Icon
    $300 Free Credits to Build on Google Cloud

    New customers can spin up VMs, build with AI, and query data at no cost.

    Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
    Start Free
  • Build Agents and Models on One Platform Icon
    Build Agents and Models on One Platform

    Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

    Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
    Start Free
  • 1
    AIDA64 Extreme

    AIDA64 Extreme

    AIDA64 Extreme: Ultimate PC diagnostics & system info tool

    AIDA64 Extreme - The Ultimate System Diagnostics & Benchmarking Tool Unlock the full potential of your PC with AIDA64, the industry-leading system information, diagnostics, and benchmarking software. Trusted by PC enthusiasts, IT professionals, and overclockers, AIDA64 provides detailed insights into your hardware, software, and system performance. Optimize your device, troubleshoot issues, and push performance to the max. Why Choose AIDA64? - Comprehensive System Info: Get...
    Downloads: 58 This Week
    Last Update:
    See Project
  • 2
    Awesome Network Analysis

    Awesome Network Analysis

    A curated list of awesome network analysis resources

    awesome-network-analysis is a curated list of resources focused on network and graph analysis, including libraries, frameworks, visualization tools, datasets, and academic papers. It covers multiple programming languages and domains like sociology, biology, and computer science. This repository serves as a central reference for researchers, analysts, and developers working with network data.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 3
    ARC-AGI

    ARC-AGI

    The Abstraction and Reasoning Corpus

    ...The dataset is structured as grid-based puzzles, where each task requires understanding transformations such as symmetry, counting, or spatial manipulation. Unlike traditional machine learning benchmarks, ARC emphasizes generalization and reasoning over statistical pattern recognition, making it particularly challenging for current AI systems. The repository also includes a browser-based interface that allows humans to attempt solving the tasks manually, providing a baseline for comparison.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 4
    LinuxHardware Suite

    LinuxHardware Suite

    Hardware Info for Linux portable AppImage + Benchmark

    LinuxHardware Suite (portable/AppImage) -------------------------------------- Ihr wollt wissen was in Eurem Rechner steckt? Genau das könnt Ihr mit LinuxHardware Suite, ohne jedes mal die Konsole zu bemühen und ohne irgendwelche zusätzlichen Installationen. Die Software ist komplett deutsch, eine englische Übersetzung ist noch geplant. Weiter Infos und Screenshots im Forum https://sourceforge.net/p/linuxhardware-info/discussion/lwi/thread/21ccc5b145/ License: Freeware (Closed...
    Downloads: 4 This Week
    Last Update:
    See Project
  • Veeam Data Platform v13.1 - Get Your Free Trial Icon
    Veeam Data Platform v13.1 - Get Your Free Trial

    Secure by design, portable by default. Recover clean, fast, anywhere. Start a free trial.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try it Free
  • 5
    q4rescue

    q4rescue

    A live linux Rescue toolkit/Emergency OS - based on q4os Trinity

    A live linux system rescue toolkit based on q4os Trinity available as a bootable iso for administrating, repairing and cloning/restoring your system and data. Check wiki for full description : https://sourceforge.net/p/q4rescue/wiki/ Main tools: -Foxclone -Rescuezilla -Clonezilla -DDrescue-gui -qtfsarchiver -G4L -Apart -Testdisk -Photorec -Boot Repair -WoeUSB -Q4OS imager -UNetbootin -usbimager -Kdirstats -Kdiskmark -Rclone & Rclone...
    Downloads: 85 This Week
    Last Update:
    See Project
  • 6
    B2B Price Waterfall Calculator

    B2B Price Waterfall Calculator

    Analyse pocket price, pocket margin and margin leakage in B2B deals.

    ...It calculates price realisation, front-end gross profit and margin, pocket gross profit and margin, margin leakage, leakage percentage, margin delta and an illustrative user-configured commercial review signal. The workbook does not provide industry benchmarks, recommended margins, acceptable leakage levels or approval policies. Published by Configure to WIN. Official online calculator: https://configure.win/resources/price-waterfall-calculator Official source repository: https://github.com/configure-to-win/b2b-price-waterfall-calculator
    Downloads: 1 This Week
    Last Update:
    See Project
  • 7
    Manufacturing Software Decision Tool

    Manufacturing Software Decision Tool

    Assess manufacturing estimating, product configuration + quoting needs

    ...It includes a Decision tool, detailed Requirements matrix, Process ownership model, Scenario library, Worked example and Definitions. The resulting scores measure only the intensity of the requirements entered. They are not vendor ratings, software benchmarks or implementation guarantees. Published by Configure to WIN. Official online decision tool: https://configure.win/resources/manufacturing-software-decision-tool Official source repository: https://github.com/configure-to-win/manufacturing-software-decision-tool
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8
    Quote Turnaround Time Calculator

    Quote Turnaround Time Calculator

    Measure active work, waiting, rework and approval delay in B2B quoting

    ...The workbook includes a Measurement log, Calculator inputs, Definitions and a fictional Worked example. The calculated results are based on the entered values. They are not industry benchmarks, recommended targets or performance guarantees. Published by Configure to WIN. Official online calculator: https://configure.win/resources/quote-turnaround-time-calculator Official source repository: https://github.com/configure-to-win/quote-turnaround-time-calculator-excel
    Downloads: 0 This Week
    Last Update:
    See Project
  • 9
    ForgePress

    ForgePress

    High-performance asynchronous Rust & WebAssembly (WASI) CMS engine.

    ...Database interactions are decoupled and optimized using atomic PostgreSQL JSONB block arrays to completely eliminate N+1 relational query bottlenecks. Read the full architecture case study and benchmarks at: https://azbrand.ca/blog/case-study-forgepress-rust-svelte-cms-rhai-wasm-sandboxed-plugins
    Downloads: 0 This Week
    Last Update:
    See Project
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • 10
    Bottleneck Calculator

    Bottleneck Calculator

    Check CPU and GPU balance with real time bottleneck analysis

    PC Bottleneck Calculator is a performance analysis tool that helps PC gamers and builders identify CPU or GPU bottlenecks in their systems. It provides accurate compatibility insights by comparing hardware data and real world benchmarks to estimate system balance. Users can instantly see how well their CPU and GPU pair together, test different configurations, and understand which component limits their gaming performance. www.pcbottleneckcalculator.io Built with a clean, responsive interface, the tool offers quick, data-driven results without requiring downloads or complex setup.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 11
    TFEL/MFront

    TFEL/MFront

    TFEL/MFront introduces DSLs based on C++ to handle material knowledge

    ...A particular focus was made on mechanical behaviours which are by essence more complex and may have significant impact on the numerical performances of mechanical simulations. Various performance benchmarks show that the code generated using MFront is in most cases on par or better than other implementations, generally written in Fortran. For mechanical behaviours, MFront introduces interfaces for various finite element or FTT solvers (Cast3M, Code-Aster, ZeBuLoN, Abaqus, Europlexus, AMITEX_FFT, etc..). The authors hope that it will prove usefull for researchers and engineers, in particular in the field of solid mechanics.
    Leader badge
    Downloads: 1 This Week
    Last Update:
    See Project
  • 12
    CPQ Metrics & KPI Dictionary

    CPQ Metrics & KPI Dictionary

    Standard CPQ metrics for quote speed, approvals and commercial control

    ...Includes an Excel workbook, CSV and JSON dictionaries, JSON Schema, dashboard specification, data requirements, implementation checklist and fictional examples. For pricing, sales operations, revenue operations, deal desk, CPQ, BI and analytics teams. No industry benchmarks, KPI targets or recommended thresholds are provided. Published by Configure to WIN. https://configure.win/resources/cpq-metrics
    Downloads: 1 This Week
    Last Update:
    See Project
  • 13
    B2B Quote Approval Workflow Template

    B2B Quote Approval Workflow Template

    Design and test B2B quote approval workflows in Excel

    ...The workbook includes Workflow overview, Approval rules, Approval groups, Workflow paths, Test scenarios, Implementation checklist, Definitions and a fictional Worked example. It does not execute live approvals, prescribe commercial thresholds, provide approval benchmarks or determine delegated authority. Published by Configure to WIN.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 14
    Cryptofolio
    Cryptofolio is an open-source, and self-hosted solution for tracking your cryptocurrency holdings. It features a web interface, an Android mobile app, and a cross-platform desktop application for Windows, macOS, and Linux. These three platforms all work using a RESTful API, which you'd have to host yourself. It can provide you with a quick glance at the market, while also keeping track of your assets and their value. It also includes a feature that allows you to share your portfolio in a...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 15
    GoModel

    GoModel

    The best AI Gateway 2026 - GoModel

    GoModel is the last AI Gateway you'll ever try. It's the most secure, the fastest, and the most reliable AI Gateway according to self-verifiable benchmarks. The best open-source alternative to LiteLLM, GoModel provides a unified OpenAI-compatible API for OpenAI, Anthropic, Gemini, Groq, xAI, Ollama, vLLM, Amazon Bedrock, OpenRouter, and many more providers. Built in Go, it's lightweight, resource-efficient, and designed for production workloads with extremely low latency. GoModel includes intelligent request routing, observability, guardrails, live logs, usage and cost tracking, rate limiting, virtual models, prompt rewriting, streaming support, and enterprise-ready authentication. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 16
    BTree implementation for Go

    BTree implementation for Go

    BTree provides a simple, ordered, in-memory data structure for Go

    ...The implementation favors minimal allocations and locality, making it attractive for indexing, query engines, and caches that need predictable iteration costs. A simple Item interface with a Less method defines ordering, keeping the API small and flexible for custom types. The library includes benchmarks and optional freelists so users can trade memory reuse for speed in hot paths.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 17
    MLPerf

    MLPerf

    Reference implementations of MLPerf™ training benchmarks

    ...The MLPerf Training working group draws on expertise in AI and the technology that powers AI from across the industry to design and create industry-standard benchmarks. Together, we create the reference implementations, rules, policies, and procedures to benchmark a wide variety of AI workloads.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 18
    DeepSeek Math

    DeepSeek Math

    Pushing the Limits of Mathematical Reasoning in Open Language Models

    ...The repository is likely to include fine-tuning routines or task datasets (e.g. MATH, GSM8K, ARB), demonstration notebooks, prompt templates, and evaluation results on math benchmarks. The goal is to push DeepSeek’s performance in domains that require rigorous symbolic steps, calculus, linear algebra, number theory, or multi-step derivations. The repo may also include modules that integrate external computational tools (e.g. a CAS / computer algebra system) or calculator assistance backends to enhance correctness. ...
    Downloads: 6 This Week
    Last Update:
    See Project
  • 19
    Evals

    Evals

    Evals is a framework for evaluating LLMs and LLM systems

    ...It includes utilities and APIs to plug in completion functions, manage prompts, wrap retries or error handling, and register new evaluation types. It also maintains a growing registry of standard benchmarks or “evals” that users can reuse (for example, tasks measuring reasoning, factual accuracy, or chain-of-thought capabilities). The design is modular so you can extend or compose new evals, integrate with your own model APIs, and capture rich metadata about each run (prompt, responses, metrics).
    Downloads: 0 This Week
    Last Update:
    See Project
  • 20
    DeepSeek VL

    DeepSeek VL

    Towards Real-World Vision-Language Understanding

    ...The model is likely used internally as the visual encoder backbone for agent use cases, to ground perception in downstream tasks (e.g. answering questions about a screenshot). The repository includes model weights (or pointers to them), evaluation metrics on standard vision + language benchmarks, and configuration or architecture files. It also supports inference tools for forwarding image + prompt through the model to produce text output. DeepSeek-VL is a predecessor to their newer VL2 model, and presumably shares core design philosophy but with earlier scaling, fewer enhancements, or capability tradeoffs.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21
    Qwen-Audio

    Qwen-Audio

    Chat & pretrained large audio language model proposed by Alibaba Cloud

    ...There is also an instruction-tuned version called Qwen-Audio-Chat which supports conversational interaction (multi-round), audio + text input, creative tasks and reasoning over audio. It uses multi-task training over many different audio tasks (30+), and achieves strong multi-benchmarks performance without task-specific fine‐tuning. It includes features such as flexible multi-run chat, audio understanding/reasoning, music appreciation, and also tool usage (e.g. voice editing).
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22
    Hiera

    Hiera

    A fast, powerful, and simple hierarchical vision transformer

    ...The repository provides installation options (from source or Torch Hub), a model zoo with pre-trained checkpoints, and code for evaluation and fine-tuning on standard benchmarks. Documentation emphasizes that model weights may have separate licensing and that the code targets practical experimentation for both research and downstream tasks. Community discussions cover topics like dataset pretrains, integration in other frameworks, and comparisons with related implementations. Security and contribution guidelines follow Meta’s open-source practices, and activity shows ongoing interest and usage across the community.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23
    Safety-Prompts

    Safety-Prompts

    Chinese safety prompts for evaluating and improving the safety of LLMs

    ...The repository also serves as a training resource for improving model alignment by providing examples of prompts that require safe reasoning and appropriate refusal behavior. In addition to evaluation prompts, the project references related tools and benchmarks for assessing model safety across different contexts.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 24
    Arthur Bench

    Arthur Bench

    Bench is a tool for evaluating LLMs for production use cases

    Bench is a tool for evaluating LLMs for production use cases. Whether you are comparing different LLMs, considering different prompts, or testing generation hyperparameters like temperature and # tokens, Bench provides one touch point for all your LLM performance evaluation.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 25
    DeepSeek LLM

    DeepSeek LLM

    DeepSeek LLM: Let there be answers

    ...The repo includes an “evaluation” folder (with results like math benchmark scores) and code artifacts (e.g. pre-commit config) that support model development and deployment. According to the evaluation files, DeepSeek LLM 67B Chat achieves strong performance on math benchmarks under both chain-of-thought (CoT) and tool-assisted reasoning modes. The model is trained from scratch, reportedly on a vast multilingual + code + reasoning dataset, and competes with other open or open-weight models. The architecture mirrors established decoder-only transformer families: pre-norm structure, rotational embeddings (RoPE), grouped query attention (GQA), and mixing in languages and tasks. ...
    Downloads: 4 This Week
    Last Update:
    See Project