Alternatives to Mercury 2.5
Compare Mercury 2.5 alternatives for your business or organization using the curated list below. SourceForge ranks the best alternatives to Mercury 2.5 in 2026. Compare features, ratings, user reviews, pricing, and more from Mercury 2.5 competitors and alternatives in order to make an informed decision for your business.
-
1
Mercury One Plus
CrisSoft
Mercury One Plus is a Medical Practice Management solution that puts the fundamentals of Revenue Cycle Management at your fingertips; it acts as a stepping stone from standard billing to intermediate billing. Mercury One Plus is offered exclusively on the cloud, with the highest level of security- you can access your data anywhere 24/7. A complete product with big functionality, Mercury One Plus includes: patient demographics input, 100 plus reports to choose from, charge entry, full history of patient activity, ERA posting, credit card acceptance, and much more. Mercury Products are HIPAA compliant with a guaranteed connection to any clearinghouse or insurance company. Mercury One Plus's automated job system will facilitate a daily system tune-up: housecleaning; folder maintenance; daily backups; 837 exports; 835 imports;HL7. All subscriptions come with the expert help of CrisSoft Support. Willing to partner/intergrate with all EMR's through REST.Starting Price: $249.00 -
2
Mercury Medical
CrisSoft
Awarded among the Top 10 MPM and RCM solutions, Mercury Medical is a robust medical billing solution. With over 400 customizable reports, built in Scheduler and Patient Portal, Mercury Medical is the perfect solution for major billing, suitable for multiple specialties and RCM processes Mercury Medical is a reliable, proven professional Accounts Receivable solution that will reduce processing times, shorten payment cycles and increase cash flow. Mercury Medical is fully configurable to any vertical or process including: Anesthesiology, University, Physical Therapy, and more. Mercury Products are HIPAA compliant with a guaranteed connection to any clearinghouse or insurance company. Mercury Medical’s automated job system will facilitate a daily system tune-up: housecleaning; folder maintenance; daily backups; 837 exports; 835 imports; HL7. All subscriptions come with the expert help of CrisSoft Support.Starting Price: $440.00 -
3
Nemotron 3 Ultra
NVIDIA
Nemotron 3 Nano is a compact, open large language model in NVIDIA’s Nemotron 3 family, designed for efficient agentic reasoning, conversational AI, and coding tasks. It uses a hybrid Mixture-of-Experts Mamba-Transformer architecture that activates only a small subset of parameters per token, enabling low-latency inference while maintaining strong accuracy and reasoning performance. It has approximately 31.6 billion total parameters with around 3.2 billion active (3.6 billion including embeddings), allowing it to achieve higher accuracy than previous Nemotron 2 Nano while using less computation per forward pass. Nemotron 3 Nano supports long-context processing of up to one million tokens, enabling it to handle large documents, multi-step workflows, and extended reasoning chains in a single pass. It is designed for high-throughput, real-time execution, excelling in multi-turn conversations, tool calling, and agent-based workflows where tasks require planning, reasoning, and more. -
4
Mercury Coder
Inception Labs
Mercury, the latest innovation from Inception Labs, is the first commercial-scale diffusion large language model (dLLM), offering a 10x speed increase and significantly lower costs compared to traditional autoregressive models. Built for high-performance reasoning, coding, and structured text generation, Mercury processes over 1000 tokens per second on NVIDIA H100 GPUs, making it one of the fastest LLMs available. Unlike conventional models that generate text one token at a time, Mercury refines responses using a coarse-to-fine diffusion approach, improving accuracy and reducing hallucinations. With Mercury Coder, a specialized coding model, developers can experience cutting-edge AI-driven code generation with superior speed and efficiency.Starting Price: Free -
5
Mercury Edit 2
Inception
Mercury Edit 2 is part of Inception Labs’ Mercury family of AI models, designed to perform high-speed reasoning, coding, and editing tasks using a fundamentally different architecture from traditional large language models. It builds on Mercury 2, a diffusion-based reasoning model that generates and refines entire outputs in parallel rather than producing text token by token, enabling significantly faster performance and more responsive editing workflows. Instead of acting like a sequential “typewriter,” the system behaves more like an editor, starting with a rough draft and iteratively improving it across multiple tokens at once, which allows for real-time interaction and rapid iteration in tasks such as code editing, content generation, and agent-based workflows. This architecture delivers throughput of up to around 1,000 tokens per second, making it several times faster than conventional models while maintaining competitive reasoning quality across benchmarks.Starting Price: $0.25 per 1M input tokens -
6
Mercury 2
Inception
Mercury 2 is the first reasoning model fast enough to pick up the phone, a reasoning diffusion language model built for real-time voice agents. Instead of making callers wait through seconds of dead air while an autoregressive model generates thinking tokens one by one, Mercury 2 uses a diffusion large language model architecture to generate tokens in parallel, decoding 1000+ tokens per second on standard NVIDIA GPUs. That speed is fast enough to run a full reasoning pass and start speaking within the latency budget of a natural conversation, reducing the cost of reasoning from seconds of silence to roughly 300 milliseconds. Mercury models work by corrupting clean text into noise, then training a standard Transformer to reverse the process and predict clean text across all positions simultaneously. Because each denoising pass touches many tokens, generation uses the GPU more efficiently than one-token-at-a-time decoding, making custom-silicon-like speed possible on NVIDIA H100s. -
7
Gemini 3.5 Flash-Lite
Google
Gemini 3.5 Flash-Lite is Google’s fastest model in the Gemini 3.5 series, designed for low-latency tasks and high-throughput developer workflows such as agentic search, document processing, coding, and large-scale data analysis. It delivers 350 output tokens per second and significantly improves on previous Flash-Lite generations in both quality and agentic performance. Developers can configure its thinking level to match the workload: minimal or low thinking supports fast execution for high-volume tasks, while higher thinking levels enable more complex, multi-step subagent workflows. Built-in computer-use capabilities allow the model to interact reliably with digital environments across supported surfaces. Gemini 3.5 Flash-Lite also advances coding, long-context understanding, and real-world task execution, outperforming Gemini 3.1 Flash-Lite across key evaluations and even surpassing Gemini 3 Flash on several agentic and software-engineering benchmarks.Starting Price: $0.30 per 1M input tokens -
8
Inception Labs
Inception Labs
Inception Labs is pioneering the next generation of AI with diffusion-based large language models (dLLMs), a breakthrough in AI that offers 10x faster performance and 5-10x lower cost than traditional autoregressive models. Inspired by the success of diffusion models in image and video generation, Inception’s dLLMs introduce enhanced reasoning, error correction, and multimodal capabilities, allowing for more structured and accurate text generation. With applications spanning enterprise AI, research, and content generation, Inception’s approach sets a new standard for speed, efficiency, and control in AI-driven workflows. -
9
Mercurial
Mercurial
Mercurial is a free, distributed source control management tool. It efficiently handles projects of any size and offers an easy and intuitive interface. Mercurial efficiently handles projects of any size and kind. Every clone contains the whole project history, so most actions are local, fast and convenient. Mercurial supports a multitude of workflows and you can easily enhance its functionality with extensions. Mercurial strives to deliver on each of its promises. Most tasks simply work on the first try and without requiring arcane knowledge. The functionality of Mercurial can be increased with extensions, either by activating the official ones which are shipped with Mercurial or downloading some from the wiki or by writing your own. Extensions are written in Python and can change the workings of the basic commands, add new commands and access all the core functions of Mercurial. -
10
MercuryDPM
MercuryDPM
MercuryDPM is an open source code for discrete particle simulations, designed to simulate the motion of particles or atoms by applying forces and torques from external body forces, such as gravity or magnetic fields, and from particle interaction laws. For granular particles, these forces are typically contact forces, including elastic, plastic, viscous, and frictional interactions, while molecular simulations can use interaction potentials such as Lennard-Jones. MercuryDPM is written as a versatile, object-oriented C++ code and is built to be understandable, flexible, and extensible for researchers and engineers who need to create new simulation models. It is developed extensively for granular applications, while remaining adaptable to other particle-based systems and long-range interactions. Its documentation guides users through installation, running simulations, visualization, analysis, and creating new MercuryDPM codes to model systems of their choice.Starting Price: Free -
11
RemObjects Mercury
RemObjects Mercury
Mercury is an implementation of the BASIC programming language that is fully code-compatible with Microsoft Visual Basic.NET™, but takes it to the next level, and to new horizons. With Mercury, you will be able to build your existing VB.NET projects and leverage your Visual Basic™ language experience to write code for any modern target platform. You can mix Mercury code with any of the other five Elements languages in the same project if you like! The Mercury language will be deeply integrated into our development environments. Develop your projects in our smart yet lightweight IDEs, Water on Windows or Fire on Mac, with project templates, code completion, integrated debugging for all platforms, and many other advanced development features. Of course, Mercury will also integrate into Visual Studio™ 2017, 2019 and 2022. With Elements, all languages are created equal. Even within the same project, you can mix Mercury, C#, Swift, Java, Oxygene and Go.Starting Price: $49 per month -
12
MercuryTel
Mercury Network
With our unique combination of phone service, technology, and expertise, and backed by our amazing PhonePro support, MercuryTel is the business phone service you’ve been longing for. Scalable to meet the needs of organizations small and large, MercuryTel is packed with features, exceptionally customizable, and incredibly affordable. Your local and long-distance calling, Cisco®, Mitel®, and Yealink® digital phones, and unlimited support are included in the fixed, monthly charge. Until now, having your own phone system was the only way to get the most advanced capabilities, but phone systems are expensive! As the technology ages, costly repairs and upgrades are required, and phone system dealers charge for every add, move and change. With MercuryTel, there is no phone system to buy and every feature is included at no extra charge! Your business will sound more professional, communicate more effectively and work more productively, without the big phone system expense.Starting Price: $10.40 per month -
13
Celeris-1
Celeris-1
Celeris-1 is a low-latency, general-purpose language model platform and a diffusion model designed to deliver frontier-level intelligence at dramatically higher speed. Instead of generating one token at a time like traditional autoregressive models, Celeris uses a diffusion-based inference architecture that enables parallel generation and response times measured in milliseconds. On its published MMLU-Pro benchmark, Celeris-1 reaches 75.9 accuracy with a 158 ms median response time and 1,664 output tokens per second, placing it within a few points of frontier models while running more than 10x faster. The model is exposed through an OpenAI-compatible API, so developers can point existing SDKs and clients at Celeris with minimal code changes. Streaming is enabled for interactive applications, with responses as low as 24 ms and no buffering or batch delay.Starting Price: $0.20 per 1M tokens -
14
Mercury Housing
RMS
Mercury Housing enables housing and conference staff to deliver revolutionary, customized content to their students and housing team. This includes custom housing applications, contracts, electronic signatures, online payments, student self-assignment, staff screens, and menus. Mercury also delivers a robust set of reporting and administrative tools, without the need for third-party plug-ins. Using drag-and-drop design features on all popular browsers, Mercury delivers business processes to students and staff in a completely customizable environment. Mercury is ready and available across all of your devices. There are no limits on the number of online business processes you can design. Even better, our Mercury Tools enables peer-to-peer sharing opportunities so that you can use online processes that other student service professionals have designed and launched. Each business process can include unlimited pages, with unlimited content components, in any order! -
15
Gemini 2.0 Flash-Lite
Google
Gemini 2.0 Flash-Lite is Google DeepMind's lighter AI model, designed to offer a cost-effective solution without compromising performance. As the most economical model in the Gemini 2.0 lineup, Flash-Lite is tailored for developers and businesses seeking efficient AI capabilities at a lower cost. It supports multimodal inputs and features a context window of one million tokens, making it suitable for a variety of applications. Flash-Lite is currently available in public preview, allowing users to explore its potential in enhancing their AI-driven projects. -
16
Mercury Network
Mercury Network
With Mercury Network, we'll help you find the best domain name for your business or personal needs. A variety of domains are available to be registered quickly, easily, and affordably. When you host the domain with Mercury Network, registration is only $15 a year! Get Exchange-level e-mail, calendaring, and collaboration for a fraction of the cost. With superior support and outstanding reliability, Mercury Network E-Mail is the best-in-class Microsoft® Exchange alternative for businesses and individuals. Developing a successful presence on the world wide web takes a combination of quality content, dynamic programming, and powerful graphics. Our web development team has the knowledge and expertise to build a complete and thoughtful presence for your business on the world wide web. WebsiteOS gives you complete management capability using a standard web browser and saves you time, money, and resources by transferring site administration to you instead of making you call technical support.Starting Price: $6.49 per month -
17
FTD Mercury
Florists Transworld Delivery
For more than 30 years, FTD has led the floral industry in bringing the best business technology systems to florists worldwide. FTD Mercury and Mercury Cloud can help you grow your business, drive sales and increase satisfaction. Thousands of florists across North America rely on Mercury Technology to help their businesses run more efficiently. Innovative. Unique options save you time and money and help increase profits. Easy to use. Manage local, florist-to-florist and FTD.com orders with one tool. Dependable. Regular updates and enhancements keep your business running smoothly. Mercury HQ is a cloud-based system that lets you manage your shop from anywhere using your phone, tablet, or computer. With real-time synching for access wherever life takes you, Mercury HQ helps you accept and transact orders like never before. Accept, send, reject and track orders whether you're in your shop or out walking the dog. -
18
Mercury Rugged Edge Servers
Mercury
Rugged subsystems engineered to bring the latest Silicon Valley technology to every inhospitable corner of the globe. No matter the end use, environment or security requirements, when your mission is critical Mercury rugged servers and embedded processing subsystems are the only options. Compute-heavy applications, including AI, signals intelligence and sensor fusion, have driven the need for real-time big data analysis at the edge. Mercury's ruggedized servers and processing solutions make the most sophisticated Silicon Valley technology profoundly more accessible to the A&D industry for actionable insights in the field. Our fully configurable systems are engineered to push computing to the cutting edge of innovation. Mercury’s airborne and mission computers accelerate complex airborne applications like artificial intelligence (AI) and streamline integration, technology refresh, and safety certification. -
19
Repositery
Repositery
SVN, Mercurial and Git cloud hosting with Trac project management. Startup or multinational, whenever you run a project requiring several programmers, version control is a must-have. Repositery is a one-stop cloud solution for your SVN, Mercurial and Git repository requirements. Repositery is packed with features to help you manage your code and your project with ease. Repositery offers fast and reliable hosting for Git, Mercurial and Subversion repositories. Mix and match unlimited repositories of any type per project. Git is the most famous version control system and you get unlimited git repos at Repositery. Repositery offers SVN hosting which is used by many entities around the world. Mercurial is a distributed revision control tool for software developers which is available at Repositery. Trac is an open-source, Web-based project management and bug tracking system.Starting Price: $3 per month -
20
Mercury Recruit
Mercury360
Mercury Recruit is a modern, AI-powered Applicant Tracking System (ATS) designed to help organisations hire smarter, faster, and more efficiently. Whether you're a growing SMB, a large enterprise, or a recruitment agency managing multiple clients, Mercury Recruit centralises your entire hiring workflow into one platform. Key capabilities include AI-powered candidate ranking that automatically scores applicants against your exact job criteria, dramatically reducing manual screening time by up to 80%. A dynamic candidate portal and personalised talent pool make it easy to manage applicants from multiple sources in one place. Customisable workflows and real-time team collaboration ensure consistent, structured evaluations across every hire. Beyond speed, Mercury Recruit helps you improve quality of hire, reduce total recruitment costs, lower reliance on job boards and agencies, and build stronger teams with best-fit talent.Starting Price: $150 per month -
21
MercuryAI
MercuryAI
MercuryAI simplifies working with multiple AI models by providing a single platform to: - Access multiple large language models (LLMs) from one interface - Orchestrate requests to the best AI model for your task - Integrate AI capabilities into apps via developer-friendly APIs - Automate AI workflows and repetitive tasks - Scale AI usage for startups and enterprises It solves the problem of fragmented AI tools and dashboards, saving time and improving efficiency for developers and teams. -
22
Mercury Office
Ease Technology
Mercury Office is a highly sophisticated travel agency database designed from the ground up with the travel agent in mind. Front office staff will find it effortlessly easy to create booking files for customers, keep track of payments and send welcome home letters. Meanwhile, back office staff can manage supplier payments, bank reconciliation, staff salaries, and company profitability. One of the best features of Mercury Office is that it's completely web-based making it accessible from anywhere in the world and means your data is safe and secure in every circumstance. Developed with real front office travel agents, each part of Mercury has been carefully designed to make it easy to use. With a comprehensive reporting engine, back office administrators will be able to find all the facilities they need from reconciliation to supplier payments.Starting Price: $65 per month -
23
Mercury xRM
Mercury xRM
We help recruiters like you place more candidates and fill vacancies faster. Our market-leading and cloud-based software streamlines your day, keeping you ahead of your competition. While they catch up with admin, you catch up with clients and candidates. As a strategic Microsoft partner, Mercury xRM fully embraces Microsoft cloud technology. Powered by Microsoft Dynamics 365, Microsoft’s Power Platform, and Office 365, Mercury xRM works with and like other Microsoft applications that your recruiters already work with every day. This makes it easier to learn and simpler to manage. Our staffing and recruiting CRM platform enables recruitment and staffing companies to take advantage of the Microsoft Power Platform’s powerful capabilities to increase placements and development. You gain access to award-winning technology that centralizes your data, improves collaboration, strengthens and automates processes, and provides actionable business insights. -
24
Gemini 3.1 Flash-Lite
Google
Gemini 3.1 Flash-Lite is Google’s fastest and most cost-efficient model in the Gemini 3 series, designed for high-volume developer workloads. It delivers strong performance at scale while maintaining affordability, with pricing set at $0.25 per million input tokens and $1.50 per million output tokens. The model significantly improves speed, offering a 2.5x faster time to first answer token and a 45% increase in output speed compared to Gemini 2.5 Flash. Despite its lower cost tier, it achieves high benchmark results, including an Elo score of 1432 and strong performance across reasoning and multimodal evaluations. Gemini 3.1 Flash-Lite supports adaptive “thinking levels,” allowing developers to control how much reasoning power is used for different tasks. It is suitable for large-scale applications such as translation, content moderation, user interface generation, and simulation building. -
25
Mercury Mail Transport System
Pegasus
Mail from the outside world is received by Mercury and placed in the addressee's mailbox, where the user can access it at any later point. Messages sent by local users to the outside world are passed to Mercury, which then takes whatever steps are necessary to deliver them, removing the burden from the user's workstation and allowing him to continue with other work. If you connect using a dialup connection, then only the mail server needs to be able to access that connection. Your workstations do not need their own modems or Internet accounts. The mail server can continue processing mail even when the client workstations are turned off, allowing functions that depend on a continuously available service, such as automatic replies and auto-forwarding. -
26
flowerSoft Silver
flowerSoft Systems
In business for over 30 years, flowerSoft Silver has always been the leader in technological advances and customer support. Recently FTD decided that we should pay them a considerable amount of money each year to allow our users to send business their way. Because of this new policy by FTD, flowerSoft will no longer interface with Mercury Direct, something it has been doing since the Mercury 3000 was first introduced. We will not be forced to charge our customer considerable more money each month so that they can send FTD their outgoing orders. There are at least 2 other wire services that are not as greedy as FTD and we are happy to offer interfaces with their systems at no cost to our customers.Starting Price: $1999 per user -
27
Claude Haiku 4.5
Anthropic
Anthropic has launched Claude Haiku 4.5, its latest small-language model designed to deliver near-frontier performance at significantly lower cost. The model provides similar coding and reasoning quality as the company’s mid-tier Sonnet 4, yet it runs at roughly one-third of the cost and more than twice the speed. In benchmarks cited by Anthropic, Haiku 4.5 meets or exceeds Sonnet 4’s performance in key tasks such as code generation and multi-step “computer use” workflows. It is optimized for real-time, low-latency scenarios such as chat assistants, customer service agents, and pair-programming support. Haiku 4.5 is made available via the Claude API under the identifier “claude-haiku-4-5” and supports large-scale deployments where cost, responsiveness, and near-frontier intelligence matter. Claude Haiku 4.5 is available now on Claude Code and our apps. Its efficiency means you can accomplish more within your usage limits while maintaining premium model performance.Starting Price: $1 per million input tokens -
28
ByteDance Seed
ByteDance
Seed Diffusion Preview is a large-scale, code-focused language model that uses discrete-state diffusion to generate code non-sequentially, achieving dramatically faster inference without sacrificing quality by decoupling generation from the token-by-token bottleneck of autoregressive models. It combines a two-stage curriculum, mask-based corruption followed by edit-based augmentation, to robustly train a standard dense Transformer, striking a balance between speed and accuracy and avoiding shortcuts like carry-over unmasking to preserve principled density estimation. The model delivers an inference speed of 2,146 tokens/sec on H20 GPUs, outperforming contemporary diffusion baselines while matching or exceeding their accuracy on standard code benchmarks, including editing tasks, thereby establishing a new speed-quality Pareto frontier and demonstrating discrete diffusion’s practical viability for real-world code generation.Starting Price: Free -
29
Kallithea
Kallithea
Kallithea, a member project of Software Freedom Conservancy, is a GPLv3'd, Free Software source code management system that supports two leading version control systems, Mercurial and Git, and has a web interface that is easy to use for users and admins. You can install Kallithea on your own server and host repositories for the version control system of your choice. Supports both Mercurial and Git wire protocols, accessible over HTTPS and SSH. A powerful access management system lets you decide who has access to the repository, and what operations they’re entitled to do. All requests are authenticated and logged, giving the administrator an ability to review users’ activity. Kallithea supports LDAP, making it easy to use your existing authentication system. You can integrate your instance with an issue tracker of your choice using the JSON-RPC API and the extensions interface. -
30
DiffusionGemma
Google
DiffusionGemma is an experimental open model that explores text diffusion, an exceptionally fast approach to text generation. Released under an Apache 2.0 license, this 26B Mixture of Experts (MoE) model moves beyond the sequential token-by-token processing of typical autoregressive Large Language Models (LLMs). Instead, it generates entire blocks of text simultaneously, delivering up to 4x faster text generation on GPUs. Built on the intelligence-per-parameter of the Gemma 4 family and Gemini Diffusion research, DiffusionGemma integrates a novel diffusion head designed to maximize generation speed. It is designed for researchers and developers exploring speed-critical, interactive local workflows such as in-line editing, rapid iteration, and non-linear text structures. By shifting the decode bottleneck from memory bandwidth to compute, it can generate more than 1,000 tokens per second on a single NVIDIA H100 and more than 700 tokens per second on an NVIDIA GeForce RTX 5090.Starting Price: Free -
31
Mercury Editor
Mercury Editor
Mercury is a full featured HTML5 editor. It was built from the ground up to help your team get the most out of content editing in modern browsers. Mercury comes bundled as a Rails Engine, so just include it in your Gemfile. Or download the current bundled package if you're not using Rails. We don't inject javascript or css into your production pages so you're free to use whatever frameworks you want without having to worry about conflicts. Easily add or remove toolbar items or create entirely new tools. Any toolbar item can be tied to an action using behaviors and the command pattern. Full HTML, Simple, Markdown, Snippet and Image regions are supported by default, but you can just extend the base regions to build your own types. Built on top of the HTML5 contentEditable features, it natively supports the all the fancy new HTML5 elements, syntax, and JavaScript APIs. -
32
Fisheye
Atlassian
Search, track, and visualize code changes. Visualize and report on activity and search for commits, files, revisions, or teammates across SVN, Git, Mercurial, CVS and Perforce. View changes with a side-by-side or unified diff tool and link your Jira Software issues directly to diffs, changeset details, or full source. Get a graphical representation of activity in your source, report on lines of code over time, and get a visual audit trail of changes. Follow what's happening throughout your projects with activity streams showing commits, Jira Software issues, and Crucible review activities across your team. Find code fast with search using any artifact in your code: file names, commit messages, authors, text, and even historical changes. Browse, index, and search all your source from all your source code management systems including SVN, Git, Mercurial, CVS and Perforce – all in one tool. Upgrade your workflow with Jira Software, Bitbucket Server, Bamboo and more.Starting Price: $10 one-time payment -
33
Perforce TeamHub
Perforce
Your code repository software is where you store your source code. This might be a Mercurial, Git, or SVN repository. Perforce TeamHub (formerly Helix TeamHub) can host your source code repository, whether it’s Mercurial, Git, or SVN. You can add multiple repositories in one project — or create a separate project for each repository. Perforce TeamHub can host more than your code repositories. You can manage and maintain all of your software assets in one spot. This includes build artifacts (Maven, Ivy) and Docker container registries. It also includes private file sharing through WebDAV repositories for your other binary files. You can use TeamHub on its own or alongside P4 to maintain a single source of truth across development teams. For example, you can keep large binary files in P4, then combine those files with Git assets from Perforce TeamHub in a hybrid workspace to achieve high build performance.Starting Price: $1.05/month -
34
Phi-4-mini-flash-reasoning
Microsoft
Phi-4-mini-flash-reasoning is a 3.8 billion‑parameter open model in Microsoft’s Phi family, purpose‑built for edge, mobile, and other resource‑constrained environments where compute, memory, and latency are tightly limited. It introduces the SambaY decoder‑hybrid‑decoder architecture with Gated Memory Units (GMUs) interleaved alongside Mamba state‑space and sliding‑window attention layers, delivering up to 10× higher throughput and a 2–3× reduction in latency compared to its predecessor without sacrificing advanced math and logic reasoning performance. Supporting a 64 K‑token context length and fine‑tuned on high‑quality synthetic data, it excels at long‑context retrieval, reasoning tasks, and real‑time inference, all deployable on a single GPU. Phi-4-mini-flash-reasoning is available today via Azure AI Foundry, NVIDIA API Catalog, and Hugging Face, enabling developers to build fast, scalable, logic‑intensive applications. -
35
Gemini Diffusion
Google DeepMind
Gemini Diffusion is our state-of-the-art research model exploring what diffusion means for language and text generation. Large-language models are the foundation of generative AI today. We’re using a technique called diffusion to explore a new kind of language model that gives users greater control, creativity, and speed in text generation. Diffusion models work differently. Instead of predicting text directly, they learn to generate outputs by refining noise, step by step. This means they can iterate on a solution very quickly and error correct during the generation process. This helps them excel at tasks like editing, including in the context of math and code. Generates entire blocks of tokens at once, meaning it responds more coherently to a user’s prompt than autoregressive models. Gemini Diffusion’s external benchmark performance is comparable to much larger models, whilst also being faster. -
36
Nemotron 3.5 Lightning
NVIDIA
NVIDIA Nemotron 3.5 Lightning is an open 30B-parameter mixture-of-experts model with 3B active parameters, designed for high-volume, low-latency execution in long-running and always-on AI agents. Built for the execution layer of agentic systems, it handles frequent tasks such as tool calls, output validation, routine commands, and subagent delegation while larger reasoning models focus on planning and orchestration. Its MoE architecture activates only a fraction of parameters for each token, combining the capacity of a larger model with lower compute requirements. The model is trained for popular agent harnesses and supports speculative decoding through multi-token prediction, DFlash, and DSpark to improve inference speed across different serving scenarios. It is available with BF16 and NVFP4 checkpoints and can run from local systems such as DGX Spark and GeForce RTX hardware to data center environments. -
37
Gemini 4
Google
Gemini 4 is Google’s next-generation Gemini model family currently in development after the release of Gemini 3.6 Flash and Gemini 3.5 Flash-Lite. Google has confirmed that pre-training for Gemini 4 has begun, positioning it as the company’s most ambitious model training effort yet. The model is expected to advance Google’s frontier AI work across reasoning, coding, multimodal understanding, agentic workflows, and enterprise AI use cases. Because Gemini 4 has not been publicly released yet, official pricing, model cards, benchmarks, API details, and availability have not been published. Gemini 4 follows Google’s broader Gemini strategy of building models for developers, enterprises, consumer apps, and AI-powered products across Google’s ecosystem. Built for the next stage of AI agents and intelligent applications, Gemini 4 is likely to become a major foundation for future Google AI products once it becomes available. -
38
Reka Flash 3
Reka
Reka Flash 3 is a 21-billion-parameter multimodal AI model developed by Reka AI, designed to excel in general chat, coding, instruction following, and function calling. It processes and reasons with text, images, video, and audio inputs, offering a compact, general-purpose solution for various applications. Trained from scratch on diverse datasets, including publicly accessible and synthetic data, Reka Flash 3 underwent instruction tuning on curated, high-quality data to optimize performance. The final training stage involved reinforcement learning using REINFORCE Leave One-Out (RLOO) with both model-based and rule-based rewards, enhancing its reasoning capabilities. With a context length of 32,000 tokens, Reka Flash 3 performs competitively with proprietary models like OpenAI's o1-mini, making it suitable for low-latency or on-device deployments. The model's full precision requires 39GB (fp16), but it can be compressed to as small as 11GB using 4-bit quantization. -
39
Kimi K2
Moonshot AI
Kimi K2 is a state-of-the-art open source large language model series built on a mixture-of-experts (MoE) architecture, featuring 1 trillion total parameters and 32 billion activated parameters for task-specific efficiency. Trained with the Muon optimizer on over 15.5 trillion tokens and stabilized by MuonClip’s attention-logit clamping, it delivers exceptional performance in frontier knowledge, reasoning, mathematics, coding, and general agentic workflows. Moonshot AI provides two variants, Kimi-K2-Base for research-level fine-tuning and Kimi-K2-Instruct pre-trained for immediate chat and tool-driven interactions, enabling both custom development and drop-in agentic capabilities. Benchmarks show it outperforms leading open source peers and rivals top proprietary models in coding tasks and complex task breakdowns, while its 128 K-token context length, tool-calling API compatibility, and support for industry-standard inference engines.Starting Price: Free -
40
Gemini 3 Flash
Google
Gemini 3 Flash is Google’s latest AI model built to deliver frontier intelligence with exceptional speed and efficiency. It combines Pro-level reasoning with Flash-level latency, making advanced AI more accessible and affordable. The model excels in complex reasoning, multimodal understanding, and agentic workflows while using fewer tokens for everyday tasks. Gemini 3 Flash is designed to scale across consumer apps, developer tools, and enterprise platforms. It supports rapid coding, data analysis, video understanding, and interactive application development. By balancing performance, cost, and speed, Gemini 3 Flash redefines what fast AI can achieve. -
41
Apache NetBeans
Apache Software Foundation
Apache NetBeans is a versatile, open-source Integrated Development Environment (IDE) used for developing applications across a wide range of programming languages, including Java, JavaScript, PHP, HTML5, and C/C++. Known for its modular architecture, NetBeans provides robust tools and features that cater to the needs of developers working on desktop, mobile, and web applications. It includes intelligent code editing, debugging, and profiling capabilities, along with a built-in visual GUI builder for designing Java-based user interfaces. NetBeans also offers support for version control systems like Git, SVN, and Mercurial, facilitating seamless team collaboration. As an Apache Software Foundation project, NetBeans benefits from an active community that continuously improves and expands its functionality, making it a reliable and flexible choice for developers across various domains.Starting Price: Free -
42
SWE-1.5
Cognition
SWE-1.5 is the latest agent-model release by Cognition, purpose-built for software engineering and characterized by a “frontier-size” architecture comprising hundreds of billions of parameters and optimized end-to-end (model, inference engine, and agent harness) for both speed and intelligence. It achieves near-state-of-the-art coding performance and sets a new benchmark in latency, delivering inference speeds up to 950 tokens/second, roughly six times faster than its predecessor Haiku 4.5 and thirteen times faster than Sonnet 4.5. The model was trained using extensive reinforcement learning in realistic coding-agent environments with multi-turn workflows, unit tests, quality rubrics, and browser-based agentic execution; it also benefits from tightly integrated software tooling and high-throughput hardware (including thousands of GB200 NVL72 chips and a custom hypervisor infrastructure). -
43
Mercury Policy & Claims Administration
Quick Silver Systems
Mercury by Quick Silver Systems allows Automobile, Property, and Casualty insurance carriers to easily rate, quote, bind, make payments, and report claims online. Minimize customer service calls through online document access, bill payments, and first notice of loss. Modular API based system allows seamless integration with new or existing data providers. Fully digital document production and 100% web-based system works on any device. Create custom, event-driven work-flows with our visual work-flow designer. Access the most up-to-date information on Written, Earned, and Unearned premiums. Automatically save every page, card, report, email, and more to review and share with associates. Collect currency in any digital format including: ACH, EFT, Electronic Checks, Credit, or Bank Card. Information Technology within an insurance company not only needs a system that provides wide accessibility. -
44
Mercurial Finance
Mercurial Finance
Mercurial is building new liquidity systems to maximise the utility and yield of stable assets on Solana. As the DeFi ecosystem on Solana grows, there will be many different variants of collateralized, wrapped, and synthetic assets in the space. Our most immediate objective is to provide the best liquidity for all the major stable and pegged assets on Solana, which we started with our Mainnet beta. Our focus will be on stable coins because they represent a major part of the DeFi demand across synthetic assets creation, swapping, and lending. Robust availability of stablecoin liquidity is crucial to any DeFi ecosystem. Moving forward, we are focused on building dynamic vaults, which are market making vaults providing low slippage swaps for stables, while also improving LP profits with dynamic fees and flexible capital allocation. -
45
Mercury
Mercury
Effective self-service and faster responses to your customers' questions, around the clock, lead to measurable improvements in your customer satisfaction. Dramatically increase your conversion rate with personalized product advice, suggestions, and purchase decision simplification. Automate your service effectively and have requests resolved before they become tickets. This significantly reduces the workload of your service team. Mercury's unique dialog technology enables an unmatched level of context, personalization, and intelligent dialog. This pays off for you by enabling more complex use cases and noticeably better UX. The active learning behavior combines two decisive advantages: It increases dialog robustness and guides your customers to the goal even with difficult questions, while leading to independent improvement of language comprehension.Starting Price: €500 per month -
46
Ideogram AI
Ideogram AI
Ideogram AI is a text to image AI image generator. Ideogram's technology is based on a new type of neural network called a diffusion model. Diffusion models are trained on a large dataset of images, and they can then generate new images that are similar to the images in the dataset. However, unlike other generative AI models, diffusion models can also be used to generate images in a specific style. -
47
Inception CRM
D3S
Inception CRM is a robust and user-friendly CRM solution for primary and specialty care as well as retail pharmacy life science field sales teams. A GDPR-compliant solution, Inception CRM is fully adapted to the needs of European life science organizations. It offers end-to-end support to reps in the field, helping them plan, execute and optimize their sales campaigns. Inception CRM guides users through their daily tasks, while providing valuable insight into their customers, territories and opportunities. Inception CRM’s powerful search helps sales reps quickly find the right customers, while detailed customer cards tell them everything they need to know. Inception’s intelligent workflow-based planner keeps field sales reps productively focused on the right tasks so that every sales campaign is a success. -
48
GPT-Realtime-2.1
OpenAI
GPT-Realtime-2.1 is OpenAI’s reasoning model with tool use for low-latency voice agents and complex speech-to-speech workflows. It updates GPT-Realtime-2 with improved alphanumeric recognition, silence and noise handling, and interruption behavior, helping applications understand spoken code, manage imperfect audio, and respond more naturally when users pause or talk over the agent. Developers can configure reasoning effort to balance deeper thinking against latency and output usage, while strong instruction following helps the model stay aligned with a defined role, tone, and workflow. It accepts and produces both audio and text, can take images as input, and supports function calling so an agent can retrieve information or perform actions during a conversation. The model has a 128,000-token context window, supports up to 32,000 output tokens, and includes reasoning-token support for extended interactions.Starting Price: $0.40 per cached input -
49
GLM-4.6V
Z.ai
GLM-4.6V is a state-of-the-art open source multimodal vision-language model from the Z.ai (GLM-V) family designed for reasoning, perception, and action. It ships in two variants: a full-scale version (106B parameters) for cloud or high-performance clusters, and a lightweight “Flash” variant (9B) optimized for local deployment or low-latency use. GLM-4.6V supports a native context window of up to 128K tokens during training, enabling it to process very long documents or multimodal inputs. Crucially, it integrates native Function Calling, meaning the model can take images, screenshots, documents, or other visual media as input directly (without manual text conversion), reason about them, and trigger tool calls, bridging “visual perception” with “executable action.” This enables a wide spectrum of capabilities; interleaved image-and-text content generation (for example, combining document understanding with text summarization or generation of image-annotated responses).Starting Price: Free -
50
AppHarbor
AppHarbor
AppHarbor is a fully hosted .NET Platform as a Service. AppHarbor can deploy and scale any standard .NET application to the cloud. AppHarbor is used by thousands of developers and businesses to host anything from personal blogs to high traffic web applications. AppHarbor lets you instantly deploy and scale .NET applications using your favorite versioning tool. Installing add-ons is just as easy. You push .NET and Windows code to AppHarbor using Git, Mercurial, Subversion or Team Foundation Server with the complimentary Git service or through integrations offered in collaboration with Bitbucket, CodePlex and GitHub. When AppHarbor receives your code it will be built by a build server. If the code compiles all unit tests contained in the compiled assemblies will be run. The result and progress of the build and unit test status can be monitored on the application dashboard. AppHarbor will call any service hooks that you add to notify you of the build result.Starting Price: $49 per month