Alternatives to Rundeck
Compare Rundeck alternatives for your business or organization using the curated list below. SourceForge ranks the best alternatives to Rundeck in 2026. Compare features, ratings, user reviews, pricing, and more from Rundeck competitors and alternatives in order to make an informed decision for your business.
-
1
JS7 JobScheduler
SOS GmbH
JS7 JobScheduler is an Open Source workload automation system designed for performance, resilience and security. It provides unlimited performance for parallel execution of jobs and workflows. JS7 offers cross-platform job execution, managed file transfer, complex no-code job dependencies and a real REST API. Platforms - Cloud scheduling from Containers for Docker®, Kubernetes®, OpenShift® etc. - True multi-platform scheduling on premises for Windows®, Linux®, AIX®, Solaris®, macOS® etc. - Hybrid use for cloud and on premises User Interface - Modern, no-code GUI for inventory management, monitoring and control with web browsers - Near real-time information brings immediate visibility of status changes and log output of jobs and workflows - Multi-client capability, role based access management High Availability - Redundancy and Resilience based on asynchronous design and autonomous Agents - Clustering for all JS7 products, automatic fail-over and manual switch-over -
2
Octopus Deploy
Octopus Deploy
Founded in 2012, Octopus Deploy enables successful deployments for over 25,000 companies around the world. Prior to Octopus Deploy, release orchestration and DevOps automation tools were clunky, limited to large enterprises and didn't deliver what they promised. Octopus Deploy was the first release automation tool to gain popular adoption by software teams, and we continue to invent new ways for Dev & Ops teams to automate releases and deliver working software to production. Runbook automation in Octopus sits side-by-side with your deployments and gives you control over your infrastructure and applications. Automate operations tasks like routine maintenance and emergency incident recovery. Flexible, role-based access control lets you manage who can deploy to production, change your deployment process, infrastructure, and more.Starting Price: Free -
3
Mint Service Desk
OPGK Software
Mint Service Desk is a comprehensive and user-friendly software solution designed to streamline and enhance the management of IT service operations within organizations. It serves as a central hub for all IT-related requests, incidents, and changes, enabling efficient communication and collaboration between IT teams and end-users. With Mint Service Desk, organizations can effortlessly track, prioritize, and resolve IT issues, ensuring minimal disruption to daily operations. The platform offers a range of powerful features, including ticket management, self-service portals, knowledge bases, asset management, and reporting capabilities. In addition to its core features, Mint Service Desk also excels in complaint management, offering robust functionality to address and resolve customer complaints efficiently. The platform understands the significance of handling complaints promptly and effectively to maintain high levels of customer satisfaction.Starting Price: $5/month/agent -
4
BigPanda
BigPanda
Aggregate data from all observability, monitoring, change and topology tools. BigPanda’s Open Box Machine Learning will correlate the data into a small number of actionable insights so incidents are detected in real-time, as they form, before they escalate into outages. Accelerate incident and outage resolution by automatically identifying the probable root cause of problems. BigPanda identifies both root cause changes and infrastructure-related root causes. Resolve incidents and outages faster. BigPanda automates and streamlines the incident response lifecycle across incident triage, ticketing, notifications, and war room creation. Accelerate remediation by integrating BigPanda with enterprise runbook automation tools. Applications and cloud services are the lifeblood of every company. When there’s an outage, everyone is impacted. BigPanda cements AIOps market leadership with $190M in funding, $1.2B valuation. -
5
PagerDuty
PagerDuty
PagerDuty, Inc. (NYSE:PD) is a leader in digital operations management. In an always-on world, organizations of all sizes trust PagerDuty to help them deliver a perfect digital experience to their customers, every time. Teams use PagerDuty to identify issues and opportunities in real time and bring together the right people to fix problems faster and prevent them in the future. PagerDuty's ecosystem of over 350+ integrations, including Slack, Zoom, ServiceNow, AWS, Microsoft Teams, Salesforce, and more, enable teams to centralize their technology stack, get a holistic view of their operations, and optimize processes within their toolsets. -
6
Ctfreak
JYP Software
Tired of maintaining your multiple crontabs? Would you like a Slack notification when one of your backups fails? CTFreak allows you to centralize and schedule various types of tasks: - Shell scripts (bash/powershell) on multiple servers via SSH - Ansible playbooks targeting multiple servers via SSH - SQL scripts on multiple databases concurrently (mysql/mariadb/postgresql) - Chart Reports from SQL queries - Webhook call - Workflow for concurrent or sequential execution of tasks Not to mention: - A mobile friendly interface - Single Sign-On via OpenID Connect - Notifications via Slack / Discord / Mattermost / Email - Issue tracking integration via Jira / Github / Linear - REST API - Incoming webhooks (Github / Gitlab / ...) - Log retrieval and consultation - User rights management by projectStarting Price: $359/year/instance -
7
ManageIQ
ManageIQ
Manage containers, virtual machines, networks, and storage from a single platform. Connect ManageIQ to your virtualization, container, network, and storage management systems, where it will discover inventory, map relationships, and listen for changes. The result is a rich, up-to-date and cross-referenced dataset that forms the basis for our advanced management capabilities. Define bundles of resources and publish them in a service catalog, from where they can be ordered by end users. Once provisioned, you can manage the full life cycle of a service, including policy, compliance, delegated operations, chargeback/showback, and retirement. Scan the contents of your VMs, hosts, and containers, and combine with auto discovery data to create advanced security and compliance policies. Content scanning works without the help of an agent and therefore works for any VM including foreign and un-cooperative ones. -
8
Jenkins
Jenkins
The leading open source automation server, Jenkins provides hundreds of plugins to support building, deploying and automating any project. As an extensible automation server, Jenkins can be used as a simple CI server or turned into the continuous delivery hub for any project. Jenkins is a self-contained Java-based program, ready to run out-of-the-box, with packages for Windows, Linux, macOS and other Unix-like operating systems. Jenkins can be easily set up and configured via its web interface, which includes on-the-fly error checks and built-in help. With hundreds of plugins in the Update Center, Jenkins integrates with practically every tool in the continuous integration and continuous delivery toolchain. Jenkins can be extended via its plugin architecture, providing nearly infinite possibilities for what Jenkins can do. Jenkins can easily distribute work across multiple machines, helping drive builds, tests and deployments across multiple platforms faster. -
9
System Frontier
Noxigen
PowerShell web front end with role based access control, auditing and remote management tools. Delegate granular permissions to manage servers, workstations, network devices and user accounts. Privileged Access Management (PAM). Let System Frontier do all the heavy lifting so you can focus on your enabling your IT teams to get more done without having more permissions than needed.Starting Price: $5 -
10
StackStorm
StackStorm
StackStorm connects all your apps, services, and workflows. From simple if/then rules to complicated workflows, StackStorm lets you automate DevOps your way. No need to change your existing processes or workflows, StackStorm connects what you already have. Community is what makes a good product great. StackStorm is used by a lot of people around the world, and you can always count on getting answers to your questions. Stackstorm can be used to automate and streamline nearly any part of your business. Here are some of the most common applications. When failures happen, StackStorm can act as Tier 1 support: It troubleshoots, fixes known problems, and escalates to humans when needed. Continuous deployment can get complex, beyond Jenkins or other specialized opinionated tools. Automate advanced CI/CD pipelines your way. ChatOps brings automation and collaboration together; transforming devops teams to get things done better, faster, and with style. -
11
Cutover
Cutover
The Cutover platform enables enterprises to simplify complexity, streamline work, and increase visibility. Cutover’s AI-powered automated runbooks connect teams, technology, and systems, increasing efficiency and reducing risk in IT disaster and cyber recovery, cloud migration, release management, and technology implementation. As a centralized system of execution, Cutover differentiates itself with scalable and proven dynamic, automated runbook technology that transforms enterprise IT operations with a new way of working. Cutover enables the creation of a template library of comprehensive, executable, and auditable runbooks covering the entire IT infrastructure. Cutover is trusted by world-leading institutions, including the three largest US banks and three of the world’s five largest investment banks. -
12
Callgoose SQIBS
ZEAZONZ TECHNOLOGIES
Callgoose SQIBS – The Future of IT Automation & Incident Management Callgoose SQIBS is a next-gen automation platform that optimizes IT operations, automates incident response, and enhances system reliability. It offers real-time alerts, on-call scheduling, incident auto-remediation, and seamless integrations to minimize downtime and improve efficiency. 🔹 Use Cases: Incident auto-remediation, on-call scheduling, process automation, IT request automation, event-driven automation, and cloud integrations. 🔹 Who Uses It? Enterprises, DevOps, MSPs, and IT teams in industries like SaaS, finance, e-commerce, telecom, and healthcare. 🔹 Key Features: Multi-channel alerts, runbook automation, no per-user fees, and full customization. 🔹 Pricing: Plans from Freemium ($0) to Dedicated ($1000/month) with automation included in every paid plan. Integrate with any ITSM, DevOps, or cloud platform. Scalable, cost-effective, and built for seamless IT automation. 🚀Starting Price: $10/month -
13
ICEFLO
Agenor Technology
ICEFLO Runbook Management (RBM) is a ServiceNow®-based platform designed to replace outdated spreadsheet runbooks with a digital solution that helps organizations manage operational resilience. It provides centralized access to runbooks, event planning, issue management, and real-time visibility into complex, multi-runbook events. -
14
FireHydrant
FireHydrant
FireHydrant is the only comprehensive incident management platform that allows you to create consistency for the entire incident response lifecycle to focus on fighting fires faster. FireHydrant is the incident management platform for businesses to manage their complex systems. Our solutions allow developers to resolve, learn, and mitigate incidents faster so they can focus on what matters most, keeping business operations running smoothly and the customers their businesses serve, happy. We're focused on building technology that thoughtfully re-engineers incident management and sets a standard for how businesses think about reliability. Our goal is to cut through manual processes and create a simple, intuitive, and best of all, delightful to use platform. Create consistency for the entire incident response lifecycle with FireHydrant, the incident management platform for teams of all sizes. Connecting integrations unlocks even more runbook automation with FireHydrant.Starting Price: $20 per user -
15
HCL HERO
HCLSoftware
Healthcheck and Runbook Optimizer that enables IT Administrator to easily monitor the health of their servers and perform informed recovery actions with specialized Runbooks. Powerful bundle offering comprising of HCL Workload Automation, HCL Clara and HCL HERO. Reduce manual labor, reduce downtime of servers, and improve IT operational efficiency across the enterprise with HCL HERO. HCL HERO effectively combines centralized application monitoring with runbook automation. It enables a single point of entry to see misconfiguration, performance or infrastructure problem on multiple environments. Users have an immediate understanding of the situation and where an action is needed with a clear and visually engaging dashboard overview. HCL HERO helps easily integrate a runbook library with customized monitors and KPIs. -
16
Runbook Studio
Kelverion
Kelverion's Runbook Studio is a graphical design application that enables organizations to harness the power of Azure Automation for developers and non-developers alike. The Studio comes packaged with integrations and solutions, making the process of creating, managing, and supporting automation runbooks accessible to all team members. It offers a drag-and-drop, code-free, graphical authoring approach, empowering users to create runbooks using a low-code/no-code capability. This approach allows users to transform manual processes into automation without the need to write any code, utilizing shapes, diagrams, and drop-down list forms. Runbook Studio provides over 800 integrations, including multi-vendor, cloud, and on-premise integrations, enabling API connections between enterprise IT systems. It also offers fully configured Runbook Solutions powered by Azure Automation for common automation use cases, ready to deploy at scale in a production environment with full logging.Starting Price: $1,095 per month -
17
HCL iAutomate
HCLSoftware
HCL iAutomate is a part of Infrastructure Automation and Orchestration offering under the HCLSoftware AI & Intelligent Operations framework. It is an Intelligent Runbook Automation product that brings Artificial Intelligence (AI) and Automation together to simplify and automate enterprise IT operation lifecycle. It leverages Machine Learning (ML) and Natural Language Processing (NLP) to comprehend issues, recommend corrective actions, and initiate automatic resolution, enabling zero-touch automation. By leveraging a repository of over 3400 configurable and reusable runbooks, it provides robust end-to-end incident remediation and task automation across the infrastructure and applications landscape. -
18
XiteiT
XiteiT
Master your cloud operation flow with a centralized platform for all production events, runbook governance, automations, operational procedures and advanced analytics. Built to improve productivity and assist every team member to achieve more. Whether you are running on-premise or cloud native, a scale-up startup or a multinational, XiteiT takes away the pain of managing the day to day complexities of your cloud operations team. A CloudOps orchestration and automation platform that integrates all of an organization’s monitoring, productivity tools and related automation platforms. Manage all your cloud operational tasks from one place to create 360o observability and operational consistency utilizing existing people and processes for a more effective incident response and production management. Drive operational visibility, so decisions are prioritized, and remediation time is dramatically reduced. -
19
Discover how to start your AIOps journey and transform your IT operations with IBM Cloud Pak for Watson AIOps. IBM Cloud Pak® for Watson AIOps is an AIOps platform that deploys advanced, explainable AI across the ITOps toolchain so you can confidently assess, diagnose and resolve incidents across mission-critical workloads. If you’re looking for IBM Netcool® Operations Insight or any previous IBM IT management offerings, IBM Cloud Pak for Watson AIOps is the evolution of your current entitlement. Correlate across all relevant data sources. Detect hidden anomalies, anticipate issues and resolve faster. Proactively avoid risks and automate runbooks for more efficient workflows. Correlate a vast amount of unstructured and structured data in real-time with AIOps tools. Keep teams focused, surfacing insights and recommendations into existing workflows. Build policy at the microservice level and automate across application components.
-
20
Shoreline
Shoreline.io
Shoreline is the Cloud Reliability platform — the only platform that lets DevOps engineers build automations in an afternoon, and fix issues forever. Shoreline reduces on-call complexity by running across clouds, Kubernetes clusters, and VMs allowing operators to manage their entire fleet as if it were a single box. Debugging and repairing issues is easy with advanced tooling for your best SREs, automated runbooks for the broader team, and a platform that makes building automations 30X faster. Shoreline does the heavy lifting, setting up monitors and building repair scripts, so that customers only need to configure them for their environment. Shoreline’s modern “Operations at the Edge” architecture runs efficient agents in the background of all monitored hosts. Agents run as a DaemonSet on Kubernetes or an installed package on VMs (apt, yum). The Shoreline backend is hosted by Shoreline in AWS, or deployed in your AWS virtual private cloud. -
21
Axcient DRaaS
Axcient
Axcient Fusion allows MSPs to consolidate and converge infrastructure and workloads in a single cloud platform. Reduce the cost, easy management, near instant recovery, and Automated Run-books. -
22
Flexible IR
Flexible IR
Planned IR skill development. Training of responders on incidents focused on domain (eg healthcare). Scenario taken from VerisDB and Flexible IR curated list. Managers can do current team evaluation and plan actions. Use of Mitre Att&ck Matrix to identify gaps that need to be practised. Evolving runbooks using Symbolic AI system integration. We provide understandable and easy baseline runbooks to handle incidents. The runbooks can be customised to your specific environment and security analyst. Expert audit of runbooks. Easily coach the less experienced members of the team in threat hunting and incident response topics. Simulate adversary use cases and practise. Plan skill development for your analysts. Move towards critical 1-10-60 rule for Incident response. Per analyst skill matrix and point systems to bring in continuous motivation and planned learning. System supports basic gamification for card based games. -
23
Doctor Droid
Doctor Droid
Doctor Droid is an AI-driven platform designed to revolutionize monitoring and troubleshooting for engineering teams. It automates complex investigations, following standard operating procedures to analyze data across multiple integrations, identify root causes, and execute standard runbooks for self-healing. By proactively listening for alerts, Doctor Droid prepares relevant data and insights, reducing on-call time by up to 80% and enabling engineers to respond swiftly. It facilitates rapid onboarding of new engineers by automating the search for documents, learning new tools, and understanding data, allowing them to become primary on-calls from day one. With the capability to perform ad-hoc investigations, such as analyzing Kubernetes clusters or checking recent deployments, Doctor Droid adapts and creates new plans based on suggestions and existing documents. It integrates seamlessly with over 40 tools across the stack.Starting Price: $99 per month -
24
iland Secure DRaaS
iland Cloud
In today’s fast-paced, global IT environment, unplanned downtime can result in irrecoverable, long-term damage to your organization. Whether from cybercrime, hardware failure, or natural disasters, the impact of a disaster event can often be felt for years in terms of revenue loss, customer churn, or the inability to continue business operations. Preparing your business for disaster events starts with combining the right people, process, and technology to ensure a quick and successful recovery. iland Secure DRaaS was designed with this in mind, providing end to end services and capabilities to meet your organization’s recovery requirements. iland Secure DRaaS with Zerto offers increased flexibility, customized runbook functionality, optimized RPOs and near-zero RTOs so you have more control over your disaster recovery plan and faster failover with automated failover and failback. -
25
ip·Solis
XenPool GmbH
ip·Solis is a self-hosted platform for IT asset lifecycle management with self-service and automation — the layer between a CMDB, IAM/HR systems, and provisioning runbooks. Employees request VDIs, shared desktops, SaaS or any IT through a self-service portal with OIDC SSO (Entra ID, Okta and more); managers approve in one click, with multi-approver quorums, conditional rules and out-of-office delegation. Chained PowerShell 7 runbooks automate provisioning against Active Directory, SCCM, XenServer and VMware. Assets are managed across their full lifecycle: pooled or personal assignment, quotas, expiry and recycle workflows. Compliance is built in — access certification (ISO 27001/SOX/PCI), append-only audit log, SCIM 2.0 and HR-webhook leaver automation. Cost reporting tracks spend by provider, consumer and department. API-first: deploys in an afternoon via Docker, on-prem or air-gapped. Free for up to 25 users; commercial license above that. Source-available.Starting Price: From €1,488/year (free up to 2 -
26
Chef
Progress Software
Chef turns infrastructure into code. With Chef, you can automate how you build, deploy, and manage your infrastructure. Your infrastructure becomes as versionable, testable, and repeatable as application code. Chef Infrastructure Management ensures configurations are applied consistently in every environment with infrastructure management automation. Chef Compliance makes it easy to maintain and enforce compliance across the enterprise. Deliver successful application outcomes consistently at scale with Chef App Delivery. Chef Desktop allows IT teams to automate the deployment, management, and ongoing compliance of IT resources. Ensure configurations are applied consistently in every environment. Powerful policy-based configuration management system software. Runbook automation to consistently define, package & deliver applications. IT automation & DevOps dashboards for operational visibility. -
27
Airplane
Airplane
Let your customer-facing teams delete accounts, change emails, issue refunds, and more. Empower your customer success team to configure accounts for new customers. Make sure you're not the only one who knows how to run that script you wrote. Make sure sensitive operations are approved by a manager or admin before being executed. Run daily reports and other periodic operations without the headache of maintaining cron or Airflow. Kick-off data backfills and other long-running tasks and get notified when they’re complete. Go beyond security checkboxes. Audit logs show who ran what so you can stop guessing and stay informed. Give teammates access upon request. Require signoff for sensitive actions. Get notifications, approve requests, and execute runbooks without leaving Slack. Go beyond security checkboxes. Audit logs show who ran what so you can stop guessing and stay informed.Starting Price: $10 per user per month -
28
ITOC360
ITOC360
ITOC360: AI-First Incident Orchestration Platform ITOC360 is an AI-First Incident Orchestration Platform that helps IT and operations teams detect, route, and resolve incidents faster, with less manual effort and fewer missed alerts. What ITOC360 Does ITOC360 centralizes alert ingestion from your entire monitoring stack and uses AI to suppress noise, correlate related events, and identify what actually needs human attention. When a real incident is detected, the platform automatically triggers the right response: notifying the right people through their preferred channels, executing runbooks, and escalating based on defined policies.Starting Price: $12/month -
29
Enov8
Enov8
End-to-end “Business Intelligence” for your IT organization. Promoting transparency, control, and productivity across environments, release and data. Promote scaled agility across your IT fabric. A complete environment and release picture supporting collaboration across teams and providing the insight that organizations require today to drive competitive innovation. Improve visibility of your complex IT fabric allowing better collaboration and decision making. Manage complex computer systems & the end-to-end IT fabric through a centralized portal. Measure test environment usage to reduce IT spend and increase project productivity. Eliminate chaotic and non-repeatable operations by establishing control via centralized runbooks and using automation on regular & time consuming tasks. Manage change and contention effectively whilst providing real time health status and powerful analytics to determine business impact.Starting Price: $8 per month -
30
Runable
Runable
Runable is an AI-automation/agent platform that lets users automate almost any digital task a human could do on a computer, using natural language instead of scripting. It supports browser, desktop, and mobile interfaces, offers connectors/integrations to common services, and allows scheduling and workflow orchestration (“runbooks”) for repetitive or multi-step tasks. Runable provides a library of example runbooks/templates (for marketing, sales, programming, research, productivity, etc.) so users can start from prebuilt automations and customize them. Use cases include things like automatically preparing meeting materials by researching companies on your calendar, generating reports with visualization, updating docs, and organizing files or data. The system includes feedback loops (you can adjust, push forward, schedule runs), permissions/connectors, and is positioned to help reduce manual work, streamline repetitive workflows, and scale productivity. -
31
Resolve
Resolve Systems
Resolve is the #1 IT automation and orchestration platform, powering more than a million automations every day from simple, high-volume tasks to incredibly complex processes that go well beyond what you imagine is automatable. With more than a decade of automation expertise under our belts, we know how to build an intelligent automation and orchestration platform to meet the growing demands faced by today’s IT Operations and Network Operations teams. In fact, millions of automations are powered by Resolve on a daily basis… many of which go well beyond what you imagine is automatable. We know it sounds impossible, but it’s true. Just ask the customers who have cracked the code on tough automations like PIM testing, updating active load balancers, CUCM onboarding in seconds, true end-to-end patch management, interacting with Watson for NLP, maintaining infrastructure in segregated networks and hybrid cloud deployments, and more. Keep reading to see how we do it. -
32
eemaan Deployment Manager
eemaan
Package and deploy software & configuration updates in seconds. Follow a 5-step wizard to package Genesys software and configuration into a portable package ready to be shared with colleagues, all from the comfort of a powerful dashboard. Deploy any shared package in a few clicks. Select the location, the package, the Genesys Application you want to update, optionally customize the deployment, and just click 'Go'. The whole process of downloading software, and updating the Genesys configuration is carried out automatically. The deployment didn't go to plan? Not to worry, just one click, and the old software and configuration are restored. The best is always saved for last. The deployment process comes with an automatic Runbook generator. In the blink of an eye, a step-by-step runbook is generated for the approval process, and for that, just in case something goes the wrong backup plan. -
33
Small Hours
Small Hours
Small Hours is an AI-powered observability platform that helps root cause server exceptions, analyze the impact, and triage to the right person or team. Use Markdown or your existing runbook to guide our assistant in debugging issues. We support OpenTelemetry for seamless integration with any stack. Hook into existing alarms and identify critical issues. Connect your codebases and runbooks as context and instructions. Your code and data are secure and never stored. Intelligently triage issues and generate pull requests. Optimized for enterprise velocity and scale. 24/7 automated root cause analysis, minimize downtime, and maximize efficiency. -
34
Our software transforms your organization into a high velocity, service-centric environment through digitalization, automation and orchestration. Classical approaches to automation focus on automating the actions, but typically this only account for 10% of your teams’ time and effort. The other 90% (the analysis, the decisions, and the management) currently lock your people into these processes. Your subject matter experts teach Cortex to intelligently orchestrate successful outcomes of automated processes. Cortex will continue to collaborate with your experts through powerful exception management, ensuring non-stop alignment with business operations. Truly self-managing, self-healing, automated operations and digitalized services, which manage all exceptions, failures and problems that may occur.
-
35
Rootly
Rootly
Rootly is an AI-native incident management platform built to help modern teams prevent and resolve incidents faster. It streamlines on-call scheduling, incident response, retrospectives, and status updates through intelligent automation and deep integrations with Slack, Teams, Jira, and Zoom. Powered by Rootly AI, the system automates root cause analysis, provides suggested fixes, and compiles incident data into clear summaries for faster recovery. Teams can manage incidents directly within their communication tools, reducing context switching and human error. With automated retrospectives and actionable insights, Rootly enables continuous improvement and reliability across engineering organizations. Trusted by global brands like Figma, Canva, Nvidia, and Webflow, it helps companies maintain uptime, minimize disruption, and create a culture of proactive resilience. -
36
Control-M
BMC Software
Control-M accelerates business outcomes by orchestrating complex workflows across hybrid, cloud, and on-prem environments. With centralized visibility and governance, it simplifies operational complexity, reduces risk, and increases agility. The platform’s predictive analytics, event-driven automation, and seamless integration with modern toolchains enable IT, data, and DevOps teams to automate data pipelines and application workflows reliably at scale. Backed by BMC’s expertise in intelligent automation, Control-M connects people, systems, and data to power sustainable growth and transformation. Through predictive analytics, event-driven automation, and seamless integration with modern toolchains, Control-M helps organizations reduce risk, increase agility, and accelerate innovation. Widely used in finance, healthcare, telecom, and manufacturing, it is valued for scalability, compliance, and operational resilience in mission-critical environments.Starting Price: $2400 / month -
37
NudgeBee
NudgeBee
NudgeBee is an AI Agents and Agentic Workflow platform built for SRE, CloudOps, and DevOps teams. It combines pre-built AI Assistants for incident troubleshooting, cloud cost optimization, and Kubernetes operations with a visual no-code Workflow Builder for custom automation. NudgeBee's AI engine auto-investigates alerts using a live semantic Knowledge Graph, grounded in your actual infrastructure topology. It queries data in place from existing tools (Prometheus, Datadog, Grafana, Loki) with zero data ingestion. The Workflow Builder supports 20+ action categories, native AWS/Azure/GCP CLI nodes, A2A and MCP protocol support, and human-in-the-loop approval gates. 49+ integrations. Enterprise-ready with RBAC, audit trails, BYOM (Bring Your Own Model), and self-hosted deployment. SOC-2 Type II and ISO 27001 compliant. -
38
Squadcast
Squadcast
Squadcast is an incident management tool that’s purpose-built for SRE. Create a blameless culture by reducing the need for physical war rooms, centralize SLO dashboards, unify internal and external SLIs and automate incident resolution and knowledge base creation with Squadcast Actions. Adopt world-class site reliability practices with a centralized SLO dashboard to view your system health. Anticipate incidents before they occur and respond proactively. The first step towards doing better incident management is adding enough context to incidents while they get detected. With Squadcast, discover everything you need, to take action and achieve best-in-class MTTD with highly configurable features like alert deduplication and tagging.Starting Price: Free -
39
Kelverion Automation Portal
Kelverion
Kelverion's Automation Portal is a lightweight, self-service interface designed to simplify IT process automation by enabling end users and IT teams to trigger, track, and manage automated tasks across various platforms. It offers a forms-driven, intuitive interface that integrates seamlessly with automation tools like Azure Automation, Power Automate, Logic Apps, and System Center Orchestrator, as well as third-party systems via a full REST API. The portal supports both on-premise and cloud-hosted deployments and can be hosted as an IIS web application. Authentication is handled through Microsoft Entra ID, ensuring enterprise-grade user security. Key features include a live dashboard displaying time and cost savings from automation, request statuses, and top requests; support for high availability via Windows Network Load Balancing (NLB), allowing users to submit and manage IT requests on the go. -
40
Tidal by Redwood
Redwood Software
The highly-scalable, highly-resilient Tidal Automation platform keeps your entire automation initiative on course, whether you’re automating foundational systems like ERP or orchestrating complex new opportunities in Big Data, IoT, AI, and more. It’s all about leveraging automation to help the enterprise meet its mission. Tidal by Redwood is an easy-to-deploy, easy-to-use, scalable solution that provides a centralized, enterprise-wide interface for planning and controlling execution of business processes, applications, data, middleware, and infrastructure. -
41
StackPilot
StackPilot
StackPilot is an AI-powered oncall copilot that automates root cause analysis and bug fixes for software engineers. It integrates directly with observability tools like Datadog, Sentry, and PagerDuty to transform alerts into actionable fixes. The platform analyzes recent commits, logs, and stack traces to pinpoint faulty code, then generates pull requests with proposed solutions. Engineers only need to review and merge, significantly cutting resolution time from hours to an average of 15 minutes. StackPilot also captures investigative steps and converts them into reusable runbooks, improving incident response over time. With strong privacy measures—no code or logs stored—it ensures secure, real-time analysis for engineering teams.Starting Price: Free -
42
Resolve AI
Resolve.ai
Operates autonomously to handle common alerts and actions, reducing escalations and preventing burnout. Dynamically adjusts thresholds and dashboards to proactively prevent incidents and adjusts runbooks with every new incident. Saves up to 20 hours per on-call engineer per week so you can get back to the building. Handles all alerts, performs root cause analysis, resolves incidents, and makes on-call stress-free. Automates root cause analysis and incident response, cutting Mean Time to Resolution (MTTR) by up to 80%. With detailed incident summaries and hypotheses available, before you log in, you'll experience faster response and significantly increased uptime. Get started in minutes with production-ready AI, which is secure and knows how to use all the production tools like an experienced software engineer. It automatically maps your production system, understands code, and captures changes without any training. -
43
Mission Cloud Secure
Mission
Mission Cloud Secure is a SaaS application that delivers 24/7 security monitoring and incident response through a powerful combination of CrowdStrike's world-class security platform and Mission's AWS expertise. Protect your cloud resources, endpoints, and credentials while maintaining compliance and operational excellence. Mission Cloud’s team of CloudOps Engineers works directly with the CrowdStrike SOC to give you 24/7 managed detection and response. We alert you to incidents and help the SOC to respond with the runbooks we’ve co-developed. CrowdStrike’s analysts also operate a continuous threat detection engine and partner with other security experts from the public and private sectors to proactively protect your environment and manage threats. In today's landscape of sophisticated cyber threats, comprehensive security requires constant vigilance, specialized expertise, and the right tooling. Never worry about when or how a security incident occurs. -
44
Runbook
Runbook
Runbook is an autonomous operations platform that turns how an operations team works into AI agents that run jobs end-to-end inside the systems the business already uses. Agents work across phone, email, SMS, Teams, Slack, web chat, ERP, CRM, spreadsheets, vendor portals, and legacy systems, including tools without APIs through browser automation. Teams describe a workflow in plain English and provide SOPs, training documents, call recordings, carrier rules, and work history so the agent learns how the operation actually runs, including edge cases and tribal knowledge. Rather than acting as a copilot, the agent owns the complete workflow, from the first inbound request through system updates and closure. It can automate appointment scheduling, accounts payable and receivable, order management, status updates, vendor coordination, calls, follow-ups, and system entries. -
45
Digitate ignio
Digitate
Transform your operations across domains using AI and Automation towards an Autonomous Enterprise for improved resilience, assurance, and superior customer experience. Digitate’s ignio helps resolve your operational woes for an Agile, Resilient and Autonomous Enterprise. Businesses can adapt to changes efficiently, evolve digitally and unleash innovation to sustain and grow. With ignio, transform your IT and business operations’ from reactive to proactive, and take a leap forward to ‘Predict, Prescribe and Prevent.’ Learn how enterprises can elevate their business and IT operation strategy to make headway into an Autonomous Enterprise. Get started on your journey from Traditional to Automated to Autonomous Operations. Powered by AI and Machine Learning, Autonomous Operations allows enterprises to reduce manual efforts, adapt to business or IT changes efficiently with minimal cost and focus on innovation. -
46
Store, optimize and protect your critical data and apps with managed hosting at a world-class global data center or your location. Managed hosting services creates a scalable, hands-free foundation that’s ready to grow with your evolving enterprise. Your mission-critical apps and data are optimized for performance and reliability and actively protected by built-in monitoring to ensure security and continuity. Get the hardware, power and bandwidth you need today and add additional services as your business evolves. Gain the peace of mind provided by redundant systems, multiple levels of security and 24/7 monitoring. Eliminate capital and operating expenses associated with maintaining and staffing a dedicated data center. Dedicated customer service manager available. 24/7 support and first-touch response. Runbook-based approach. Trend analysis and engineer consulting and support.
-
47
Hyground
Hyground
Hyground is an AI-powered DevOps and SRE co-pilot — not a chatbot wrapper, but a full-stack operational intelligence system that runs inside the customer's Kubernetes cluster with no data egress. The agent connects to 21+ enterprise systems and investigates incidents across logs, metrics, traces, and K8s events. Engineers ask questions in plain language and get answers grounded in their own data — no new query languages to learn. AutoRCA turns an alert webhook into an autonomous root-cause investigation, then posts findings back to Slack or Teams. Investigation starts the instant an alert fires, not when an engineer wakes up. Customers report up to 85% MTTR reduction. Built on Google's Agent Development Kit, Hyground uses a multi-agent architecture and learns from your infrastructure over time. Resolved incidents extend the knowledge base, so runbooks stay current. -
48
Nuphos
Nuphos
Nuphos is an AI-native DevOps workspace where engineering teams and AI agents operate production systems together without giving up control. Agents learn your infrastructure, investigate issues, and work across AWS, GCP, Kubernetes, Cloudflare, and the rest of your stack while using fine-grained IAM permissions, human approvals, and full audit trails. Each agent session can be scoped with native IAM roles and short-lived, least-privilege credentials, and anything that changes infrastructure is first proposed as a plan for approval. Agents can inspect resources, open dashboards, read logs, generate plans, request approval, and take safe actions while building memory of services, environments, workflows, runbooks, and operational history. Engineers and agents share the same DevOps workspace instead of jumping between terminals, cloud consoles, dashboards, and documentation.Starting Price: $29 per month -
49
Dock
Dock
Dock is the AI workspace for you, your team, and every agent you run. It gives humans and AI agents the same shared cloud workspace, where everyone can read and write the same state in real time instead of working across scattered chats, files, and one-off outputs. Dock is built around tables with typed columns, rich-text docs, and agents as first-class identities, each with their own API keys, permissions, and audit trail rather than delegated human tokens. Teams can use Dock to plan, research, decide, and ship with humans and AI on the same surface, with use cases across engineering, go-to-market, research, operations, solo work, and agency workflows. Engineering teams can manage sprint planning, spec docs, and incident response; GTM teams can organize content calendars, sales pipelines, and customer success; research teams can track interviews, themes, and competitive intelligence; and operations teams can manage runbooks, recruiting, compliance, and onboarding.Starting Price: $19 per month -
50
Automic Automation
Broadcom
Enterprises need to automate a complex and diverse landscape of applications, platforms and technologies to deliver services in a competitive digital business environment. Service Orchestration and Automation Platforms are essential scale your IT operations and derive greater value from automation: You have to manage complex workflows across platforms, ERP systems, business apps from mainframe to microservices and multi-cloud. You need to streamline your big data pipelines, enabling self-services for data scientists while providing massive scale and strong governance on data flows. You're required to deliver compute, network and storage resources on-prem and in the cloud for development and business users. Automic Automation gives you the agility, speed and reliability required for effective digital business automation. From a single unified platform, Automic centrally provides the orchestration and automation capabilities needed accelerate your digital transformation.