Roadmap
Erfahren Sie, woran wir aktuell arbeiten, um Checkmk weiter zu verbessern.
Fehlt Ihnen eine Funktion? Dann fragen Sie es auf unserem Ideas Portal an.
-
Set-up cloud monitoring in minutes
- Configure the entire monitoring of AWS, Azure and GCP in one go
- Immediate feedback during configuration on errors by testing the connection during the configuration
Done -
You'll love notification with Checkmk!
- Central notifications hub to have all information and relevant rules in one place
- Guided configuration of notifications
- Overview mode for notification configuration
- On-the-fly creation and editing of notification parameters
- New notification emails
- Updated notification plug-ins for Microsoft Teams, Jira, OpsGenie, ServiceNow, Slack
Done -
OpenTelemetry & Prometheus exporter integrations
- Monitor custom applications by
- pushing OpenTelemetry protocol (OTLP) metrics to Checkmk sites
- pulling Prometheus metrics from exporters or natively directly from Checkmk sites
- Flexible architecture accommodating almost any use case with cascading OpenTelemetry collectors
Done - Monitor custom applications by
-
Better handling of large highly dynamic environments
- Re-architecture of dynamic host management into a queue-based system
- Parallel processing of connectors for high-speed data ingestion
- Stable & safe processing of configuration changes and activation
Done - Re-architecture of dynamic host management into a queue-based system
-
Enhanced distributed monitoring
- New inter-site communication bus also allowing remote-remote communication
- Distribution of piggyback data across sites
Done -
Automated in-depth Synthetic Monitoring
- Managed robots
- Linux support
- Building test environment on offline systems
- KPI monitoring
Done -
Non-root agent
- Deployment of all agent files within one customizable folder
- Agent can be run as unprivileged user
Done -
Improve platform & security
- Improve REST API performance
- Replacement of automation user in internal features
- Improve runtime stability of Checkmk with config validation layer
- Checkmk agent can run on FIPS-enabled OS
- Checkmk site can run on FIPS-enabled OS
- Certificate overview page
- Enforce Two-Factor Authentication for users
- Software architecture & build system rework
- Distributed tracing of Checkmk itself for performance analysis using OpenTelemetry & Jaeger
Done -
And many more smaller improvements
- Update service discovery parameters
- Faster loading of service discovery page, effective parameters of host/service, computation of effective labels, and likely much more across entire Checkmk
- Many plug-ins incl. mainlining of Redfish and improvements to mk_sql, check_cert and check_httpv2
Done
-
AI
Root Cause Analysis
- AI chat assistant embedded in Checkmk: Ask about multiple hosts and services via chat conversation, and iterate with follow-up questions to reveal root causes faster.
- Incident clustering: Collapse alert storms into a concise, ranked, clustered list of incidents.
MCP server
- AI agent access: Bring your own Model (BYOM) to connect Checkmk’s live system data to AI tools you already use. Troubleshoot issues quickly with natural language queries.
- Site-level investigation: Save time, effort, and reduce MTTR by letting your AI agent run diagnostic workflows across your entire infrastructure stack.
- Secure, efficient data processing: MCP server reuses Checkmk’s own logic, and inherits native authentication, authorization, and RBAC.
In progress -
Advancing Full-Stack Observability
Introducing Network Flow Monitoring (3.0)
- Real-time and historical netflow data: Analyze and visualize inbound and outbound network flow data, with support for all major flow formats NetFlow (v5, v9), sFlow, IPFIX.
- Root-cause analysis: Follow a journey from an incident on a host or interface straight into the flows behind it.
- Built-in dashboards: Reduce manual investigation with Flow overview, Traffic, Autonomous systems, traffic analysis and Flow explorer dashboards. Drill down into hosts and autonomous systems to pinpoint the source of sudden traffic and bandwidth spikes.
Application Observability
- Quick OTel setup: Onboard and configure OpenTelemetry data in minutes using a guided, step-by-step flow.
- API automation: Configure OpenTelemetry data collection automatically via API, using Infrastructure as Code tools like Terraform and Ansible.
- Guided Service-to-Alert Flow: Custom service creation leads straight into alert configuration, with thresholds set directly on the metric graph so impact is visible before saving.
- Intuitive metrics filtering: Use regex to find and isolate which metrics are ingested, without needing to know exact metric names in advance.
- Customizable visualization: Dig deeper into the data, and fine-tune how complex insights are presented with flexible display options.
- Time-series aggregations: Group or slice data to evaluate application health, from high-level summaries to granular time slices.
Monitor Anything in the Cloud
- Extensible Azure monitoring: Extend the out-of-the-box monitoring to cover any Azure resource using the native OTel integration (not included in Pro).
- Improved out-of-the-box monitoring for AWS: Connect your entire AWS organization with a simplified setup that scales, and manage and visualize hosts in a hierarchy that natively maps to your AWS account structure.
Kubernetes Monitoring
- Quick setup: Monitor your clusters quickly and easily with Kubernetes guided setup.
- Push or pull agents: Monitor securely in push mode without exposing your API, or pull data directly with our new agent.
- Auto-scalable workloads: Monitor large, dynamic clusters without Checkmk configuration churn, even as pods come and go.
- Native OTel support: Gain additional insight into your cluster's performance by drilling down to container-level metrics (not included in Pro).
Synthetic Monitoring
- Built-in dashboards: visualize individual test executions over time, across locations, environments, and more in the new Pulse View.
In progress -
Reimagining Analytics, Visualization & UX
Views 3.0
- Responsive search: See search results in real-time, as you type.
- Slide-out panels: Click individual hosts to display their service overview, parameters, event history (and more).
- In-View Actions: Execute problem acknowledgments and schedule downtimes right from the slide-out panel.
Intuitive Graphing
- Global time-picker: Change the time range once and every dashboard updates together.
- Frictionless navigation: Pinch-to-zoom, inspect mode, and smarter pinning make insights easier to find.
Connected, Drill down Dashboards
- Build your own: Build custom drill down links between dashboards.
- Responsive display: Pan, zoom, brush, and select custom ranges.
Modern NagVis Successor (Orbvis)
- Native maps integration: Map with a modern, open-source engine that replaces the 20-year-old NagVis codebase.
Foundations
- Linux agent: Unify the updater and controller in a clean, default one-directory deployment.
- Enterprise identity management: Authenticate with SAML across distributed sites, with unified user identities.
- Oracle monitoring: proper multi-tenant DB support.
- REST API version tracking: Check Veeam REST and Nutanix v4 APIs directly.
In progress -
Beyond 3.0
- Behavioral threat detection: Flag unusual activity like sudden traffic surges or new destinations, with a built-in alert library that includes security checks.
- Modernized Views: Visualize host relationships and impact paths.
- Log management: Search all your logs in one place, natively connected to AI that diagnoses and remediates issues.
Network Flow Monitoring expansion (3.1)
- Security threat and anomaly detection: Behavioral checks automatically detect threshold breaches on volume, unusual traffic patterns, and known bad protocols.
- Probe mode with deep packet inspection: Analyze traffic inside internal networks, and recognize 300+ apps and protocols (Teams, YouTube, etc.).
- NetFlow in distributed setups: Map flow data from multi-tenancy setups to a central Checkmk site.
- Cloud flow log ingestion: Natively ingest AWS, Azure and GCP flow logs.
Planned
-
Enhanced Infrastructure Monitoring
- Improved and expanded monitoring of Virtualisation Platforms, including Proxmox, Hyper-V and Podman. Expect a more unified experience.
- Significant expansion and improvements to plugins for visibility into advanced Azure Cloud services.
Done -
Modern Application Monitoring
- Effortlessly analyze & visualize for application data.
- A scalable and robust data backend to store and query modern application metrics.
- Integrate OpenTelemetry Protocol (OTLP) or Prometheus metrics with ease.
- Smarter host name handling: Automatic and consistent host name computation.
- More flexible and secure ways to authenticate data sources.
- Better Windows logwatch monitoring.
- Oracle Monitoring: deeper database insights.
Done -
Powerful Dashboarding
- Easily create, test and edit powerful and visually appealing dashboards.
- A smarter and more guided process for dashboard creation.
- Better filtering: set and apply context search filters, and time ranges.
- More flexibility in editing and reshaping dashboards.
- Responsive dashboards: view dashboards on any screen.
Done -
Navigation and onboarding
- Frictionless first steps for faster onboarding.
- More meaningful status updates about what is happening.
- In-page validation – fix issues before they become a problem.
- Less context switching.
Done -
User experience improvements
- Activate changes on the fly without context-switching.
- Further performance improvements.
- A more interactive UI.
- Dynamic pages.
- A faster user interface.
Done -
Monitor segmented networks with Checkmk Relay
- Checkmk Relay. A lightweight, hands-off and fully centrally manageable way to monitor segmented networks.
- Deploy Checkmk Relay for more centrally manageable distributed monitoring setups.
- Securely connect and exchange data with a site.
Done
Checkmk Cloud Roadmap
-
Host Checkmk sites in a US region
- Giving users the possibility to host their Checkmk site in the US or in the EU. All site specific data will remain in that region
Done -
SSO via SAML
- Support SSO via Microsoft EntraID (formerly known as Azure AD)
In progress -
Kubernetes Monitoring
- Allow monitoring of managed or unmanaged Kubernetes clusters, e.g. EKS, GKE, AKS or Redhat OpenShift.
Done -
Network Device Monitoring & monitor systems not accessible from the internet
- Allow monitoring of devices and environments that can only be monitored via SNMP or are not reachable from the internet
- Allow monitoring of systems that are not accessible from outside your network and are being monitored via a special agent (e.g. virtualization platforms like vSphere, Proxmox, storage solutions)
- Focus on a solution for “smaller environments” first
In progress -
Monthly subscriptions
- Flexibly use Checkmk Cloud without the commitment of an annual contract.
- Scale your subscription only if and when your monitoring needs grow
Done