Open-source infrastructure monitoring

The monitoring stackyou own outright.

Uptime, pagespeed, hardware and incident alerts in one self-hosted dashboard. No per-monitor pricing. And you find out before your users do.

open source, no asterisks

Demo credentials: demouser@demo.com / Demouser1!

Monitor types
HTTP · Ping · TCP · gRPC · DNS · SSL
Integrations
Slack · Telegram · Teams · PagerDuty
Deploy
Docker · Helm · Bare metal
The problem

Monitoring shouldn't costa fortune.

Enterprise tools meter every host. Free tools hit a wall fast. Checkmate sits in the middle: self-hosted, fully featured and not trying to upsell you at 2am.

Per-host pricing adds up

Datadog and PagerDuty charge per host, per seat and per feature. Checkmate has one tier: all of it, free.

Your metrics, their servers

SaaS monitoring ships telemetry off your network. Self-host Checkmate and nothing leaves your perimeter. There's no third party to trust, because there's no third party.

Stacks that take a week

Prometheus, Grafana, Alertmanager, exporters. Checkmate gets you to a first alert in about 5 minutes.

What you monitor

10 monitor types,one dashboard.

An endpoint can return 200 while the cert expires, the disk fills and a container crash-loops. A green dot can't see any of that. These can.

HTTPStatus codes, response times and body matching
PingICMP reachability and latency
TCPIs the port accepting connections?
gRPCStandard gRPC health protocol checks
WebSocketConnections accepted and responsive
DNSRecords resolve where they should
PageSpeedLighthouse scores and Core Web Vitals
InfraCPU, memory and disk via the Capture agent
DockerContainer state, checked like any monitor
Game100-plus game types via native queries
Features

Everything you need,under one roof.

Uptime, pagespeed, servers, containers and alerts behind a single login. Outages love the gaps between tools, so we didn't leave any.

Global uptime monitoring

Monitor HTTP, ping and TCP endpoints from 6 continents with GlobalPing. Know when things go down, anywhere in the world.

North AmericaEuropeAsiaOceaniaSouth AmericaAfrica

Server monitoring

Track CPU, memory, disk and network usage across all your servers.

CPU
24%
Memory
67%
Disk
45%

Docker monitoring

Know the moment a container stops, with health and resource views on the way.

nginx
postgres
redis
api

Multi-channel alerts

Get notified on 12 channels, from email to PagerDuty. You hear about it before your users do.

EmailSlackDiscordTelegramTeamsTwilio SMSPushoverntfyMatrixRocket.ChatPagerDutyWebhooks

Public status pages

Beautiful, customizable status pages to keep your users informed during incidents.

APIWebDB

Game server monitoring

Monitor Minecraft, CS2, Valheim and more than 100 game types.

Response time tracking

Track response times and latency with historical charts and performance insights.

Team collaboration

Invite your team with role-based access control. Everyone stays in the loop.

AdminEditorViewerInvite
Inside the product

The dashboard,as it ships.

These are the four views you'll spend the most time in. Each one loads fast, works from the keyboard and shows you what changed without making you dig for it.

checkmate.soUp
https://checkmate.so · checked every 60 s
Uptime / 30 days
99.98%
2 incidents
Avg response
142 ms
p95 380 ms
Cert expiry
84 days
Mar 12, 2027
Domain expiry
213 days
Jul 19, 2027
Response times
DayWeekMonth
00:0006:0012:0018:00Now
Recent checksFrankfurt · London · Virginia
12:04:32200 OK138 msFrankfurt
12:03:31200 OK129 msVirginia
12:02:32200 OK151 msLondon
11:56:12504 Gateway timeout30,000 msFrankfurt
Uptime

Know the moment something breaks.

Per-endpoint uptime, incident counts and response time history. Chart views span a day, week, or month so patterns show up before they become outages.

  • Real-time uptime percentage and incident tracking
  • Response time charts with day, week and month views
  • SSL certificate expiry monitoring
PageSpeed

Lighthouse scores, tracked over time.

Performance, accessibility, SEO, best practices, plus Core Web Vitals. Every run recorded so regressions don't ship silently.

  • Lighthouse performance, accessibility and SEO scores
  • Core Web Vitals: LCP, FCP, CLS and more
  • Score history to track improvements over time
checkmate.soDesktop
Lighthouse · last run 6 min ago
Performance
Accessibility
Best practices
SEO
Core Web Vitals
LCP1.2 s
FCP0.9 s
CLS0.02
TBT210 ms
Performance score · last 30 runsDaily · desktop
30 runs agodip: image regression, fixed same dayToday: 98
prod-web-01Online
Capture agent · reporting every 15 s
CPU
Memory
Disk
Temp
Network throughput
InOut
Per-core usage
C1C2C3C4C5C6C7C8
Host
Uptime42 days
Load avg0.61 · 0.54 · 0.48
Swap0.2 / 4 GB
Infrastructure

Server health, down to the core.

CPU, memory, disk and network across every box, reported by the Capture agent. Live gauges for right-now, history for capacity planning.

  • CPU usage, temperature and core frequency
  • Memory and disk usage with visual gauges
  • Historical charts for capacity planning
Status pages

A status page users trust.

Public dashboard with per-service uptime bars and your own branding. One shareable URL for customers, on-call and stakeholders.

  • Public-facing status dashboard with custom branding
  • Per-service uptime bars with historical data
  • Shareable link for customers and stakeholders
Acme status
status.acme.com
All systems operationalUpdated 12:01 UTC
API100.00% uptime
Web app99.98% uptime
Database99.95% uptime
CDN100.00% uptime
45 days agoToday
Past incidents
ResolvedDatabase latency above thresholdAug 12 · lasted 22 min
ResolvedElevated error rate after deployJul 30 · lasted 4 min
Incidents1 active
Last 30 days
Open now
1
api.checkmate.so
Resolved / 30d
4
across 3 monitors
Avg time to resolve
0.4 h
last 30 days
Most incidents
2
api.checkmate.so
Resolution detailsdb-primary · Aug 12
ManualResolved by ops@checkmate.so
"Failed over to the replica, root cause was a full disk. Cleaned up and re-synced."
Timeline
Activeapi.checkmate.so · HTTP 50412 min
HTTP/1.1 504 Gateway Timeout · latest failure kept on the incident
Started at 12:04 · ongoing for 12 min · alerts went out on this monitor's channels
Resolveddb-primary · connection refusedAug 12 · 22 min
Resolvedcheckout.acme.com · timeoutAug 2 · 4 min
Resolvedcdn-edge · SSL handshake failedJul 30 · 9 min
Incidents

Every outage, on the record.

Checkmate opens an incident on the first failed check, tracks how long it lasts and keeps the timeline after recovery. The history is your SLA math and your postmortem notes in one place.

  • Incidents created automatically from failed checks
  • Timelines with start, duration and resolution
  • Filter by monitor, date range or status
Maintenance

Planned downtime, without the noise.

Schedule one-off or recurring windows and Checkmate holds the alerts while you work. Checks keep running and get recorded; nobody gets paged for the reboot you planned.

  • One-off and recurring schedules
  • Alerts suppressed, checks still recorded
  • Windows scoped to the monitors you pick
Maintenance windows1 running
Alerts hold while a window runs
DB failover drill · running18 min left · 3 monitors muted
db-primarydb-replicapgbouncer
Next 7 days3 windows scheduled
Mon
Tue
Wed
Thu
Fri
Sat
Sun
ScheduledTimes in UTC
Weekly patch windowRecurringSun 02:0045 min
Cert rotationOne-offMar 12, 22:0030 min
Storage migrationOne-offMar 18, 01:002 h
Router firmwareRecurringFirst Tue 03:0020 min
Works with

Plugs into the toolsyou already run.

Email
Slack
Discord
Telegram
Twilio SMS
Pushover
Matrix
ntfy
Rocket.Chat
PagerDuty
Microsoft Teams
Webhooks
Use cases

Who runsCheckmate?

Homelabs

A Pi and a compose file watching the NAS, the reverse proxy and 40 containers.

SaaS founders

One founder, one server and a status page that answers "is it down?" before support has to.

Agencies

Every client site in one dashboard, with a branded status page for each.

Sysadmins and DevOps

TCP, DNS, ping and hardware metrics without standing up a Prometheus stack.

E-commerce

Checkout uptime during the sale, certificate expiry visible long before it matters.

Open-source maintainers

Free monitoring for project infra, run on the project's own box.

Deployment

Your servers,your rules.

one docker compose up
Install

Docker Compose

A single compose file. Up and running on your box in under 5 minutes.

Scale

Bare metal or k8s

Drop the containers on a VM, a Pi, or a Helm chart. No artificial monitor limits.

Own

AGPL-3.0

Source is yours to read, fork and audit. Data never leaves your perimeter.

Open source

Built in the open,for everyone.

AGPL-3.0, no vendor lock-in, no paywalled features. Read the code, file an issue, or send a PR. 130-plus contributors already have. It has run in homelabs and prod racks since 2024.

GitHub stars
10,774
Forks
1,194
Contributors
158
License
AGPL-3.0
Contributors

Shipped byreal humans.

FAQ

Frequently askedquestions.

A Linux box (or any Docker host) with about 1 GB of RAM and a few GB of disk. Checkmate ships as a small set of containers: pull the compose file, run it, done.

AGPL-3.0. You can run it for free, fork it and modify it. If you expose a modified version as a network service, you need to share your changes under the same license.

No. Everything (metrics, logs, incidents, user accounts) lives in the database you point Checkmate at. No phone-home, no telemetry pipeline.

A ping tells you the server answered. It can't tell you the cert is valid next week, DNS still points where it should, the disk isn't filling or a container isn't restart-looping. That's why there are 10 monitor types instead of one.

Commercial tools meter by host or monitor. Checkmate doesn't. You give up the managed cloud and some polish; you get unlimited monitors, your own data and a smaller bill.

There's no one-click import, but most existing monitors map cleanly: HTTP, ping, TCP, Docker, status pages. For bulk setup the API accepts JSON, so a short script usually does it.

Discord is the fastest path; maintainers and contributors hang out there. GitHub issues work for bugs and feature requests. There is no paid support tier.

Get started

Stop worryingabout downtime.

A single compose file, a few minutes and the next incident reaches you first, not your users. Free forever, self-hosted.