Features
Everything you need to monitor your fleet
Real-time metrics, container control, uptime checks, alerts, groups, and teams - all in one dashboard.
Core monitoring
Live metrics from every machine - CPU, GPU, RAM, disk, network, and load refresh every second. History runs from 5 minutes to 365 days with average and P95 series, a live mode, and CSV export on every chart.
- CPU per core with load 1 / 5 / 15 and frequency; GPU for NVIDIA, AMD, Intel, and Apple silicon
- Online / Connected / Offline states with last-seen timestamps
- History windows from 5m to 365d with average and P95, plus live mode
- CSV export on every chart; Free keeps 14 days at 60 s, Pro keeps 180 days at 1 s
CPU
42%
GPU
18%
RAM
68%
Disk
31%
Machines
Rows, grid, or cards - drag to reorder, nest into groups, and drill into every machine with six dedicated tabs.
- Per-core CPU, memory pressure and swap, disk per partition with read / write rates, ZFS pools
- Hardware inventory on the Device tab: P / E cores, GPU VRAM, DIMM type and speed, disks, NICs with link speed
- Per-machine tabs: Dashboard · History · Containers · Control · Device · Management
- Inline rename, group assignment, unregister, forget history, delete
prod-server-01
Ubuntu 22.04
Mac Mini M2
macOS 15.3
dev-workstation
Windows 11
backup-node
Debian 12
Docker & containers
See every container's CPU, memory, health, and status across every host, refreshed every 3 seconds. Start, stop, restart, pause, pull images, remove, and stream logs - from any device.
- Containers grouped by machine, nested under group hierarchy
- Start / Stop / Restart / Pause / Unpause / Pull / Remove per container
- Details popup with state, image, ports, health, and IO rates
- Live logs viewer - no SSH or terminal needed
Proxmox guests
Run statsd on a Proxmox node and its VMs and LXC containers show up beside the host. Auto-detected - no agent inside each guest.
- VMs and LXC containers refreshed every 5 seconds
- Detected automatically on the node - nothing to configure
- Guests listed under their host in the same dashboard
Hosts
Watch any endpoint over HTTP, ICMP ping, or a MongoDB connection - public sites, internal services, third-party APIs. Track response times, status codes, and uptime.
- HTTP, ICMP, and MongoDB checks, as often as every 10 seconds
- Custom method, headers, expected status codes, body match, and timeout
- Success and failure thresholds before a host changes state
- Uptime bars for 24h / 7d / 30d / 90d, response-time and connectivity charts
Hosts
5 watched · 4 up · 1 down
Alerts
ProMetric rules held for 1 to 15 minutes, status rules for machines, containers, and hosts. Every incident gets a timeline and can be resolved by hand. Delivered by email, with flood control.
- Metric rules on CPU / GPU / RAM / Disk with >, <, =, held 1 to 15 minutes
- Status rules for machine, container, and host
- Scope to all machines, a group, or specific ones - with exclusions
- Incidents with a timeline and manual resolve
- Flood control: 30-minute dedup, at most 20 emails an hour, then an hourly digest
New alert rule
Recent incidents
Groups
Nest groups to mirror your infrastructure - regions, environments, teams. Drag to reorder, and deleting a group reparents its contents.
- Parent / child hierarchy, nested as deep as you need
- Drag-reorder on Dashboard and Machines pages
- Group cards show online / health ratios at a glance
Groups
Teams and sharing
Share machines and hosts with a team instead of sharing a password. Owner and admin roles decide who can change what.
- Share individual machines and hosts with a team
- Owner and admin roles
- Pro includes 3 seats - you and two more; Team has unlimited seats
Team · Infra
3 of 3 seats
Shared with this team
Machines
6
Hosts
4
Two-factor and sessions
TOTP two-factor on every plan. Agents link with a device code, so your password never goes on a machine, and changing it signs every other session out.
- TOTP with QR setup and 10 recovery codes, on all plans
- Agents authenticate with a device code - statsd login prints a link, you approve it
- Refresh tokens rotate; every session is revoked on password change
Two-factor authentication
Scan with your authenticator
TOTP · 6 digits · 30s
Dashboard
One overview with fleet health, a ranked list of what needs attention, clickable summary cards, active incidents, group cards, and the last day of incidents.
- Needs attention: disks filling up with a days-to-full forecast, unhealthy containers, memory pressure, sustained CPU, silent agents, hosts down. No rules needed, on every plan
- Recent incidents across alerts and hosts in one feed, resolvable in place
- System-wide health dot: green / orange / red / gray
- Clickable cards for machines, containers, health, alerts
- Command palette (Cmd / Ctrl K), light / dark / system theme, installable as a PWA
Dashboard
Overview of all systems
Machines
14
13 online · 1 offline
Containers
47
42 running · 5 stopped
Health
1 unhealthy
41 healthy · 1 unhealthy
Alerts
1 active
6 configured · 1 firing
Needs attention
Active incidents
Install the agent and see it all
Free up to 5 machines forever. No credit card. Two minutes from landing here to live metrics streaming in.