Cognitum/Cluster Health Monitorv1.0.0

📟 Cluster Health Monitor

Real-time dashboard of all seeds in the swarm. Monitors uptime, vector count, CPU temperature, WiFi signal, disk usage, and installed cogs — and alerts when a seed goes offline or exceeds thresholds.

Live
Seeds online
11 / 12
1 offline · 92% reachable
Active alerts
3
1 critical · 2 warning
Avg CPU temp
54.2°C
↓ 2.1°C vs 1h ago
Vectors indexed
2.41M
across 12 seeds
Avg WiFi signal
-62 dBm
good signal strength
Fleet uptime
14d 6h
median across seeds

Swarm seeds

Edge Seed Fleet · 12 registered · 11 online
Pi Zero 2 W seeds running cognitum-seed agents on the Tailscale network
Fleet JSON

Active alerts

3 firing
threshold_exceeded · seed_offline

Event stream

Recent events
seed_online · seed_offline · threshold_exceeded · health_report

Threshold configuration

Alert thresholds
Trigger threshold_exceeded events when any seed exceeds these limits
Metric Warning Critical Sustained for Status
CPU temperature°C 70 85 5m 1 breaching
Disk usage% 85 95 1m 1 breaching
WiFi signaldBm (lower = weaker) -75 -85 2m nominal
Memory usage% 80 92 3m nominal
Offline timeoutmissing heartbeats 2 5 1 critical