- We already have Prometheus. What is left to do?
- Usually three things: retention, which is where most self-built stacks quietly stop at fifteen days; alerting, which tends to be a set of thresholds nobody has revisited since installation; and the logs and traces that were never wired in, so an incident still ends in kubectl logs. The install is the easy part and it is rarely where the value is.
- What does the free health check actually involve?
- A read-only look at what you run today, then a written list of findings ordered by what would help most: signal-to-noise, alert hygiene, retention, storage and cost. It takes a short call and access to your dashboards. You keep the findings whether or not we work together, and there is no obligation attached to it.
- Does this only work on Kubernetes?
- No. Kubernetes is where most of the work happens, but the same collectors run on plain virtual machines and bare metal, and managed cloud services are scraped through their own exporters. A hybrid estate is normal rather than a complication.
- Self-hosted or SaaS?
- Depends on your volume, your team, and what you are required to keep in-house. Self-hosted is usually cheaper past a certain telemetry volume and costs you the operational burden of running it. We have built and run both, so what you get is an estimate of what each costs in practice rather than a preference. If Datadog is the answer, that has its own page.
- How long does it take?
- One week for a single cluster monitored properly, three to four weeks for a full production stack with logs, traces, profiles and long-term storage. Multi-cluster work is quoted after a short scoping call, because the answer depends on how much of it already exists.
- Do you work with teams outside the Netherlands?
- Yes. We work remotely with teams across Europe.