Unified logs and status across a multi-site fleet — self-hosted, mesh-private, and filtered at the edge so volume (and cost) stay low.
First-party build — our own multi-region database + application platform. TODO: founder to confirm attribution.
Figures pending verification — placeholders until confirmed.
A multi-site fleet — databases across sites plus application pods on k3s — had no single place to see health and errors. The obvious SaaS options price per host and per ingested GB, meaning visibility gets more expensive exactly as the fleet grows.
We ship with Vector — a single lightweight agent — and filter at the edge with VRL so only warn/error/fatal log lines and 60-second status probes ever leave a host. That cuts ingest volume, and therefore cost and noise, at the source instead of paying to store and index everything. It all lands in OpenObserve — one backend for both the database-infra streams and the application logs — reached over a Tailscale mesh, so no agent or backend is exposed to the public internet. WireGuard/headscale were considered and deliberately rejected in favour of a single Tailscale mesh.
The pipeline unifies multi-site database infrastructure and application logs into one self-hosted, mesh-private backend, with edge filtering holding volume (and cost) down. TODO: founder to fill with the real result — actual monthly cost vs. a comparable Datadog/SaaS quote, and ingest volume (GB/day or events/sec). These need real numbers before they are stated publicly.
Bring it to a senior US engineer who'll scope it, build it, prove it, and document it. Start with a call — or prove the approach in a one-week sprint first.