Build a Central Observability Platform
Plan a central observability platform across on-premises and Azure: ownership, Collector blueprints, change control and measurable pilot acceptance criteria.
Plan a central observability platform across on-premises and Azure: ownership, Collector blueprints, change control and measurable pilot acceptance criteria.

Services drift on attribute names until your queries quietly miss data. How to normalize OpenTelemetry semantic conventions at the collector, fleet-wide.

Most sizing guides get CPU right and memory wrong. A ratio-based method for sizing an OpenTelemetry gateway, and why scaling out won't fix memory.

A Splunk-to-Sentinel move only pays off if data volume changes on the way. What a lift-and-shift costs, and where the pipeline layer belongs.

A reference architecture for telemetry under revDSG and FINMA: classify at the edge, a Swiss gateway tier, split destinations, a control plane you host.

The Collector stopped being a protocol shim and became the enforcement point for cost, privacy and routing. What that shift changes about operating it.

Attribute-based routing in the OpenTelemetry Collector — a priority-ordered, first-match-wins cascade you build with a picker instead of raw OTTL.

Keep credentials out of committed Collector config: env vars, file substitution, Kubernetes Secrets, Vault — and how to rotate without downtime.

Every other data domain grew a middle tier. Observability never did — because agents shipped with backends. What that historical accident still costs.

Observability explores telemetry; AIOps applies ML to detect, correlate and automate. Where the line sits, and why your pipeline decides both.

Every backend cost lever acts on data you already transmitted. Where the spend is really decided, and why commitment tiers make it permanent.

The 10 storage metrics worth tracking — capacity, performance, health — and how to collect them from servers and arrays with one OTel Collector.

Splunk, Sentinel, Grafana and Dynatrace in one estate is normal, not a mistake. Collect once and fan out, instead of one agent per backend.

Dual-ship telemetry to two backends to validate a migration — without double-counting hosts, doubling metrics or duplicating every log line.

Seven Cribl alternatives compared for 2026 — self-hosted, OpenTelemetry-native and SaaS — with an honest look at when Cribl is still the right call.

Isolation, tenant identification and per-tenant routing for a shared OpenTelemetry Collector fleet — dedicated vs shared pipelines, quotas, RBAC.

A single Collector is a single point of failure. HA across agent and gateway tiers: load-balanced pools, sending queues, and persistent queues.

A practical playbook to cut observability spend: measure per-route volume, filter noise at the edge, sample, and route before the invoice lands.

OTLP, the Collector and OpAMP make collection portable and give you a real exit path. What OpenTelemetry decouples — and what it doesn't.

Five ways to deploy an OpenTelemetry Collector — agent-per-host, DaemonSet, sidecar, gateway, hybrid — compared, plus the no-collector option.

The reference architecture for OpenTelemetry at scale: agent collectors, a central gateway tier, and OpAMP as the control plane that configures it.

A telemetry pipeline sits between your systems and your backends — where you control cost, routing and PII. What it is and how to run your own.