Every product here meters on volume: hosts monitored, gigabytes ingested, events captured, sessions recorded. A per-user mental model — the one that works everywhere else on this site — produces a forecast that is wrong by an order of magnitude.
Datadog publishes three different meters inside one platform: APM and Infrastructure per host, Log Management per GB ingested plus per million events indexed, RUM per thousand sessions. Model your volumes against the meter, not the rate.
Already decided — What this page decides
Still yours to weigh
Observability is the practice of keeping enough telemetry — logs, metrics, traces — that you can answer a question you did not anticipate. Monitoring answers questions you wrote a check for in advance. That distinction sounds academic until you see the bill: keeping the detail is exactly what you pay for.
The twenty-three products below split by meter more usefully than by feature. Eight are per-host or per-GB platform modules. Five are event-metered and materially cheaper for the same job at small volumes. Five are quote-only. And two — Mixpanel’s pair — are product analytics rather than infrastructure observability, included because the meter and the buying team overlap, and labelled so you do not buy one for the other. And one — Kentik — watches the network rather than the application, on a published annual tier sized in flows per second.
The row to get exactly right
The pricing meter. This is the category’s equivalent of a firewall’s throughput figure: a per-host quote and a per-GB quote for the same estate can differ by an order of magnitude, and no comparison on the internet models it against your volumes. Ask every vendor to price your actual host count, ingest volume and event rate — then compare the totals, never the rates.
Often confused withDatabases & Data Tools — why the query plan changed →·Developer Tools — where the code and the pipeline live →·SIEM & Log Management — the same logs, asked a security question →
These are not tiers. Observability is not monitoring done better — it answers questions you never wrote a check for, from telemetry you pay to keep, at a materially different cost.
Monitoring vs observability
Did you know in advance what to check?
Observability vs APM
The whole estate, or the application request?
Logs vs metrics vs traces
Which one answers your question, and what does it cost to keep?
APM vs RUM vs error tracking
Server-side performance, device experience, or exceptions?
Seven variables move the shortlist. The first one moves it more than the other six combined.
The pricing meter
Per host, per GB ingested, per million events indexed, per session, per user — or quote-only. Same estate, different meters, order-of-magnitude difference.
Retention, and what keeping it costs
The licence covers the first tranche. Shortening retention to control cost is the standard reflex, and the incident then falls outside the window.
Language and runtime coverage
For the stack you actually run, including the older service nobody wants to touch.
Cardinality limits
The failure mode nobody forecasts: one well-meant custom tag turns a metric into millions of time series.
Alerting and on-call integration
Whether alerts reach the rota you already run, or need a second product.
Open-source viability
Prometheus, Grafana and OpenTelemetry are genuinely credible here. A page pretending otherwise is not worth reading.
India region and residency
Documented for Dynatrace (AWS Mumbai) and Elastic (Mumbai and Pune regions); Datadog has no India site; the rest do not document one — logs frequently carry personal data, so confirm rather than assume.
Pick what the telemetry has to cover and how it has to be billed. Products drop out with the reason stated, never silently.
What it has to see
How it has to work
What the bill counts
India
India residency is annotated rather than used to eliminate: only Dynatrace and Elastic document an India telemetry region, so every other product is flagged for you to confirm rather than ruled out.
engine free under AGPL plus a subscription tier; Elastic Cloud Hosted priced on PROVISIONED RESOURCES rather than per host or per GB (reported ~$99/mo Standard to ~$184/mo Enterprise at entry size), or Serverless on consumption
Teams already running the ELK stack for logs — this formalises what they have rather than migrating — and estates large enough that per-host pricing has become the problem.
The catch: You run the cluster unless you buy Elastic Cloud, and operating Elasticsearch well at scale is real engineering work. The correlation UX is less polished than Datadog’s and dashboards need more assembly.
per host / month, published list; distributed tracing across services with the span-level detail that names the slow hop — billed on top of Infrastructure Monitoring rather than instead of it
Estates running microservices that need to follow one request across many services, and that already accept a platform bill rather than a per-tool one.
The catch: The meter is per host, so the bill tracks your infrastructure rather than your traffic — and it compounds with every other Datadog module you enable. Cost management is the standing operational task, not a one-time negotiation. Datadog runs no India site: its Asia-Pacific sites are AP1 (Japan) and AP2 (Australia).
per host / month, published list; the base layer of the platform — host, container and cloud-service metrics with the dashboards and alerting most estates start from
Teams that want one place for infrastructure health across cloud accounts, and are choosing the platform deliberately rather than by accident.
The catch: Metrics only: it will tell you the p99 got worse and not which span caused it. Custom metrics and high cardinality are separately metered, and cardinality is the line that multiplies a bill without traffic multiplying. Datadog runs no India site: its Asia-Pacific sites are AP1 (Japan) and AP2 (Australia).
per GB ingested PLUS per million events indexed, published list; ingest and indexing are separate meters, so what you keep searchable costs more than what you merely collect
Estates whose debugging happens in logs and who are prepared to decide, deliberately, which logs are worth indexing.
The catch: Two meters on one product is where forecasts break: teams model ingest and forget indexing. Almost every runaway observability bill in this category is a log bill. Datadog runs no India site: its Asia-Pacific sites are AP1 (Japan) and AP2 (Australia).
per 1,000 sessions / month for RUM, with synthetic tests metered separately by run; what the browser or mobile app actually experienced, including the network and third-party scripts the backend never sees
Teams whose users report slowness that server-side metrics insist is not happening — the gap between what the server did and what the device experienced.
The catch: Session-metered, so a traffic spike is a bill spike with no infrastructure change to explain it. Synthetics are a third meter again. Datadog runs no India site: its Asia-Pacific sites are AP1 (Japan) and AP2 (Australia).
per month on the Team tier (about ₹2,158), Business about $80, both annual; metered on events captured rather than hosts run — exceptions grouped by fingerprint with the release, the commit and the users affected
Teams whose actual question is “why did that crash, in which release, for how many people” — answered at a fraction of full APM pricing.
The catch: Errors and performance for the application, not the infrastructure: no host metrics, no cloud-service integration, no log platform. It is deliberately one job done well rather than a platform.
metered on transactions/spans within the same published Sentry tiers; developer-first tracing joined to the errors and the release that introduced them
Engineering teams that want tracing tied to the commit that caused the regression, without adopting a per-host platform.
The catch: Sampled by design and scoped to the application — it will not replace infrastructure monitoring, and deep cloud-service telemetry is outside it.
metered per replay within the Sentry tiers; a recording of the session that produced the error, so the reproduction step is watching rather than guessing
Teams losing time to bugs that cannot be reproduced from a stack trace alone.
The catch: Replays are a separate meter and privacy masking must be configured deliberately — an unmasked replay of a checkout page is a data-protection problem, not a debugging win.
per host / month billed annually, published starting prices: Infrastructure $15, App & Infra $60, End-to-End $75 (adds RUM and browser synthetics); a full-fidelity platform built on OpenTelemetry with no-sample tracing
Estates that want every trace kept rather than sampled, and that are already in the Cisco/Splunk relationship — a distinct SKU from Splunk Enterprise Security, which is carded on the Security category.
The catch: The published figures are starting prices per host, so the bill tracks infrastructure like Datadog's, and the final number is still a quote. Full-fidelity tracing is a genuine differentiator and a genuine volume commitment. Splunk documents no India realm for Observability Cloud.
quoted on the Splunk platform; service-level health modelling and AIOps event correlation over data the platform already holds — the layer that turns thousands of alerts into a handful of service states
Large estates already ingesting into Splunk whose problem is alert volume rather than missing telemetry.
The catch: It assumes the Splunk platform underneath and its value is proportional to how much you already ingest there. Service modelling is configuration work measured in weeks, not a switch.
quoted per monitored resource; SaaS-delivered infrastructure and application monitoring from the vendor whose depth is databases — the cloud front end to a Foglight estate
Estates already running Foglight for databases that want the same console over cloud infrastructure rather than a second platform.
The catch: Narrower application-tracing depth than the observability specialists; its strength is the database line, and buying it as a general APM is buying the wrong half of the portfolio.
quoted per monitored host or VM; virtual and cloud infrastructure monitoring with capacity planning — built for VMware-shaped estates and their migrations
Estates with a large virtualised footprint that need capacity forecasting alongside health, particularly through a hypervisor migration.
The catch: Infrastructure-scoped: no application tracing and no log platform. Its natural buyer is the virtualisation team, not the engineering team this page is written for.
about $0.28 per 1,000 events over the free allowance, Enterprise from roughly $25–30k a year; product analytics — what users did in the product, funnels and retention, not whether the server was healthy
Product and engineering teams asking which features are used and where people drop out — a different question from whether the system is up.
The catch: This is product analytics, not infrastructure observability: it will not tell you a service is down or a query is slow. Included here because the event meter and the buying team overlap, not because it substitutes for APM.
metered within the Mixpanel event tiers; replays joined to the analytics events, so a funnel drop-off can be watched rather than inferred
Teams that have the funnel data and still cannot explain the drop-off at one step.
The catch: Tied to the Mixpanel platform and to product questions rather than engineering ones. Privacy masking is a deliberate configuration, not a default.
per host / month published: $29 for infrastructure monitoring, $58 per 8 GiB of host memory for full-stack — drawn down from one annual Dynatrace Platform Subscription; one agent instruments hosts, processes and code, and a live dependency map is what the root-cause analysis reasons over
Estates where incident time is lost finding the cause rather than fixing it, and that want one agent instead of per-service instrumentation.
The catch: Full-stack is priced per 8 GiB of host memory, so large-memory hosts count several times over. Root cause is only as good as the coverage — hops the agent does not see are gaps in the map. The Mumbai region is documented; a written storage commitment is a contract question.
per GiB published, split three ways: $0.20 to ingest, $0.0007 per GiB-day to retain and $0.0035 per GiB scanned by queries — or $0.02 per GiB-day with queries bundled; logs sit in Grail beside the traces and metrics they explain
Teams that want logs in the same store as their traces, and that will set retention per source rather than keep everything for a year.
The catch: Pay-per-query charges per GiB scanned, so wide queries over long windows add up — choose between pay-per-query and bundled on how the team actually searches. DQL is new to most teams.
per 1,000 real-user sessions published ($4.50 with Session Replay), synthetic browser tests at $4.50 per 1,000 actions; each session is joined to the backend trace behind it. A Leader in the 2025 Gartner Magic Quadrant for Digital Experience Monitoring
Digital and product teams that need a slow page to arrive with the backend service that caused it, and replay evidence behind a conversion drop.
The catch: Cost scales with sessions, so a high-traffic consumer app needs a sampling policy decided up front, and replay is safe only once masking is configured and reviewed.
per node a month for network and infrastructure; APM $27.50 per service, logs $5 per GB, databases $70 per instance, synthetics and RUM priced apart — multi-year contracts billed annually
Teams that want network, infrastructure, APM and logs in one SaaS with every module on a published price, and a bridge to self-hosted SolarWinds.
The catch: Data sits in North America, Frankfurt or Australia — no India region; each module is priced separately, so model the whole bill; a Niche Player in Gartner’s Observability Platforms Magic Quadrant.
per node a month on multi-year contracts billed annually (Essentials; Advanced $14, Premier $17.50); Enterprise from 500 nodes; NPM, SAM and other modules still sold singly on quote
Network and infrastructure teams that must keep monitoring data on their own servers — including in India — with the Orion-era modules in one platform.
The catch: You run and patch it (critical flaws fixed in September 2026), and when the subscription lapses polling of new data stops; SUNBURST (2020) came through this platform’s build system — the SEC case was dismissed in November 2025.
quote-only in two editions — Express for private cloud, Premium for private or public cloud with real user monitoring — plus paid add-ons for integrations, event and metric analytics, GenAI incident management, observability, automation and discovery/CMDB; SaaS, hybrid, on-premises or air-gapped
Operations teams running many monitoring tools who want one event view above them — 1,000+ prebuilt integrations, AI correlation OpenText says cuts event volume by 30–95%, and runbooks gated by approvals.
The catch: An event and AIOps layer, not deep APM: code-level traces and logs come from the Application Observability add-on or the tools it ingests. CVE-2025-3476 (CVSS 9.4) hit the event console in 2023.05 to 24.4; no OpenText-hosted India region is published.
Pro from US$2,000 a month billed annually (US$24,000 a year) with 50 users, 1,000 flows per second and 25 monitored devices included, a 30-day trial and Premier by quote; network flow, device metrics, synthetic tests and cloud flow logs; Infoblox acquired Kentik in August 2026
Network and platform teams that need to see traffic across data centres, clouds and the internet — flow, routing and synthetic tests — rather than application traces.
The catch: Network observability, not APM: it does not trace code. Data is stored in Ashburn, Virginia or Frankfurt, with no India region, full-resolution flow is kept 45 days by default, and the published price is a US list for annual contracts only.
per managed virtual server (MVS, roughly a host) / month on SaaS; $0.03 per MVS-hour pay-per-use; self-hosted $385.20 per MVS / year; 14-day trial; SaaS in Mumbai since April 2024 (IBM staff blog)
Teams that want one-second, automatically discovered traces and dependency maps without hand instrumentation, priced per host rather than per GB.
The catch: MVS metering follows infrastructure size, not traffic, and IBM should confirm in writing how containers and Kubernetes nodes are counted; retention is not published.
logs from $0.49 per GiB ingested, traces from $0.59 per GiB, metrics from $0.008 per data point per minute; compute, unlimited users and 30-day retention for logs and traces included (metrics 13 months); committed-volume subscription
Teams whose Datadog or Splunk bill grew faster than their traffic and who want logs, metrics and traces in an open lakehouse with an AI SRE on top.
The catch: A committed-ingest subscription, so you size the commitment up front; no Indian hosting region is documented, and RUM and session replay are not listed. Snowflake bought Observe in 2026, so packaging is still settling.
distributed tracesRules out Datadog Infrastructure Monitoring, Datadog Log Management, Datadog Digital Experience Monitoring, Sentry Error Monitoring, Sentry Session Replay, Splunk ITSI, Quest Foglight Cloud, Quest Foglight Evolve, Mixpanel Product Analytics, Mixpanel Session Replay, Dynatrace Log Analytics, Dynatrace Digital Experience Monitoring, SolarWinds Observability Self-Hosted, OpenText AI Operations Management (formerly Operations Bridge) and Kentik (an Infoblox company) — no distributed tracing. That leaves Elastic Observability, Datadog APM, Sentry Performance Monitoring, Splunk Observability Cloud, Dynatrace Application & Infrastructure Observability, SolarWinds Observability SaaS, IBM Instana Observability and Observe by Snowflake.
log searchRules out Datadog APM, Datadog Infrastructure Monitoring, Datadog Digital Experience Monitoring, Sentry Error Monitoring, Sentry Performance Monitoring, Sentry Session Replay, Quest Foglight Cloud, Quest Foglight Evolve, Mixpanel Product Analytics, Mixpanel Session Replay, Dynatrace Digital Experience Monitoring and OpenText AI Operations Management (formerly Operations Bridge) — not a log platform. That leaves Elastic Observability, Datadog Log Management, Splunk Observability Cloud, Splunk ITSI, Dynatrace Application & Infrastructure Observability, Dynatrace Log Analytics, SolarWinds Observability SaaS, SolarWinds Observability Self-Hosted, Kentik (an Infoblox company), IBM Instana Observability and Observe by Snowflake.
real user monitoringRules out Datadog APM, Datadog Infrastructure Monitoring, Datadog Log Management, Sentry Error Monitoring, Sentry Performance Monitoring, Splunk ITSI, Quest Foglight Cloud, Quest Foglight Evolve, Mixpanel Product Analytics, Dynatrace Application & Infrastructure Observability, Dynatrace Log Analytics, SolarWinds Observability Self-Hosted, Kentik (an Infoblox company) and Observe by Snowflake — server-side only: it cannot see what the device experienced. That leaves Elastic Observability, Datadog Digital Experience Monitoring, Sentry Session Replay, Splunk Observability Cloud, Mixpanel Session Replay, Dynatrace Digital Experience Monitoring, SolarWinds Observability SaaS, OpenText AI Operations Management (formerly Operations Bridge) and IBM Instana Observability.
error trackingRules out Elastic Observability, Datadog APM, Datadog Infrastructure Monitoring, Datadog Log Management, Datadog Digital Experience Monitoring, Sentry Performance Monitoring, Sentry Session Replay, Splunk Observability Cloud, Splunk ITSI, Quest Foglight Cloud, Quest Foglight Evolve, Mixpanel Product Analytics, Mixpanel Session Replay, Dynatrace Application & Infrastructure Observability, Dynatrace Log Analytics, Dynatrace Digital Experience Monitoring, SolarWinds Observability SaaS, SolarWinds Observability Self-Hosted, OpenText AI Operations Management (formerly Operations Bridge), Kentik (an Infoblox company), IBM Instana Observability and Observe by Snowflake — no dedicated error grouping with release and commit context. That leaves Sentry Error Monitoring.
OpenTelemetry-nativeRules out Datadog APM, Datadog Infrastructure Monitoring, Datadog Log Management, Datadog Digital Experience Monitoring, Sentry Error Monitoring, Sentry Performance Monitoring, Splunk ITSI, Dynatrace Digital Experience Monitoring, OpenText AI Operations Management (formerly Operations Bridge) and IBM Instana Observability — OpenTelemetry is supported but the vendor agent is the primary path; Sentry Session Replay, Quest Foglight Cloud, Quest Foglight Evolve, Mixpanel Product Analytics, Mixpanel Session Replay, SolarWinds Observability Self-Hosted and Kentik (an Infoblox company) — OpenTelemetry support is not documented. That leaves Elastic Observability, Splunk Observability Cloud, Dynatrace Application & Infrastructure Observability, Dynatrace Log Analytics, SolarWinds Observability SaaS and Observe by Snowflake.
one platformRules out Sentry Error Monitoring, Sentry Performance Monitoring, Sentry Session Replay, Quest Foglight Cloud, Quest Foglight Evolve, Mixpanel Product Analytics, Mixpanel Session Replay and Kentik (an Infoblox company) — focused on one job rather than covering the estate. That leaves Elastic Observability, Datadog APM, Datadog Infrastructure Monitoring, Datadog Log Management, Datadog Digital Experience Monitoring, Splunk Observability Cloud, Splunk ITSI, Dynatrace Application & Infrastructure Observability, Dynatrace Log Analytics, Dynatrace Digital Experience Monitoring, SolarWinds Observability SaaS, SolarWinds Observability Self-Hosted, OpenText AI Operations Management (formerly Operations Bridge), IBM Instana Observability and Observe by Snowflake.
on-premisesRules out Elastic Observability, Datadog APM, Datadog Infrastructure Monitoring, Datadog Log Management, Datadog Digital Experience Monitoring, Sentry Session Replay, Splunk Observability Cloud, Quest Foglight Cloud, Mixpanel Product Analytics, Mixpanel Session Replay, Dynatrace Log Analytics, Dynatrace Digital Experience Monitoring, SolarWinds Observability SaaS, SolarWinds Observability Self-Hosted, Kentik (an Infoblox company), IBM Instana Observability and Observe by Snowflake — SaaS only. That leaves Sentry Error Monitoring, Sentry Performance Monitoring, Splunk ITSI, Quest Foglight Evolve, Dynatrace Application & Infrastructure Observability and OpenText AI Operations Management (formerly Operations Bridge).
published pricingRules out Elastic Observability, Splunk ITSI, Quest Foglight Cloud, Quest Foglight Evolve and OpenText AI Operations Management (formerly Operations Bridge) — quote-only: no published list, so the quote is the only real number. That leaves Datadog APM, Datadog Infrastructure Monitoring, Datadog Log Management, Datadog Digital Experience Monitoring, Sentry Error Monitoring, Sentry Performance Monitoring, Sentry Session Replay, Splunk Observability Cloud, Mixpanel Product Analytics, Mixpanel Session Replay, Dynatrace Application & Infrastructure Observability, Dynatrace Log Analytics, Dynatrace Digital Experience Monitoring, SolarWinds Observability SaaS, SolarWinds Observability Self-Hosted, Kentik (an Infoblox company), IBM Instana Observability and Observe by Snowflake.
not per hostRules out Datadog APM, Datadog Infrastructure Monitoring, Splunk Observability Cloud, Dynatrace Application & Infrastructure Observability, SolarWinds Observability SaaS, SolarWinds Observability Self-Hosted and IBM Instana Observability — per host, so the bill tracks infrastructure and autoscaling moves it. That leaves Elastic Observability, Datadog Log Management, Datadog Digital Experience Monitoring, Sentry Error Monitoring, Sentry Performance Monitoring, Sentry Session Replay, Splunk ITSI, Quest Foglight Cloud, Quest Foglight Evolve, Mixpanel Product Analytics, Mixpanel Session Replay, Dynatrace Log Analytics, Dynatrace Digital Experience Monitoring, OpenText AI Operations Management (formerly Operations Bridge), Kentik (an Infoblox company) and Observe by Snowflake.
event meteringRules out Elastic Observability, Datadog APM, Datadog Infrastructure Monitoring, Datadog Log Management, Splunk Observability Cloud, Splunk ITSI, Quest Foglight Cloud, Quest Foglight Evolve, Dynatrace Application & Infrastructure Observability, Dynatrace Log Analytics, SolarWinds Observability SaaS, SolarWinds Observability Self-Hosted, OpenText AI Operations Management (formerly Operations Bridge), IBM Instana Observability and Observe by Snowflake — not event-metered. That leaves Datadog Digital Experience Monitoring, Sentry Error Monitoring, Sentry Performance Monitoring, Sentry Session Replay, Mixpanel Product Analytics, Mixpanel Session Replay, Dynatrace Digital Experience Monitoring and Kentik (an Infoblox company).
India regionRules nothing out on published terms. It flags Datadog APM — An India data region for telemetry is not documented on the vendor's pages, Datadog Infrastructure Monitoring — An India data region for telemetry is not documented on the vendor's pages, Datadog Log Management — An India data region for telemetry is not documented on the vendor's pages, Datadog Digital Experience Monitoring — An India data region for telemetry is not documented on the vendor's pages, Sentry Error Monitoring — An India data region for telemetry is not documented on the vendor's pages, Sentry Performance Monitoring — An India data region for telemetry is not documented on the vendor's pages, Sentry Session Replay — An India data region for telemetry is not documented on the vendor's pages, Splunk Observability Cloud — An India data region for telemetry is not documented on the vendor's pages, Splunk ITSI — An India data region for telemetry is not documented on the vendor's pages, Quest Foglight Cloud — An India data region for telemetry is not documented on the vendor's pages, Quest Foglight Evolve — An India data region for telemetry is not documented on the vendor's pages, Mixpanel Product Analytics — An India data region for telemetry is not documented on the vendor's pages, Mixpanel Session Replay — An India data region for telemetry is not documented on the vendor's pages, SolarWinds Observability SaaS — An India data region for telemetry is not documented on the vendor's pages, SolarWinds Observability Self-Hosted — An India data region for telemetry is not documented on the vendor's pages, OpenText AI Operations Management (formerly Operations Bridge) — An India data region for telemetry is not documented on the vendor's pages, Kentik (an Infoblox company) — An India data region for telemetry is not documented on the vendor's pages and Observe by Snowflake — An India data region for telemetry is not documented on the vendor's pages — marked on the cards, not removed.
The meter decides the bill, not the rateA per-host quote and a per-GB quote for the same estate can differ by an order of magnitude. Datadog alone runs three meters — per host for APM and infrastructure, per GB ingested plus per million events indexed for logs, per session for RUM. Model your own volumes against each meter before comparing any two vendors.
Splunk Observability now publishes per-host starting pricesSplunk lists Observability Cloud from $15 per host a month (Infrastructure), $60 (App & Infra) and $75 (End-to-End), billed annually. They are starting points, so the quote is still the final number, and ITSI remains quote-only.
Cardinality is the line nobody forecastsAdd a user ID or request ID as a metric tag and one metric becomes millions of time series. It is the most common cause of a bill that multiplies while traffic does not, and no vendor stops you doing it.
Open source is genuinely viable herePrometheus for metrics, Grafana for dashboards, OpenTelemetry for instrumentation, Loki or Elastic for logs. Unlike most categories on this site, the open-source path is credible for real production estates. It costs engineering time instead of licence — typically an owner, not a side project — and that trade is worth making explicitly rather than by default.
If one of these is your sentence, the shortlist is short.
Why: Exceptions grouped by fingerprint with the release, the commit and the affected users — event-metered, published from about $26 a month.
The trade-off: No infrastructure monitoring and no log platform. If you also need host health, this is half the answer.
Why: Distributed tracing is the only signal that attributes time per hop across service boundaries.
The trade-off: Datadog meters per host and compounds with its other modules; Splunk keeps every trace and is also per host, from $60 for App & Infra.
Why: Move from a host or ingest meter to an event meter, and cut what you index rather than what you collect.
The trade-off: Event-metered tools are narrower. The real fix is usually cardinality and log indexing discipline, not a new vendor.
Why: Only real user monitoring sees the device, the network and the third-party scripts the backend never touches.
The trade-off: Session-metered, so a traffic spike is a bill spike with no infrastructure change to explain it.
Why: OpenTelemetry-native ingest keeps the instrumentation portable, so changing vendor does not mean reinstrumenting.
The trade-off: Per host from $15 a month, billed annually, and full-fidelity tracing is a real volume commitment.
Why: All three offer a self-hosted path where residency or policy forbids SaaS telemetry.
The trade-off: You own the upgrades, the storage and the scaling — the operational burden the SaaS price was covering.
Why: Service-level health modelling collapses alert volume into a handful of service states.
The trade-off: Assumes the Splunk platform underneath, and the service modelling is weeks of configuration.
Why: Genuinely credible for real production estates in this category, unlike most others on this site. Named here because pretending otherwise would waste your time.
The trade-off: It costs an owner rather than a licence. Budget the person, the storage and the upgrade path — and revisit when that person leaves.
Every other category on this site prices per user, per device or per instance — numbers that move when you hire, buy hardware or open an office. You can forecast those. Observability prices on how much your software is used, and that number moves when a marketing campaign works.
The same estate on three different meters:
Model all three at your current volume, at double, and at ten times. The vendor that wins at today’s volume frequently loses badly at ten times, and the contract you sign is usually multi-year.
Ask before signature
Scale here means telemetry volume, not team size — which is the whole point of the category.
One application, modest traffic
Put this in your PoC
Price the event meter before looking at any platform.
Several services, real traffic
Put this in your PoC
Decide what gets indexed versus merely collected.
Microservices at scale
Put this in your PoC
Model the bill at 2× and 10× before signing multi-year.
Very high volume
Put this in your PoC
Name the person who owns cardinality. Nobody does until the bill arrives.
Where a vendor does not publish list pricing, this page says so rather than repeating a third-party figure.
Instrumentation is the lock-in, not the data.
Instrumentation
Vendor agents mean reinstrumenting; OpenTelemetry means changing an endpoint
Dashboards
Rebuilt in the new tool every time — no interchange format exists
Historical telemetry
Exportable in principle, rarely worth the cost of moving
Alert rules and on-call routing
Re-authored, and the tuning that made them quiet is re-learned
The practical consequence: instrument with OpenTelemetry from the start if you expect to change vendor, even if you use a vendor agent today. It is the single cheapest insurance in this category.
By meter, in USD and INR, modelled at more than one volume.
Four checks, in the order most likely to return a yes.
This is the one category on the site where the open-source answer is genuinely competitive for production estates. Saying otherwise would cost you credibility with the engineer reading this.
Three meters, and the published rates that exist.
TechBag quotes every one of these in INR with GST, and models your actual volumes against each meter rather than comparing rates. Where a vendor publishes no list price, this page says so instead of repeating a third-party figure.
TechBag gives INR pricing, GST, PO cycle, minimums and tier-matched quotes. The INR above is conversion for scale at ≈₹83/$; the tier-matched INR quote is ours.
The largest hidden line. The licence covers a tranche; growth and retention are billed on top.
Frequently metered separately. One tag can multiply a metric into millions of series.
Not a licence, but real and recurring. Budget an owner, storage and an upgrade path.
Unless you instrumented with OpenTelemetry, changing vendor means changing every service.
Five ways this purchase goes wrong. Every one of them is a cost surprise rather than a capability gap.
The second-year bill after traffic grew
The forecast was built on today's volume against a meter that tracks usage. Model at 2× and 10× before signing a multi-year contract.
A cardinality explosion from one well-meant tag
Someone adds a user ID to a metric label. One metric becomes millions of time series and the bill multiplies with no traffic change.
Retention shortened to control cost
The reflex works until the incident that matters falls outside the window, and the post-mortem has no data.
Instrumenting everything, alerting on nothing actionable
Full telemetry and an on-call rota that ignores the pages. The tool is not the problem; nobody owns alert quality.
Buying APM when the requirement was error tracking
Ten times the price for a question that error tracking answers better. Read the boundary section above before shortlisting.
Our retail & e-commerce guide maps CERT-In, DPDP, PCI DSS and what gateways, marketplaces and ONDC demand to the controls a store chain or online seller needs. This category answers:
Vendor-neutral. No gated content. · Last reviewed