{"id":4866,"date":"2026-08-19T21:15:52","date_gmt":"2026-08-19T21:15:52","guid":{"rendered":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/2026\/08\/19\/what-it-really-takes-to-run-opentelemetry\/"},"modified":"2026-08-19T21:15:52","modified_gmt":"2026-08-19T21:15:52","slug":"what-it-really-takes-to-run-opentelemetry","status":"publish","type":"post","link":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/2026\/08\/19\/what-it-really-takes-to-run-opentelemetry\/","title":{"rendered":"What It Really Takes to Run OpenTelemetry"},"content":{"rendered":"<div><img data-opt-id=177232841  fetchpriority=\"high\" decoding=\"async\" width=\"770\" height=\"330\" src=\"https:\/\/devops.com\/wp-content\/uploads\/2026\/08\/opentelemetry_observability_770x330.jpg\" class=\"attachment-large size-large wp-post-image\" alt=\"\" \/><\/div>\n<p><img data-opt-id=162896923  fetchpriority=\"high\" decoding=\"async\" width=\"150\" height=\"150\" src=\"https:\/\/devops.com\/wp-content\/uploads\/2026\/08\/opentelemetry_observability_770x330-150x150.jpg\" class=\"attachment-thumbnail size-thumbnail wp-post-image\" alt=\"\" \/><\/p>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\" data-doc-id=\"7434132000065822436\" data-doc-type=\"writer\">OpenTelemetry solved a real problem. Before it, every APM vendor had you install a proprietary agent, which meant switching vendors meant re-instrumenting everything. OTel gave engineering teams a vendor-neutral way to generate traces, metrics, and logs once and send them wherever they wanted.<\/p>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">That part of the pitch is true, and open-source deserves credit for it. What gets left out of most OTel adoption conversations is what happens after the SDKs are wired up. The framework is free. Running it well is not.<\/p>\n<h3 class=\"zw-paragraph heading2\" data-header=\"2\" data-margin-bottom=\"14.94pt\" data-margin-top=\"14.94pt\" data-hd-info=\"2\" data-line-height=\"1.2\">The Pain Points That Show Up After Rollout<\/h3>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">Collector sprawl. A production OTel deployment usually means running collector instances per region or per cluster, tuning batch and memory limiter settings, and watching for the collector itself becoming a bottleneck under load. This is the infrastructure your team now owns and patches, on top of the infrastructure it was supposed to help you monitor.<\/p>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">Storage and retention decisions become your job. OTel defines how telemetry is generated and transported, not where it lives. Teams end up choosing and operating a backend, commonly a trace store, a time-series database (TSDB) for storing and querying metrics over time, and a log index, then building the queries and dashboards to make that data usable. Every version upgrade across that chain is a coordination exercise.<\/p>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">Cross-signal correlation requires continuous engineering. Traces, metrics, and logs arriving in three different systems does not automatically mean an engineer can jump from a slow span to the exact log line or the query that caused it. Building that correlation layer, and keeping it working as schemas evolve, is ongoing engineering work, not a one time setup.<\/p>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">The on call burden shifts internally. With a managed platform, a vendor\u2019s SRE team is paged when ingestion breaks. With a self-run OTel stack, it\u2019s usually the same engineers who were supposed to be using the telemetry to <a href=\"https:\/\/www.manageengine.com\/it-operations-management\/application-observability\/?utm_source=devops&amp;utm_medium=article3&amp;utm_campaign=appObservability\">fix application problems<\/a>, not maintaining the pipeline that delivers it.<\/p>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">Upgrade churn. OTel semantic conventions and SDKs still move quickly. Staying current across every instrumented service, especially in a polyglot environment, is recurring work that never shows up in a headcount plan but consumes real hours every quarter.<\/p>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">None of this means self-managed OTel is a bad choice. For teams with the platform engineering capacity to run it well, and a real need for full control over the pipeline, it\u2019s a legitimate architecture. The honest accounting is that free software still has a labor cost, and that cost scales with the number of services, languages, and regions you\u2019re instrumenting.<\/p>\n<h3 class=\"zw-paragraph heading2\" data-header=\"2\" data-margin-bottom=\"14.94pt\" data-margin-top=\"14.94pt\" data-hd-info=\"2\" data-line-height=\"1.2\">Where the Labor Actually Goes<\/h3>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">Ask any team that\u2019s run this for two years and the pattern is consistent. Initial instrumentation is the easy part. The ongoing cost lives in three places, keeping collectors healthy under changing load, keeping the backend\u2019s storage and indexing performant as volume grows, and rebuilding correlation logic every time a service boundary changes.<\/p>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">That last one matters most for incident response. A trace showing high latency in one service tells you where. It doesn\u2019t tell you whether the cause was a missing database index, a downstream API timeout, or a resource constraint on the host. Getting from where to why usually requires stitching context across systems that weren\u2019t designed to talk to each other, which is exactly the work a managed backend takes off your plate.<\/p>\n<h3 class=\"zw-paragraph heading2\" data-header=\"2\" data-margin-bottom=\"14.94pt\" data-margin-top=\"14.94pt\" data-hd-info=\"2\" data-line-height=\"1.2\">What a Converged Platform Changes<\/h3>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">ManageEngine\u2019s OpManager Nexus <a href=\"https:\/\/www.manageengine.com\/it-operations-management\/application-observability\/open-telemetry.html?utm_source=devops&amp;utm_medium=article3&amp;utm_campaign=appObservabilityopnTM\">accepts OpenTelemetry data natively<\/a>. Teams keep their existing OTel instrumentation and export traces, metrics, and logs directly to the platform without operating a separate collector and storage layer. The platform handles ingestion, correlation, root cause analysis, real user monitoring, and anomaly detection on top of that same OTel data, so the instrumentation investment a team already made doesn\u2019t get thrown away.<\/p>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">That\u2019s the actual trade being made. You keep the vendor-neutral instrumentation layer that makes OTel valuable in the first place. You hand-off the operational weight of running collectors, tuning storage, and building correlation logic, which is the part that consumes engineering hours long after the initial rollout is done. But that trade isn\u2019t free of its own limits.<\/p>\n<h3 class=\"zw-paragraph heading2\" data-header=\"2\" data-margin-bottom=\"14.94pt\" data-margin-top=\"14.94pt\" data-hd-info=\"2\" data-line-height=\"1.2\">Where Self Hosting Still Wins<\/h3>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">Teams with mature platform engineering practices, a specific compliance reason to own the storage layer, or deep enough Kubernetes native maturity that a collector fleet is just another workload they already know how to run, will find a managed backend more constraining than a stack they built themselves. Sampling logic, retention policy, and pipeline level customization go deeper in a self-hosted setup than any managed platform will expose. The community exporter and processor ecosystem around OpenTelemetry is also larger than what a single vendor surfaces natively, simply because it\u2019s not bounded by one company\u2019s roadmap.<\/p>\n<p class=\"zw-paragraph heading0\" data-header=\"0\" data-textformat='{\"fgc\":\"rgb(0, 0, 0)\",\"size\":\"12.00\",\"type\":\"text\"}' data-margin-bottom=\"12pt\" data-margin-top=\"12pt\" data-hd-info=\"0\" data-line-height=\"1.2\">So the decision isn\u2019t open-source versus managed, it\u2019s where you want your engineering hours to go. Every team instrumenting with OTel already made the right call on the instrumentation layer. The question worth revisiting is whether the collector, storage, and correlation work sitting underneath that instrumentation is still the best use of the team\u2019s time, or whether it\u2019s become a second job nobody signed up for.<\/p>\n<p>If you\u2019re already running OTel instrumentation and want to see what happens when that same data feeds into a platform that handles ingestion, correlation, and root cause analysis for you, <a href=\"https:\/\/www.manageengine.com\/it-operations-management\/download.html?utm_source=devops&amp;utm_medium=article3&amp;utm_campaign=telemtrydwnlad\">setup OpManager Nexus within minutes<\/a> without re-instrumenting anything you\u2019ve already built.<\/p>\n<p><a href=\"https:\/\/devops.com\/what-it-really-takes-to-run-opentelemetry\/\" target=\"_blank\" class=\"feedzy-rss-link-icon\">Read More<\/a><\/p>\n<p>\u200b<\/p>","protected":false},"excerpt":{"rendered":"<p>OpenTelemetry solved a real problem. Before it, every APM vendor had you install a proprietary agent, which meant switching vendors [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":4867,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"site-sidebar-layout":"default","site-content-layout":"","ast-site-content-layout":"default","site-content-style":"default","site-sidebar-style":"default","ast-global-header-display":"","ast-banner-title-visibility":"","ast-main-header-display":"","ast-hfb-above-header-display":"","ast-hfb-below-header-display":"","ast-hfb-mobile-header-display":"","site-post-title":"","ast-breadcrumbs-content":"","ast-featured-img":"","footer-sml-layout":"","ast-disable-related-posts":"","theme-transparent-header-meta":"","adv-header-id-meta":"","stick-header-meta":"","header-above-stick-meta":"","header-main-stick-meta":"","header-below-stick-meta":"","astra-migrate-meta-layouts":"default","ast-page-background-enabled":"default","ast-page-background-meta":{"desktop":{"background-color":"var(--ast-global-color-4)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"tablet":{"background-color":"","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"mobile":{"background-color":"","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""}},"ast-content-background-meta":{"desktop":{"background-color":"var(--ast-global-color-5)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"tablet":{"background-color":"var(--ast-global-color-5)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"mobile":{"background-color":"var(--ast-global-color-5)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""}},"footnotes":""},"categories":[5],"tags":[],"class_list":["post-4866","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-devops"],"_links":{"self":[{"href":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/wp-json\/wp\/v2\/posts\/4866","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/wp-json\/wp\/v2\/comments?post=4866"}],"version-history":[{"count":0,"href":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/wp-json\/wp\/v2\/posts\/4866\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/wp-json\/wp\/v2\/media\/4867"}],"wp:attachment":[{"href":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/wp-json\/wp\/v2\/media?parent=4866"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/wp-json\/wp\/v2\/categories?post=4866"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/rssfeedtelegrambot.bnaya.co.il\/index.php\/wp-json\/wp\/v2\/tags?post=4866"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}