Previous Job
Previous
Senior Observability Engineer
Ref No.: 26-00741
Location: Virginia
Job Description – Senior Observability Engineer
Location : McLean, VA
ONLY H1B AND W2 NO GC WITH EMPLOYER


Job Summary

We are seeking a highly skilled Senior Observability Engineer with 9+ years of experience
in designing, implementing, and optimizing enterprise observability solutions. The ideal
candidate will possess deep expertise in modern observability platforms, application
performance monitoring, Open Telemetry implementation and cloud technologies, with a
strong focus on improving system reliability, operational efficiency, and user experience.

Key Responsibilities
 Analyze the existing observability solution deployed in Elastic cloud and understand
the gap;
 Document ideal scenario versus existing deployment and recommend the changes
required to bring the Observability solution to improve overall application monitoring
 implement end-to-end observability solutions for distributed and cloud-native
applications; Work with development, Infrastructure and application support team to
streamline the application monitoring using Elastic Cloud
 Develop comprehensive monitoring strategies covering infrastructure, applications,
logs, traces, metrics, and user experience.
 Migrate application monitoring from legacy monitoring platforms (Dynatrace, Splunk,
Prometheus ) to modern observability platforms such as Dynatrace and Elastic.
 Design and implement Elastic-based monitoring architectures, including data
pipelines, storage, APM, dashboards, and advanced analytics.
 Build custom extensions, automated workflows, and synthetic monitoring solutions
using Python and JavaScript.
 Integrate observability platforms with CI/CD pipelines (Jenkins and related tools) to
automate monitoring, alerting, and incident management.
 Configure OpenPipeline, Business Events, anomaly detection, and AI-driven
analytics to improve operational visibility.
 Optimize observability platform licensing, data ingestion, and storage costs while
maintaining monitoring effectiveness.
 Collaborate with Development, Infrastructure, and Operations teams to improve
application reliability, performance, and operational excellence.
 Conduct dashboard reviews, monitoring assessments, and observability maturity
improvements across enterprise applications.
 Support proactive monitoring, root cause analysis, incident response, and continuous
service improvement initiatives.
Required Skills & Qualifications

 9+ years of experience in Application Performance Monitoring (APM), Observability,
and Monitoring Engineering.
 Excellent knowledge about Deploying observably using Open Telemetry framework
 Strong expertise in onboarding application in Elastic Search (www.elastic.co)
 Strong knowledge of:
o Distributed tracing
o Log analytics
o Infrastructure and application monitoring
o Synthetic monitoring
o Real User Monitoring (RUM)
 Experience developing automation using Python and JavaScript.
 Experience integrating monitoring platforms with Jenkins and CI/CD pipelines.
 Hands on experience in implementing Observability for Container based application
 Strong understanding of observability architecture, SRE principles, and cloud-native
monitoring practices.
 Hands-on experience with AWS cloud platforms.
 Knowledge of anomaly detection, Open Pipeline configuration, Business Events, and
AI-assisted observability.
 Strong analytical, troubleshooting, and root cause analysis skills.
 Experience designing scalable enterprise observability architectures.
 Knowledge of license optimization and observability cost management.
 Experience implementing AI-driven observability and automated incident
management.
 Excellent communication, stakeholder management, and cross-functional
collaboration skills.
 Passion for driving operational excellence through automation, proactive monitoring,
and observability best practices.