NextRaiseNextRaiseFind jobs
Sign inSign up free
Jobs / Site Reliability Engineer in United States of America
2 days ago
Apply with autofill
Apply with autofill
Ensono·2 days ago
2 days ago

Expert Automation & Observability Engineer

Remote - United StatesRemoteMid · 5-7 yearsSite Reliability Engineer

Sign up free to see how well your resume matches this role.

Boost your chances at ensono

How you compare FREE

?
Your scoreYour score: not yet known
→
44
Top 10%Top 10%: 44 out of 100

Top 10% of NextRaise users matched against Site Reliability Engineer roles in United States.

Must-have skills for this role

  • ibm instana
  • grafana
  • opentelemetry
  • kubernetes

PDF or DOCX · no account needed

Apply faster with autofill FREEensono uses Greenhouse - autofill it instead of retyping.careers.example.com/applyAutofillingFull namePriya SharmaEmailpriya.sharma@example.comPhone+49 30 1234567LocationBerlGet the extension

What you'll do

  • Architect and govern a unified observability framework covering metrics, logs, traces, and events using IBM Instana, Grafana, OpenTelemetry, Telegraf, and InfluxDB.
  • Lead First-of-a-Kind (FOAK) implementations—evaluating new observability tech and converting them into secure, repeatable, production-ready patterns.
  • Define enterprise standards for telemetry pipelines, data retention, high-cardinality controls, and observability cost management.
  • Define and govern Service Level Indicators (SLIs), Objectives (SLOs), and error budgets.
  • Serve as the senior technical escalation point, leading major P1/P2 incident war rooms and conducting evidence-based Root Cause Analysis (RCA).
  • Drastically reduce MTTD/MTTR and alert noise through event correlation, dynamic thresholds, and dependency mapping.
  • Drive Observability-as-Code and infrastructure automation using Ansible, Terraform, Python, and GitOps.
  • Automate the deployment, configuration, and self-healing workflows for monitoring agents and telemetry collectors.
  • Integrate observability platforms seamlessly with ITSM (ServiceNow), Netcool, and CI/CD pipelines.
  • Design deep observability for Docker, Kubernetes, microservices, and multi-cloud environments (Azure/AWS/GCP).
  • Correlate application APM telemetry with Kubernetes control planes, pods, nodes, and infrastructure dependencies.
  • Ensure secure-by-design telemetry pipelines (RBAC, TLS, secrets management, and image scanning).

What they're looking for

  • 12+ years of total IT experience, with a minimum of 5 to 7 years functioning as a Lead Architect, SRE, or Principal Observability Engineer in a massive enterprise environment.
  • Proven track record of migrating organizations from legacy monitoring to proactive, automated observability platforms.
  • Hands-on expertise in building scalable, secure telemetry pipelines and time-series databases.
  • Extensive experience leading FOAK rollouts and complex vendor/operations transition (KT) programs.

Nice to have

  • CKA (Certified Kubernetes Administrator), Cloud Architect (AWS/Azure), or specific APM/Observability vendor certifications.

Summarised by NextRaise from the employer’s description, which follows in full below.

Full description from employer

At Ensono, our Purpose is to be a relentless ally, disrupting the status quo and unleashing our clients to Do Great Things!  We enable our clients to achieve key business outcomes that reshape how our world runs. As an expert technology adviser and managed service provider with cross-platform certifications, Ensono empowers our clients to keep up with continuous change and embrace innovation.

 

We can Do Great Things because we have great Associates. The Ensono Core Values unify our diverse talents and are woven into how we do business. These five traits are the key to achieving our purpose:

 

Honesty, Reliability, Curiosity, Collaboration, and Passion.

 

About the role and what you'll be doing: 

 We are seeking an Expert Observability Engineer to serve as the strategic technical lead and architect for our enterprise Observability, APM, and Telemetry ecosystems. You will lead the transformation from decentralized, reactive monitoring to a unified, automated, and proactive observability framework. Operating across hybrid cloud, Kubernetes, and legacy environments, you will design scalable architectures, drive Site Reliability Engineering (SRE) practices, and lead First-of-a-Kind (FOAK) technology implementations to ensure maximum service reliability.

 

We want all new Associates to succeed in their roles at Ensono. That's why we've outlined the job requirements below. To be considered for this role, it's important that you meet all Required Qualifications. If you do not meet all of the Preferred Qualifications, we still encourage you to apply. 


Core Responsibilities

1. Enterprise Architecture & Strategy

  • Architect and govern a unified observability framework covering metrics, logs, traces, and events using IBM Instana, Grafana, OpenTelemetry, Telegraf, and InfluxDB.
  • Lead First-of-a-Kind (FOAK) implementations—evaluating new observability tech and converting them into secure, repeatable, production-ready patterns.
  • Define enterprise standards for telemetry pipelines, data retention, high-cardinality controls, and observability cost management.

2. SRE & Service Reliability

  • Define and govern Service Level Indicators (SLIs), Objectives (SLOs), and error budgets.
  • Serve as the senior technical escalation point, leading major P1/P2 incident war rooms and conducting evidence-based Root Cause Analysis (RCA).
  • Drastically reduce MTTD/MTTR and alert noise through event correlation, dynamic thresholds, and dependency mapping.

3. Platform Engineering & Automation

  • Drive Observability-as-Code and infrastructure automation using Ansible, Terraform, Python, and GitOps.
  • Automate the deployment, configuration, and self-healing workflows for monitoring agents and telemetry collectors.
  • Integrate observability platforms seamlessly with ITSM (ServiceNow), Netcool, and CI/CD pipelines.

4. Cloud-Native & Kubernetes Observability

  • Design deep observability for Docker, Kubernetes, microservices, and multi-cloud environments (Azure/AWS/GCP).
  • Correlate application APM telemetry with Kubernetes control planes, pods, nodes, and infrastructure dependencies.
  • Ensure secure-by-design telemetry pipelines (RBAC, TLS, secrets management, and image scanning).

5. Technical Leadership & Transition Management

  • Lead complex Knowledge Transfer (KT) programs, vendor transitions, and operational readiness handovers for global 24x7 teams.
  • Mentor cross-functional engineering teams and influence enterprise technology roadmaps.

 

Required Technical Stack

 

  • Observability & APM: IBM Instana, Grafana (Enterprise & Alloy), Prometheus, OpenTelemetry, Telegraf, InfluxDB.
  • Legacy/Traditional Monitoring: SolarWinds, Netcool, Elastic/Splunk.
  • Cloud & Containerization: Kubernetes, Docker, OpenShift, AWS/Azure/GCP.
  • Infrastructure: Linux (RHEL), Windows Server, VMware, Citrix VDI, load balancers, and edge proxies.
  • Automation & DevOps: Ansible, Terraform, Python, Bash, Webhooks, CI/CD (GitHub Actions/GitLab/Jenkins).
  • ITSM/Operations: ServiceNow, ITIL 4, advanced Major Incident Management.

Qualifications & Experience

  • 12+ years of total IT experience, with a minimum of 5 to 7 years functioning as a Lead Architect, SRE, or Principal Observability Engineer in a massive enterprise environment.
  • Proven track record of migrating organizations from legacy monitoring to proactive, automated observability platforms.
  • Hands-on expertise in building scalable, secure telemetry pipelines and time-series databases.
  • Extensive experience leading FOAK rollouts and complex vendor/operations transition (KT) programs.
  • Preferred Certifications: CKA (Certified Kubernetes Administrator), Cloud Architect (AWS/Azure), or specific APM/Observability vendor certifications.

 

Why Ensono?

 

Ensono is a place to make better happen – for our clients and for your career. You can do great things through innovation or collaboration, by learning or volunteering, or to promote diversity and inclusion. You can do great things for your own health or for a healthier planet. Whatever it means to you to do great things we want Ensono to be the place you can do it. 

 

We are a client-facing business, but we do encourage clients to allow us to work remotely most of the time so if you are not required to be on a client site, you can choose to work from home or in our Ensono offices.

 

 

Some of our benefits include:

  • Unlimited Paid Days Off
  • Three health plan options
  • 401k with company match
  • Eligibility for dental, vision, short and long-term disability, life and AD&D coverage, and flexible spending accounts
  • Family Forming Benefit including fertility coverage and adoption/surrogacy reimbursement
  • Paid childbearing and paternal leave
  • Education Reimbursement, Student Loan Assistance or 529 College Funding
  • Sabbatical leave
  • Wellness program
  • Flexible work schedule

 

As of the date of this posting, a good faith estimate of the current pay scale for this role is $140,000 to $180,000 annually based on a full-time schedule. Please note that placement in the range may vary based on numerous factors including but not limited to skills, experience, internal equity, and business needs. In addition to base salary, other compensation programs, depending on eligibility, include an annual bonus plan based on company and individual performance and an equity grant under our Associate Equity Appreciation Program.

 

 

Ensono is an Equal Opportunity/Affirmative Action employer. We are committed to providing equal employment to our Associates and building a diverse and inclusive workforce. All qualified applicants will be considered without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability, or other legally protected basis, in accordance with applicable law.

 

Pay transparency nondiscrimination statement/posting OFCCP’s pay transparency policy can be found on OFCCP’s website.

 

If you need accommodation at any point during the application or interview process, please let your recruiter know or email USTalentAcquisition@ensono.com.

Company

Ensono
Remote - United States

Company facts come from this company's own listings. We only show what the postings themselves carry.

Sourced from Ensono's careers site·first seen 18 Sept 2026·last verified 18 Sept 2026·How we source jobs

Similar jobs

  • Sr. Staff Site Reliability Engineer at earlywarningScottsdale, United States of America–match not yet calculated
  • Silicon Photonics Quality & Reliability Engineer at IntelAlbuquerque, United States of America–match not yet calculated
  • Site Reliability Engineer II at JPMorgan ChaseTampa, United States of America–match not yet calculated
  • Lead Site Reliability Engineer at JPMorgan ChaseOH, United States–match not yet calculated
  • Reliability Engineer at Finning InternationalElkford, United States of America–match not yet calculated

Browse more jobs

  • Site Reliability Engineer jobs in United States
  • Systems Engineer jobs in United States
  • Network Engineer jobs in United States
  • Platform Engineer jobs in United States
  • Site Reliability Engineer jobs in India
  • Site Reliability Engineer jobs in United Kingdom