Ensono

Expert Automation & Observability Engineer

Ensono$140K — $180K *
US-AnywhereRemote in United States
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 12+ years of total IT experience, including 5-7 years as a Lead Architect, SRE, or Principal Observability Engineer.
  • Proven experience migrating from legacy monitoring to automated observability.
  • Hands-on expertise in scalable telemetry pipelines and time-series databases.
  • Experience with leading First-of-a-Kind (FOAK) implementations and vendor transitions.
  • Preferred certifications include CKA, AWS or Azure Cloud Architect, or APM/Observability vendor certifications.

Responsibilities

  • Architect and govern a unified observability framework.
  • Lead First-of-a-Kind technology implementations for secure, production-ready patterns.
  • Define standards for telemetry pipelines and observability cost management.
  • Govern Service Level Indicators (SLIs), Objectives (SLOs), and error budgets.
  • Lead major incident resolution and conduct Root Cause Analysis (RCA).
  • Drive Observability-as-Code and automate infrastructure workflows.
  • Design observability for Docker, Kubernetes, and multi-cloud environments.

Benefits

  • Unlimited Paid Days Off
  • Three health plan options
  • 401k with company match
  • Dental, vision, short/long-term disability coverage
  • Family Forming Benefit for fertility coverage and adoption reimbursement
  • Paid childbearing and paternal leave
  • Education Reimbursement and Student Loan Assistance
  • Sabbatical leave
  • Wellness program
  • Flexible work schedule
Full Job Description
About the role and what you'll be doing:

We are seeking an Expert Observability Engineer to serve as the strategic technical lead and architect for our enterprise Observability, APM, and Telemetry ecosystems. You will lead the transformation from decentralized, reactive monitoring to a unified, automated, and proactive observability framework. Operating across hybrid cloud, Kubernetes, and legacy environments, you will design scalable architectures, drive Site Reliability Engineering (SRE) practices, and lead First-of-a-Kind (FOAK) technology implementations to ensure maximum service reliability.

We want all new Associates to succeed in their roles at Ensono. That's why we've outlined the job requirements below. To be considered for this role, it's important that you meet all Required Qualifications. If you do not meet all of the Preferred Qualifications, we still encourage you to apply.

Core Responsibilities

1. Enterprise Architecture & Strategy
  • Architect and govern a unified observability framework covering metrics, logs, traces, and events using IBM Instana, Grafana, OpenTelemetry, Telegraf, and InfluxDB.
  • Lead First-of-a-Kind (FOAK) implementations-evaluating new observability tech and converting them into secure, repeatable, production-ready patterns.
  • Define enterprise standards for telemetry pipelines, data retention, high-cardinality controls, and observability cost management.

2. SRE & Service Reliability
  • Define and govern Service Level Indicators (SLIs), Objectives (SLOs), and error budgets.
  • Serve as the senior technical escalation point, leading major P1/P2 incident war rooms and conducting evidence-based Root Cause Analysis (RCA).
  • Drastically reduce MTTD/MTTR and alert noise through event correlation, dynamic thresholds, and dependency mapping.

3. Platform Engineering & Automation
  • Drive Observability-as-Code and infrastructure automation using Ansible, Terraform, Python, and GitOps.
  • Automate the deployment, configuration, and self-healing workflows for monitoring agents and telemetry collectors.
  • Integrate observability platforms seamlessly with ITSM (ServiceNow), Netcool, and CI/CD pipelines.

4. Cloud-Native & Kubernetes Observability
  • Design deep observability for Docker, Kubernetes, microservices, and multi-cloud environments (Azure/AWS/GCP).
  • Correlate application APM telemetry with Kubernetes control planes, pods, nodes, and infrastructure dependencies.
  • Ensure secure-by-design telemetry pipelines (RBAC, TLS, secrets management, and image scanning).

5. Technical Leadership & Transition Management
  • Lead complex Knowledge Transfer (KT) programs, vendor transitions, and operational readiness handovers for global 24x7 teams.
  • Mentor cross-functional engineering teams and influence enterprise technology roadmaps.


Required Technical Stack

  • Observability & APM: IBM Instana, Grafana (Enterprise & Alloy), Prometheus, OpenTelemetry, Telegraf, InfluxDB.
  • Legacy/Traditional Monitoring: SolarWinds, Netcool, Elastic/Splunk.
  • Cloud & Containerization: Kubernetes, Docker, OpenShift, AWS/Azure/GCP.
  • Infrastructure: Linux (RHEL), Windows Server, VMware, Citrix VDI, load balancers, and edge proxies.
  • Automation & DevOps: Ansible, Terraform, Python, Bash, Webhooks, CI/CD (GitHub Actions/GitLab/Jenkins).
  • ITSM/Operations: ServiceNow, ITIL 4, advanced Major Incident Management.

Qualifications & Experience
  • 12+ years of total IT experience, with a minimum of 5 to 7 years functioning as a Lead Architect, SRE, or Principal Observability Engineer in a massive enterprise environment.
  • Proven track record of migrating organizations from legacy monitoring to proactive, automated observability platforms.
  • Hands-on expertise in building scalable, secure telemetry pipelines and time-series databases.
  • Extensive experience leading FOAK rollouts and complex vendor/operations transition (KT) programs.
  • Preferred Certifications: CKA (Certified Kubernetes Administrator), Cloud Architect (AWS/Azure), or specific APM/Observability vendor certifications.


We are a client-facing business, but we do encourage clients to allow us to work remotely most of the time so if you are not required to be on a client site, you can choose to work from home or in our Ensono offices.

Some of our benefits include:
  • Unlimited Paid Days Off
  • Three health plan options
  • 401k with company match
  • Eligibility for dental, vision, short and long-term disability, life and AD&D coverage, and flexible spending accounts
  • Family Forming Benefit including fertility coverage and adoption/surrogacy reimbursement
  • Paid childbearing and paternal leave
  • Education Reimbursement, Student Loan Assistance or 529 College Funding
  • Sabbatical leave
  • Wellness program
  • Flexible work schedule


As of the date of this posting, a good faith estimate of the current pay scale for this role is $140,000 to $180,000 annually based on a full-time schedule. Please note that placement in the range may vary based on numerous factors including but not limited to skills, experience, internal equity, and business needs. In addition to base salary, other compensation programs, depending on eligibility, include an annual bonus plan based on company and individual performance and an equity grant under our Associate Equity Appreciation Program.

About Ensono

Ensono is a hybrid IT services provider that helps clients transform their IT infrastructure, operations and service delivery to improve business agility, accelerate growth and optimize their IT expense management. Ensono delivers technology solutions across all platforms and for all types of IT environments, helping clients to drive innovation and enhance their customer experiences. The company has over 2,000 associates and is headquartered in Chicago, Illinois, with operations in North America, Europe, and Asia.
Learn more about Ensono
Size
2,000 employees
Industry
Net Income
$20 million
Founded
2000
5 Year Trend
+5%
Revenue
$550 million

Similar Jobs

More Jobs at Ensono

More Information Technology Jobs

Find similar Expert Automation & Observability Engineer jobs: