OSS Telemetry Observability Engineer

E-Space

$160K — $200K *
Telecommunications & Hardware
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 10+ years of experience in telecom or similar roles
  • Expertise in monitoring tools such as Prometheus, Kibana, Grafana, ELK Stack
  • Strong knowledge of distributed systems and microservices architecture
  • Basic knowledge of Python, Linux, and Bash
  • Experience with cloud platforms like AWS, GCP, or Azure
  • Understanding of SLO/SLI/SLA concepts
  • Excellent written and verbal communication skills

Responsibilities

  • Design and implement observability strategies across distributed systems
  • Deploy and maintain monitoring solutions using relevant tools
  • Develop automated alerting systems with AI-powered anomaly detection
  • Create and maintain real-time system visibility dashboards
  • Implement distributed tracing and log aggregation solutions
  • Provide technical support to NOC engineers in issue resolution
  • Develop KPIs and metrics to monitor mobile network performance
  • Create training material related to KPIs and dashboards

Benefits

  • Opportunity to work in fast-paced, high-impact environments
  • Chance to collaborate with a highly experienced engineering team
  • Flexibility to work during critical periods, including evenings and weekends
  • Innovative work redefining satellite technology
  • Potential for personal and professional growth in a cutting-edge field
Full Job Description
We9re looking for OSS Telemetry Observability Engineer to grow our team. This role will be hands on and you will work with our core engineering team and customer support escalation groups, to monitor and properly escalate real-time core network or infrastructure issues that could affect customers.

Your role will be within a dynamic and busy Satellite communications company and will work alongside individuals who are highly experienced and experts in their field. You will maintain and improve OSS network and provide technical support, fault investigation and troubleshooting of all issues on the network. You will keep the performance of our telecom network at optimum levels by ensuring that network/application problems are detected and corrected according to agreed KPI and SLA - and your work will directly shape the reliability and scalability of the networks our partners depend on.

What you will do:

    • Design and implement comprehensive observability strategies across distributed systems

    • Deploy and maintain monitoring solutions using tools like Prometheus, Kibana, ELK Stack, Grafana

    • Develop automated alerting systems with AI-powered anomaly detection

    • Create and maintain dashboards for real-time system visibility

    • Implement distributed tracing and log aggregation solutions

    • Creating on-demand dashboards to monitor specific metrics or to validate restoring of services after failures

    • Technical support to NOC engineers to analyze and solve issues.

    • Development of KPIs and other metrics to monitor key aspects in mobile network performance

    • Develop and maintain dashboards, alerts, and visualizations to track key performance metrics.

    • Development of training material related to KPIs / Dashboards

    • Ability to troubleshoot, document, and assess proper escalation channel and team or group.


What you bring to this role:

    • 10+ years of experience working in a telecom environment or similar roles

    • Expertise in monitoring tools (Prometheus, Kibana, Grafana, ELK Stack)

    • Strong knowledge of distributed systems and microservices architecture

    • Basic Python / Linux / Bash knowledge

    • Experience with cloud platforms (AWS, GCP, Azure)

    • Understanding of SLO/SLI/SLA concepts

    • Excellent written and verbal communication skills

    • Excellent interpersonal skills


Bonus points for the following:

    • Experience with 3GPP NTN core adaptations (Rel-17/18): satellite access node (SAN) architecture, timer and window parameter adaptation for high-latency links, or UPF placement in non-terrestrial topologies.

    • Experience in a startup or fast-paced R&D environment where architectural decisions move at product speed.


Additional Requirements

This is a fast-paced, high-impact environment - flexibility to occasionally work extended hours or weekends during critical periods is expected.

This is a full time, exempt position, based out of our Arlington, TX office.

The total compensation packaged will be determined by various factors such as your relevant job-related knowledge, skills, and experience.

We are redefining how satellites are designed, manufactured and used-so we9re looking for candidates with passion, deep knowledge and direct experience on LEO satellite component development, design and in-orbit activities. If that9s your experience - then we9ll be immediately wow-ed.

$160,000 - $200,000 a year

This is a full time, exempt position, based out of our Saratoga, CA office.

The total compensation packaged will be determined by various factors such as your relevant job-related knowledge, skills, and experience.

We are redefining how satellites are designed, manufactured and used-so we9re looking for candidates with passion, deep knowledge and direct experience on LEO satellite component development, design and in-orbit activities. If that9s your experience - then we9ll be immediately wow-ed.

Similar Jobs

More Jobs at E-Space

More Telecommunications & Hardware Jobs

Find similar OSS Telemetry Observability Engineer jobs: