Our New TeammateThis role with our Central Operations team provides direct ownership of the Linux infrastructure, the automation practice, and the day-to-day direction of two engineers on your team. You will work directly alongside the Central Operations Manager as a technical leadership partner, helping shape how the team operates, where it can improve, and how it delivers support to internal teams and external stakeholders.
The systems you support underpin transportation and public safety operations that require reliable infrastructure.
You can expect to spend your time accomplishing the following:
- 40% of the time on Objective 1: Infrastructure Automation and Systems Operations
- 30% of the time on Objective 2: Observability and Monitoring
- 20% of the time on Objective 3: Team Leadership and Development
- 10% of the time on Objective 4: Standards Documentation and Collaboration
Job Responsibilities - What to Expect- Develop and improve operational tooling, workflows, and automation that reduce manual effort and improve the reliability and scalability of day-to-day systems operations.
- Define and drive GitOps adoption across the entire Operations team, building the habits and tooling required.
- Use Ansible as the primary configuration management tool, including playbook and role development, inventory management, and automated remediation workflows.
- Bash and Python scripting to build out the broader automation practice, covering system deployment, scheduled validation runs, and testing against defined baselines.
- Administer and tune KVM and CloudStack environments to meet performance and capacity needs
- Work alongside the Skyline Platform Team to support Kubernetes environments as a secondary resource.
- The expectation is a solid working understanding of the environment, how workloads are structured, and how deployments are managed.
- Own and improve the CMDB practice end to end, from automated discovery, auditing and validation through change tracking and reporting
- Storage architecture awareness and backup support for our storage team
- Maintain a working understanding of storage architecture and how different storage types interact with the systems you support
- Ongoing improvement of monitoring and alerting practices to support early detection of issues before they become service-impacting. This means building toward meaningful coverage, well-tuned alerts, and monitoring standards that reduce noise and surface what matters.
- Ensure telemetry is high quality, consistent, and operationally useful across environments through the development and refinement of dashboards, alerts, and supporting configurations.
- Partner with stakeholders to review observability needs, improve visibility into system and service health, and translate operational requirements into effective monitoring solutions.
- Support the ongoing maintenance, enhancement, and reliability of observability systems that serve both internal operational needs and broader engineering use cases.
- Define tasking and support for the two engineers on your team
- Serve as a mentor, escalation point, and peer reviewer for the broader Operations team, with a focus on automation practices and configuration management
- Develop and maintain clear operational documentation, including runbooks, standards, and implementation guidance for automation and observability systems
Your Knowledge & Expertise- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field
- Red Hat Certified Engineer (RHCE) preferred
- ITIL Foundation preferred
- 5 or more years of advanced Linux system administration in an enterprise or managed services environment
- Strong scripting and automation skills in Bash, Python, or both
- Experience with Ansible for configuration management and automation
- Familiarity with CMDB practices and configuration tracking
- Solid understanding of TCP/IP networking, routing, and firewall concepts
- Working knowledge of Grafana, OpenNMS, and Zabbix
- Experience with Infrastructure as Code tools such as Terraform or OpenTofu
- KVM virtualization Experience
- Experience with CloudStack administration and cloud orchestration
- Kubernetes administration experience
- Experience with Linux security hardening and compliance practices
- Comfort leading a small team and driving process adoption
Benefits Included:- Medical Insurance
- Vision Insurance
- Dental Insurance
- FSA Plan
- Paid Time Off
- 401K Retirement Savings Plan
- Training & Tuition Assistance
- Disability & Life Insurance