Site Reliability Engineer

Compunnel

$100K — $120K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • Minimum 5 years in DevOps, SRE, or related fields.
  • Strong hands-on experience with Kubernetes, Docker, and Linux/UNIX.
  • Proficient in CI/CD tools like Jenkins and artifact management.
  • Experience with YAML, Ansible, or equivalent automation frameworks.
  • Knowledge of networking concepts including DNS and TCP/IP.
  • Proven ability to troubleshoot production issues and perform root-cause analysis.
  • Familiarity with Agile methodologies and collaboration tools.

Responsibilities

  • Design and maintain CI/CD pipelines and automated deployment solutions.
  • Deploy and support applications in Kubernetes and container environments.
  • Develop automation scripts with Python and Linux shell scripting.
  • Manage Git repositories, branching strategies, and workflows.
  • Build and maintain Kubernetes configurations and Helm charts.
  • Integrate testing, security scanning, and quality checks in CI/CD pipelines.
  • Troubleshoot complex application and infrastructure issues.

Benefits

  • Participation in after-hours production support and on-call rotation as needed.
  • Opportunities for collaboration across teams including cybersecurity and production support.
  • Potential for personal development and upskilling in new technologies.
  • Engagement in challenging projects that enhance platform reliability and operational resilience.
Full Job Description
Job Summary

We are seeking an experienced DevOps, Kubernetes, and Site Reliability Engineer with 5+ years of relevant experience to design, automate, deploy, and support reliable application platforms and CI/CD capabilities. This role focuses on improving software delivery, platform stability, operational resilience, and production reliability.

Key Responsibilities

  1. De sign, build, and maintain CI/CD pipelines and automated deployment solutions.
  2. Deploy and support applications across Kubernetes and containerized environments.
  3. Develop automation using Python, Linux shell scripting, YAML, and related technologies.
  4. Manage Git/GitHub repositories, branching strategies, pull requests, and automated workflows.
  5. Build and maintain Kubernetes deployment configurations and Helm charts.
  6. Integrate automated testing, security scanning, code-quality checks, and policy controls into CI/CD pipelines.
  7. Troubleshoot complex application, Kubernetes, container, infrastructure, network, and deployment issues.
  8. Investigate production incidents, perform root-cause analysis, and implement corrective solutions.
  9. Apply SRE principles to improve availability, scalability, monitoring, observability, and operational readiness.
  10. Collaborate with application, infrastructure, network, database, cybersecurity, and production support teams.
  11. Participate in after-hours production support and on-call rotation as required.


Required Qualifications

  1. Min imum 5 years of experience in DevOps, SRE, Platform Engineering, Production Engineering, Infrastructure Engineering, or a related field.
  2. Strong hands-on experience with Kubernetes, Docker/Podman, Linux/UNIX, Python, Bash/KornShell, Git, and GitHub.
  3. Experience with CI/CD and artifact-management tools such as Jenkins and Artifactory, or equivalent platforms.
  4. Experience with YAML, Ansible, or similar automation frameworks.
  5. Strong knowledge of networking concepts including DNS, TCP/IP, HTTP/HTTPS, TLS, proxies, firewalls, routing, and load balancing.
  6. Experience supporting production applications, troubleshooting incidents, and performing root-cause analysis.
  7. Understanding of SRE practices, monitoring, incident response, release automation, and deployment strategies.
  8. Experience with automated testing, code-quality tools, security scanning, and CI/CD policy controls.
  9. Experience working in Agile environments and using tools such as Jira.
  10. Strong communication, collaboration, troubleshooting, and prioritization skills.


Preferred Qualifications

  1. Bachelor's degree in Computer Science, Engineering, Information Technology, or a related discipline, or equivalent practical industry experience.
  2. Experience with OpenShift or other managed Kubernetes platforms.
  3. Experience with Azure, AWS, GCP, infrastructure-as-code, Helm, and GitOps.
  4. Knowledge of service mesh, container networking, ingress controllers, and API gateways.
  5. Experience with observability and telemetry platforms, including Grafana.
  6. Knowledge of high availability, disaster recovery, capacity management, and production resiliency.
  7. Experience with relational databases such as DB2, Sybase, or Oracle.
  8. Experience within financial services or another regulated enterprise environment.
  9. Familiarity with secure software supply-chain practices, secrets management, certificate management, and vulnerability remediation.

Similar Jobs

More Jobs at Compunnel

More Information Technology Jobs

Find similar Site Reliability Engineer jobs: