JOB SUMMARY
We are seeking a highly technical, hands-on Senior Site Reliability Engineer (SRE) / DevOps Engineer to join our engineering organization. This individual will play a critical role in designing, automating, and maintaining our development and production environments while partnering closely with software engineering teams to improve application reliability, scalability, and deployment efficiency. The ideal candidate has a strong software engineering mindset combined with deep infrastructure expertise. They should understand how applications are built and deployed, be passionate about automation, and have experience leveraging modern AI-powered development tools to accelerate engineering workflows.
Key Responsibilities
Design, implement, and maintain CI/CD pipelines using Jenkins and related DevOps tooling.
Partner closely with software engineers to improve build processes, deployment automation, and application reliability.
Develop automation scripts and tooling using Python.
Troubleshoot complex application, infrastructure, and deployment issues across development, test, and production environments.
Manage and optimize Linux-based infrastructure supporting enterprise applications.
Configure and support load balancing technologies including F5 BIG-IP and HAProxy.
Monitor system health, performance, and availability while proactively identifying opportunities for automation and optimization.
Support infrastructure modernization initiatives and implement Infrastructure as Code (IaC) best practices.
Collaborate with networking, infrastructure, and application development teams to ensure highly available and scalable solutions.
Utilize AI-assisted engineering tools (GitHub Copilot, Cursor, ChatGPT, Claude Code, or similar) to improve development productivity, troubleshooting, documentation, and automation.
Participate in production support, incident response, root cause analysis, and continuous improvement initiatives.
Required Qualifications
8+ years of experience in Site Reliability Engineering, DevOps Engineering, Platform Engineering, or Infrastructure Engineering.
Strong Python scripting and automation experience.
Extensive experience with Jenkins and CI/CD pipeline development.
Deep understanding of software development lifecycles, application build processes, and deployment methodologies.
Experience supporting Java or other enterprise application environments.
Strong Linux systems administration experience.
Experience with source control systems such as Git.
Strong troubleshooting skills across infrastructure, networking, and applications.
Experience working in Agile software development environments.
Infrastructure & Networking Experience
Strong understanding of networking fundamentals, including: TCP/IP DNS HTTP/HTTPS SSL/TLS Routing Firewalls Load balancing
Hands-on experience with: F5 BIG-IP HAProxy Reverse proxy technologies
Experience supporting highly available production environments.
Preferred Qualifications
Kubernetes and container orchestration experience.
Docker.
Terraform or other Infrastructure as Code tools.
AWS, Azure, or GCP cloud experience.
Monitoring and observability tools such as Splunk, Prometheus, Grafana, Datadog, or Dynatrace.
Experience implementing SRE best practices, reliability engineering, and production automation.
Preferred Characteristics
Extremely hands-on technical contributor.
Strong collaboration and communication skills.
Automation-first mindset.
Passion for continuous improvement and engineering excellence.
Comfortable working across development, infrastructure, networking, and operations teams.
Curious about emerging technologies and experienced using AI-powered engineering tools to improve productivity and software delivery.