Senior Site Reliability Engineer, AWS/Datacenter Hybrid

Cognitiv

$160K — $210K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • Deep knowledge of AWS infrastructure and networking practices.
  • 10+ years of experience in operations, software engineering, or as an SRE.
  • Working knowledge of modern datacenter practices with support for multi-DC deployments.
  • Proven experience with infrastructure as code methodologies.
  • Proficiency with scripting languages such as Python and Bash.

Responsibilities

  • Design, implement, and maintain infrastructure across the AWS environment.
  • Evaluate existing AWS architecture for scalability and growth.
  • Collaborate with engineering and product teams to define project scopes.
  • Drive service management improvements in deployments and monitoring.
  • Support and maintain co-located datacenter deployments along with SRE team.

Benefits

  • Medical, Dental, and Vision plan for US employees; Extended Health Benefits for Canadian employees.
  • 12 weeks paid parental leave plus 4 additional weeks working from home.
  • Unlimited PTO with a unique Work-From-Anywhere August.
  • Equity options available for all employees.
  • Hybrid work model with daily team lunches.
Full Job Description
We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service management across Cognitiv. Our immediate challenge is to scale and harden our AWS environment as we continue expanding our hybrid cloud footprint. Past that, we are a rapidly growing organization and need to continue to move toward industry best practices. This role demands an experienced engineer with the interest to rapidly learn our environment and help drive our long-term service management roadmap. This role partners closely with our datacenter-focused SRE, so while deep, hands-on AWS expertise is the priority, working familiarity with datacenter operations is important so you can help provide multi-DC coverage when needed. Location: This position will be in our Bellevue, WA office with a hybrid work schedule of 3 days in office (Mon/Tue/Wed) and 2 days remote (Thursday/Friday). Responsibilities • Design, implement, and maintain infrastructure across our AWS environment, serving as the primary owner of our cloud footprint. • Evaluate our existing AWS architecture (compute, networking, security) and ensure we are set up for long-term scalability and growth. • Work across engineering and product teams to scope projects tightly to core business requirements. • Drive engineering-wide efforts to improve company service management around deployments, monitoring, and disaster recovery. • Support and help maintain our co-located datacenter deployments alongside our datacenter-focused SRE, providing coverage as needed. Our Stack • Hosting: AWS and Equinix colocation. • Monitoring: Datadog (migrating off Prometheus). • Infrastructure as Code: Terraform and Ansible. • Compute: Primarily raw EC2 instances and bare-metal machines, with a small footprint of Kubernetes. Requirements • Deep knowledge of AWS infrastructure, networking, and management practices. • 10+ years of experience in operations, software engineering, or as an SRE. • Working knowledge of modern datacenter practices, with the ability to support multi-DC deployments as needed. • Proven experience with infrastructure as code • Proficiency with Python and Bash. • An independent self-starter who looks at the big picture, takes ownership, and independently seeks out new challenges with creative solutions. • A constructive, supportive team player, with good communication and interpersonal skills. Preferred Qualifications • AWS certifications (e.g., Solutions Architect, SysOps Administrator) • Experience with hybrid cloud/on-prem solutions. • Hands-on experience building out datacenters. • Willingness and interest to travel 1-2 times per quarter. Salary: $160,000 - $210,000 USD Base Salary + Equity What We Offer Compensation is based on experience, skills, and other factors. Base salary is just one part of your total rewards at Cognitiv-you'll also receive equity and a comprehensive benefits package. Highlights include: • Medical, Dental and Vision plan for US employees & Extended Health Benefits for Canadian employees • 12 weeks paid parental leave + 4 weeks WFH • Unlimited PTO + Work-From-Anywhere August • Career development with clear advancement paths • Equity for all employees • Hybrid work model & daily team lunch • Health & wellness stipend + cell phone reimbursement • 401(k) & RRSP with employer match • Parking (CA, WA, Vancouver offices) & pre-tax commuter benefits • Employee Assistance Program • Comprehensive onboarding (Cognitiv University) • ...and more!

Similar Jobs

More Jobs at Cognitiv

More Information Technology Jobs

Find similar Senior Site Reliability Engineer, AWS/Datacenter Hybrid jobs: