Senior Site Reliability Engineer, Platform Infrastructure

Cricut, Inc. • $120K — $145K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 4+ years of experience in software engineering or site reliability engineering.
  • Fluency in AI-assisted development tools (e.g., Claude Code, GitHub Copilot) with practical prompt engineering skills.
  • Strong software engineering background, ideally with backend microservices experience (.NET preferred).
  • Deep expertise with AWS core services (e.g., EC2, S3, RDS).
  • Proven experience using Infrastructure-as-Code tools like Terraform or CloudFormation.
  • Solid understanding of SRE principles and experience in improving site reliability.
  • Experience with modern observability and monitoring platforms, ideally Datadog and OpenTelemetry.

Responsibilities

  • Own architecture, reliability, and scalability of critical AWS infrastructure.
  • Partner with Software and Lead Engineers to shape the technical strategy for platform infrastructure.
  • Drive best practices in security, cost management, and scalability within the AWS environment.
  • Champion and expand the 'infrastructure-as-code' philosophy across the organization.
  • Utilize AI-assisted development tools to accelerate delivery and evaluate AI/ML infrastructure.
  • Oversee production monitoring, incident response, and continuous improvement processes.
  • Develop and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs) for critical production systems.

Benefits

  • Medical, Dental, and Vision coverage.
  • 401(k) match.
  • Generous PTO and tuition reimbursement.
  • Yearly lifestyle stipend for wellness and passions.
  • Exclusive employee discounts.
Full Job Description
Job Description

We're looking for a Senior Site Reliability Engineer, Platform Infrastructure to take hands-on technical ownership of the architecture, reliability, and scalability of our entire AWS infrastructure. Reporting to the Engineering Manager, Platform Infrastructure & SRE, you'll set technical direction, review designs, and raise the bar for reliability engineering across a growing and globally distributed engineering organization.

This is a senior individual contributor role. You'll work side by side with our onsite SRE team, Software Engineers, and other Lead Engineers to support seamless 24/7 reliability. It's ideal for an AI-forward engineer with a strong software engineering background who uses AI-assisted development tools to move faster, has a passion for infrastructure-as-code, and a proven track record of mentoring engineers to build highly reliable, scalable, and performant systems.

Key Responsibilities
  • Own the architecture, reliability, and scalability of critical AWS infrastructure, working hands-on across the full stack.
  • Partner with Software Engineers and other Lead Engineers to shape the roadmap and technical strategy for Cricut's platform infrastructure.
  • Take ownership of our AWS environment, driving best practices in security, cost management, and scalability.
  • Champion and expand our "infrastructure-as-code" philosophy across the organization.
  • Use AI-assisted development tools (e.g., Claude Code, GitHub Copilot) to accelerate delivery, applying prompt engineering and context management practices to get reliable results, and evaluate AI/ML infrastructure (e.g., model serving, vector databases, LLM tooling) as it becomes part of the platform.
  • Oversee production monitoring, incident response, and blameless post-mortem processes to continuously improve system reliability.
  • Develop and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs) for critical production systems.
  • Act as a key consultant for our feature-focused pillar and pod teams, ensuring they have the infrastructure resources and support required to deliver their projects successfully.
  • Mentor software engineers who have an affinity for infrastructure, helping them grow their skills in reliability engineering.
  • Collaborate closely with the onsite SRE team, sharing the on-call rotation to ensure seamless 24/7 reliability coverage.


Qualifications
  • 4+ years of experience in a software engineering or site reliability engineering role.
  • An AI-forward mindset: fluency with AI-assisted development tools (e.g., Claude Code, GitHub Copilot, Cursor) to accelerate delivery, practical prompt engineering skills, and experience managing context (e.g., structuring prompts, memory, and retrieved data) to keep LLM-based workflows accurate and reliable, plus hands-on exposure to AI/ML infrastructure (e.g., model serving, vector databases, LLM operations).
  • A strong background in software engineering, ideally with experience in backend microservices (.NET is a strong plus).
  • Deep, hands-on expertise with AWS and its core services (e.g., EC2, S3, RDS, Kinesis, VPC, IAM).
  • Proven experience building and managing infrastructure with Infrastructure-as-Code (IaC) tools like Terraform or CloudFormation.
  • A solid understanding of SRE principles and a proven track record of improving site reliability.
  • Experience with modern observability, monitoring, and logging platforms, ideally Datadog and OpenTelemetry.
  • A Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent industry experience.

Soft Skills
  • Communicates complex technical concepts with clarity-written and verbal-to diverse audiences across engineering, product, and leadership.
  • Mentors with intent, with a demonstrated ability to develop engineers and help them grow in their careers.
  • Influences without authority, building genuine alignment on reliability and infrastructure standards across teams that don't report to you.
  • Stays calm and decisive under production pressure, leading incident response and blameless postmortems that turn outages into durable fixes.
  • Stays genuinely curious, tracking advances in cloud infrastructure, reliability engineering, and AI tooling, and pulls the best of what's new into the team's everyday practice.


Additional Information

We've Got You Covered

At Cricut, we take care of our people. Enjoy competitive Medical, Dental, and Vision coverage, a 401(k) match, generous PTO, tuition reimbursement, and a yearly lifestyle stipend to support your wellness and passions. You'll also receive exclusive employee discounts-and best of all, you'll be surrounded by some of the most talented, creative, and curious minds out there.

A Quick Note Before You Apply...

Cricut is in an exciting chapter of transformation. We're evolving fast-refining our strategy, growing our teams, and raising the bar across everything we do. This is an incredible opportunity for the right kind of person-but it's not for everyone.

We're looking for A-players-people who thrive in dynamic environments, turn challenges into momentum, and consistently deliver their best work. If that sounds like you, read on.

Here's what makes someone a great fit for this role (and for this moment at Cricut):
  • You move with urgency. You don't wait for perfect clarity to act-you start, learn, and adjust.
  • You set high standards. You take ownership, deliver quality, and hold yourself accountable.
  • You stay focused when things move fast. You prioritize what matters most and tune out the noise.
  • You collaborate like a pro. You elevate others, communicate clearly, and bring a low-ego, high-output energy.
  • You embrace AI as part of your toolkit. From idea exploration to data analysis and creative problem-solving, you leverage AI to accelerate innovation and amplify impact-because technology and creativity go hand-in-hand here.

One More Thing (It's a Big One)

This role is in-office at least 4-5 days per week. We believe real collaboration, innovation, and culture are built face-to-face. If you're energized by working alongside smart, kind, creative people-and love those hallway conversations that spark the next great idea-you'll feel right at home.

If you're looking for a fully remote role, this may not be the right fit. But if you're excited by challenge, purpose, and building something better-let's make something amazing together.

Relocation Statement:
• This position is eligible for relocation assistance.

What to Do Next: Please attach your resume, cover letter and/or include links to your portfolio or other social presence. If you want to show your super powers in other ways - include that information too. You can be sure that Cricut® is an employer who values individuality, equality and diversity, so tell us what you're all about. If you are a Maker or a DIY enthusiast, whether you think you are a good one or not, we would love to hear about it when you send us your information.

About Cricut, Inc.

Cricut, Inc. is a technology company that specializes in creating innovative crafting products. The company's flagship product is the Cricut cutting machine, which allows users to create intricate designs and shapes for a variety of crafting projects. Cricut also offers a range of accessories and materials to complement its cutting machines, as well as software and apps to help users design and customize their projects. Founded in 1969, Cricut has been at the forefront of the crafting industry for over 50 years, and continues to innovate and inspire crafters around the world.
Learn more about Cricut, Inc.
Size
1,100 employees
Market Cap
$1.9 billion
Industry

Similar Jobs

More Jobs at Cricut, Inc.

More Information Technology Jobs

Find similar Senior Site Reliability Engineer, Platform Infrastructure jobs: