Site Reliability Engineer

Stefanini$126K — $137K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 4+ years of experience in software development
  • Hands-on experience with BigQuery and Dynatrace
  • Experience with Google Cloud Platform (GCP)
  • Proficient in monitoring/observability tools, especially Dynatrace or similar
  • Familiar with ITSM tools like ServiceNow

Responsibilities

  • Collaborate with Infrastructure teams to automate routine tasks
  • Monitor and manage production environments to proactively resolve issues
  • Build advanced tooling for access monitoring and reliability across data centers
  • Engage with engineering to enhance on-call efficiency and incident management
  • Perform capacity planning to support increasing demand
  • Maintain monitoring and alerting systems for proactive health checks
  • Continuously optimize system performance, stability, and security through data analysis
  • Create and maintain documentation and diagrams for knowledge sharing

Benefits

  • Opportunity to work in a hands-on SRE role with cloud infrastructure and services
  • Engagement in advanced tooling and observability projects
  • Focus on impact at scale within a GCP environment
  • Collaborative atmosphere with Infrastructure and engineering teams
  • Opportunities for professional development and knowledge sharing
Full Job Description
Details:

Stefanini is looking for a Site Reliability Engineer, Dearborn, MI (Onsite)



Seeking an experienced SRE who is responsible for ensuring availability, reliability and performance of cloud and network systems and services by automating routine manual tasks

This position is focused on observability, monitoring, and technical consulting across our GCP-based data platforms. working hands-on with cloud infrastructure, BigQuery workloads, CI/CD pipelines, and enterprise monitoring tools to keep critical systems healthy, performant, and reliable at scale.

Responsibilities
  • Collaborate with Infrastructure teams in implementing critical solutions by automating routine tasks
  • Monitor and manage production environments, proactively identifying and resolving issues.
  • Participate in building advanced tooling for system access monitoring, log session recording, and administration of reliability across multiple geographically distributed data centers.
  • Engage with engineering teams to improve on-call efficiency, drive incident management and post-mortem analysis.
  • Perform capacity planning and optimization to support growing demand and traffic patterns.
  • Maintaining, monitoring and alerting systems for proactive system health checks.
  • Continuously improve system performance, stability, and security through data-driven analysis and optimization.
  • Facilitate knowledge sharing by creating and maintaining comprehensive documentation & diagrams.


Job Requirements

Details:

Experience Required
  • 4+ years of experience in development
  • Hands on experience with Big Query & Dynatrace
  • Hands-on experience with Google Cloud Platform (GCP).
  • Proficiency with monitoring/observability tools, ideally Dynatrace (or comparable, e.g., Datadog, New Relic).
  • Familiarity with ITSM tools such as ServiceNow (incident, problem, change management)


Experience Preferred
  • Familiarity with the use of AI tools - agents, skills, LLMs, copilot. Experience defining and tracking SLAs/SLOs/SLIs
  • Experience with GCP Cloud Run, Python, Troubleshooting (Problem Solving)


Education Required

  • Bachelor's Degree


**Listed salary ranges may vary based on experience, qualifications, and local market. Also, some positions may include bonuses or other incentives***

Stefanini takes pride in hiring top talent and developing relationships with our future employees. Our talent acquisition teams will never make an offer of employment without having a phone conversation with you. Those face-to-face conversations will involve a description of the job for which you have applied. We also speak with you about the process, including interviews and job offers.

#LI-AK3

#LI-ONSITE

Pay Range:

$ 61.00 - $ 66.00

Similar Jobs

More Jobs at Stefanini

More Information Technology Jobs

Find similar Site Reliability Engineer jobs: