Mizuho Financial

Site Reliability Engineer

Mizuho Financial$111K — $160K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or related field.
  • Experience as a Site Reliability Engineer or similar role.
  • Proficiency in automation tools like Ansible, Terraform, or Jenkins.
  • Advanced skills in monitoring and visualization, particularly with Grafana.
  • Hands-on experience with cloud services such as AWS, Azure, or Google Cloud.
  • Knowledge of containerization tools like Docker and Kubernetes.
  • Strong scripting abilities in languages like Python, Bash, or Go.

Responsibilities

  • Design automated deployment, monitoring, and alerting systems.
  • Build scalable infrastructure using Infrastructure as Code techniques.
  • Monitor system reliability and performance with Grafana and other tools.
  • Collaborate with development and operations teams for system efficiency.
  • Quickly diagnose and resolve production issues to reduce downtime.
  • Establish best practices for Site Reliability Engineering processes.
  • Enhance observability by refining logging and monitoring systems.

Benefits

  • Hybrid working program allowing for flexible remote work opportunities.
  • Comprehensive employee benefits package.
  • Discretionary bonus eligibility based on performance.
Full Job Description
In this role you will play a crucial role in maintaining the reliability, scalability, and overall performance of our production systems. This position collaborates closely with development, operations, and product teams to automate workflows, monitor system health, and maintain robust services. Expertise in Grafana is vital for creating insightful visualizations and analyzing performance metrics.

Key Responsibilities:
  • Design, implement, and manage automated deployment, monitoring, and alerting solutions.
  • Build and support scalable infrastructure through Infrastructure as Code (IaC) tools.
  • Use Grafana and other monitoring platforms to track system reliability and performance.
  • Partner with development and operations for ongoing improvements to system reliability and efficiency.
  • Diagnose and resolve production issues quickly to minimize downtime.
  • Create and maintain best practices and guidelines for SRE processes.
  • Enhance observability by improving logging, monitoring, and alert systems.
  • Participate in on-call rotations to ensure round-the-clock support for critical systems.
  • Lead post-incident reviews and put preventative measures in place.
  • Mentor and educate team members on SRE methodologies and technologies.


Qualifications:
  • Bachelor's degree (or equivalent experience) in Computer Science, Engineering, or a related area.
  • Demonstrated experience as a Site Reliability Engineer (SRE) or in a similar capacity.
  • Strong background in automation tools and methodologies such as Ansible, Terraform, or Jenkins.
  • Advanced skills in monitoring and visualization with Grafana.
  • Experience working with cloud providers like AWS, Azure, or Google Cloud.
  • In-depth knowledge of containerization and orchestration tools (e.g., Docker, Kubernetes).
  • Familiarity with CI/CD pipelines and associated tools.
  • Proficient scripting or programming abilities in languages like Python, Bash, or Go.
  • Exceptional problem-solving and troubleshooting capabilities.
  • Excellent communication and teamwork skills.
  • Comfortable working in a fast-paced, ever-changing environment.


Preferred Qualifications:
  • Hands-on experience with Prometheus or comparable time-series databases.
  • Solid understanding of networking and security best practices.
  • Knowledgeable in database administration and optimization strategies.


The expected base salary ranges from $111k-$160k. Salary offers are based on a wide range of factors including relevant skills, training, experience, education, and, where applicable, certifications and licenses obtained. Market and organizational factors are also considered. In addition to salary and a generous employee benefits package, successful candidates are eligible to receive a discretionary bonus.

Other requirements

Mizuho has in place a hybrid working program, with varying opportunities for remote work depending on the nature of the role, needs of your department, as well as local laws and regulatory obligations. Roles in some of our departments have greater in-office requirements that will be communicated to you as part of the recruitment process

About Mizuho Financial

Mizuho Financial Group, Inc. is a Japanese banking holding company headquartered in the ?temachi district of Chiyoda, Tokyo, Japan. The name "mizuho" literally means "abundant rice" in Japanese. It holds assets in excess of $1.8 trillion US dollars through its control of Mizuho Bank, Mizuho Corporate Bank, and other operating subsidiaries. The company's combined holdings form the second largest financial services group in Japan. Its banking businesses rank third in Japan after Mitsubishi UFJ Financial Group and Sumitomo Mitsui Financial Group. It is the 15th largest banking institution in the world by total assets as of December 2018.
Learn more about Mizuho Financial
Size
54,492 employees
Market Cap
$35 billion
Industry
Net Income
$84.4 billion
5 Year Trend
-0.6%
NASDAQ

Similar Jobs

More Jobs at Mizuho Financial

More Information Technology Jobs

Find similar Site Reliability Engineer jobs: