H-E-B

Sr Software Engineer (Site Reliability) Austin or Dallas, TX

H-E-B$120K — $145K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years in designing/troubleshooting distributed systems
  • 3+ years SRE experience with GKE, K8s, or AWS
  • 2+ years Java (Spring) programming preferred
  • 3+ years using Terraform for cloud infrastructure
  • 3+ years CI Pipeline experience with GitLab or GitHub
  • Experience with microservices architecture patterns
  • Proficiency with scripting languages like Python or Bash

Responsibilities

  • Develop and maintain environment monitoring and task automation tools
  • Enhance the full lifecycle of services from inception to refinement
  • Analyze and establish efficient software configurations and setups
  • Collaborate with teams on service architecture and deployment plans
  • Monitor SLOs and SLAs, addressing any identified gaps
  • Act as a technical subject matter expert for cross-functional teams
  • Troubleshoot systems-related issues and provide maintenance support

Benefits

  • Hybrid work model with preferred location in Austin
  • Opportunity to shape complex code solutions
  • Exposure to a variety of cloud technologies like GCP and Kubernetes
  • Chance to work on a highly collaborative team
  • Access to advanced monitoring and visualization tools
  • Engagement in a culture of mentoring and career growth
Full Job Description
Responsibilities

Job Summary: As a Senior Software Engineer-Site Reliability on the Digital Fullfilment team, you'll deliver complex code solutions. You'll support the build and deployment pipeline and when necessary, diagnose / solve production support or on-call issues. You'll contribute to overall system design, architecture, security, scalability, reliability, application performance and provide end-to-end support.

 

Location: Austin (preferred), open to Dallas, TX (Hybrid)

Key Responsibilities & Essential Functions:

  • Develops and maintains tooling used for environment monitoring and task automation 
  • Engages in and improves whole lifecycle of services, including inception and design, deployment, operation, and refinement 
  • Analyzes and establishes efficient configurations for software and servers, DB connections / indexes, drivers, etc. 
  • Collaborates with development teams to design service architectures, software platforms and frameworks, capacity planning, release plans and launch reviews
  • Monitors internal and vendor service level objectives (SLOs) and agreements (SLAs); identifies / resolves SLO / SLA gaps 
  • Serves as technical subject matter expert (SME) for cross-functional engineering Teams; assists with / troubleshoots systems-related issues and maintenance 

The responsibilities and essential functions outlined above describe the general nature and level of work assigned to this position. This is not an exhaustive list of all duties, responsibilities, and skills required. Duties and responsibilities may be modified at any time based on business needs. Employees may be required to perform other job-related tasks as requested by their supervisor, subject to reasonable accommodations.Qualifications & Key Requirements:Work Experience:

  • 5+ years experience designing, analyzing, developing, or troubleshooting distributed systems 
  • 3+ years of SRE experience managing Google Kubernetes Engine (preferred), K8s, or AWS environments
  • 2+ years of Java (Spring) programming experience preferred
  • 3+ years of using Terraform to maintain cloud infrastructure
  • 3+ years of CI Pipeline experience with either Gitlab Pipelines, or GitHub Actions
  • Experience with tools such as Gitlab, JIRA, Slack, Confluence and Intellij is preferred
  • Experience with microservices architecture patterns
  • Experience working with PostgreSQLKubernetesDockerLinuxGCPTerraform, and APIs using REST and GraphQL
  • Experience working with monitoring and visualization tools such as DatadogGrafana, or New Relic 
  • Strong proficiency with scripting languages such as PythonRubyGroovyBash
  • Proven track record of researching, understanding, and effectively applying Scalability and High Availability principles

Knowledge/Skills/Abilities:

  • Advanced knowledge in system and data architecture, data modeling, and design and capable of architecting and designing at the application or service level using well-accepted design patterns -
  • Able to review platform designs for strength of engineering solutions, namely performance, sustainability, and iterative development potential. -
  • Comprehensive knowledge of Computer Science fundamentals: data structures, algorithms, design patterns, system architecture and design patterns -
  • Advanced understanding of development methodologies and processes -
  • High degree of personal accountability to self and team for continued growth -
  • Adjust - Leverages Agile metrics to improve team performance and deliverables. Evaluates and adjusts resources, self, and team as necessary. -
  • Collaborate - Ability to work on tasks which span multiple domains, requiring cross-team collaboration, which have a high impact on your project. -
  • Agility - Embraces risk, change, and helps team manage ambiguity within the team's scope of work. -
  • Able to drive progress without having a complete picture and can articulate potential tradeoffs and prioritize when faced with ambiguity. -
  • Connect - Delivers clear, concise, effective messages across different levels; can tailor communication based on intended audience. -
  • Growth Mindset - Fosters a culture of mentoring and coaching across multiple technical teams and other stakeholders. -
  • Relate - Fosters a culture within their team where people are encouraged to share their opinions and contribute to discussions in a respectful manner, approach disagreement non-defensively with inquisitiveness, and use contradictory opinions as a basis for constructive, productive conversations. -

Education:

  • A Computer Science degree or comparable formal training, certification, or work experience -involving software / systems engineering

Physical Demands & Working Conditions:

  • Travel by car or plane with overnight stays
  • Work extended hours; sit for extended periods
  • Work rotating and on-call schedules, as needed

The work environment characteristics described here are representative of those a Partner encounters while performing the essential functions of this job. Reasonable accommodations may be made to enable individuals with disabilities to perform the essential functions.JDENGINEERING

DEV3232

About H-E-B

H-E-B is a privately held supermarket chain based in San Antonio, Texas, with more than 340 stores throughout the U.S. state of Texas, as well as in northeast Mexico. The company also operates Central Market, an upscale organic and fine foods retailer. As of 2021, the company has a total revenue of $32 billion. H-E-B was named Retailer of the Year in 2010 by Progressive Grocer.
Learn more about H-E-B
Size
120,000 employees
Industry

Similar Jobs

More Jobs at H-E-B

More Information Technology Jobs

Find similar Sr Software Engineer (Site Reliability) Austin or Dallas, TX jobs: