Infrastructure Engineer

MatX

$175K — $400K *
Enterprise Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 5-7 years of experience in Linux systems and networking
  • Strong understanding of troubleshooting OS, network, and cloud issues
  • Familiarity with infrastructure as code, preferably with Terraform
  • Ability to write scripts and tools to enhance productivity
  • Demonstrated capability in operational thinking and risk assessment
  • Experience in handling shared infrastructure and security considerations
  • Proficiency with git workflow and collaboration practices

Responsibilities

  • Diagnose and resolve issues related to the OS and network
  • Develop tools and scripts to improve engineering workflows
  • Collaborate on and support existing infrastructure setups
  • Manage changes to production environments with caution
  • Help onboard new engineers and diagnose system-related issues
  • Engage with complex troubleshooting across multiple system layers

Benefits

  • 4 weeks of accrued PTO plus 12 company holidays
  • Company-subsidized health insurance for employees and dependents
  • Up to 5% company contribution for retirement plans
  • $1500 annual budget for professional development
  • Onsite meals provided Monday through Friday
  • Fully covered commuting costs via Uber or train
  • Monthly stipends for cell and internet expenses
  • 100% of mental health benefits covered
  • 12 weeks paid parental leave and family support programs
  • Robust AI resources and internal tooling support
Full Job Description
What You'll Do Here
  • Day to day, the work spans:
    Linux and networking work (the core of the role)
    • Diagnose and fix issues across the OS, network, and cloud stack
    • Reason about routing, DNS, firewalls, VPCs, private connectivity, and trust boundaries
    • Track down "permission denied" that's actually a mount option, or "build is slow" that's actually a metadata-server timeout
    • Improve, harden, and extend the network and host configuration we already have

    Building tools and integrations
    • Write internal tools, scripts, and small services that make the engineering team faster
    • Pick up unfamiliar protocols and codebases and ship working integrations against them

    Supporting the infrastructure stack
    • Pair with the system owner on compute, CI, shared storage, developer VMs, and the Terraform-managed cloud setup; take ownership of areas as you grow into them, and cover when they're out
    • Execute and review production changes carefully - a bad apply can take down the shared filesystem

    Helping engineers
    • Onboard new hires and debug their environment problems
    • Solve the kind of problems that start with "X is broken" and end with a fix three layers down the stack

Who You Are
  • We care more about instincts and pattern recognition than a checklist of tools. The right person has seen enough systems like ours to know which questions to ask
  • Deep Linux systems knowledge. You can debug from userspace down to syscalls and routing tables, and you've spent enough time with namespaces, mounts, and process semantics to recognize their failure modes on sight
  • Deep networking. VPCs, DNS, firewalls, shared filesystems, private connectivity. Has opinions on when to reach for peering vs a private-service endpoint vs an identity-aware proxy vs an overlay network - and can articulate which choices expand the trust boundary and which don't
  • Strong generalist instincts. You don't need a paved path to make progress. You'll learn enough of a build system to debug a remote-cache miss, ship a small service against a protocol you've never seen, or read upstream source to verify a claim - preferring the source over the docs when it matters
  • Infrastructure-as-code experience on a major cloud. Comfortable in production: reading plans, reasoning about drift, executing migrations without taking the cluster down. We use Terraform on GCP; depth there is a plus, but the principles transfer and we'll happily talk to people coming from AWS, Azure, or other IaC tools
  • Conservative about new patterns. When introducing a new module or tool, reads a few siblings first to pick up conventions. Spots and questions inherited patterns that don't apply to the new use case
  • Threat-modeling instincts for shared infrastructure. Reasons about who can talk to what, what gets cached and trusted by whom, and the blast radius when something goes wrong. Distinguishes load-bearing security choices from defense-in-depth
  • Operational thinking. Reasons about apply ordering, coordination windows, and "what fails first if X is misconfigured"
  • Surgical git workflow. Knows the rebase tooling well enough that rewriting a branch isn't scary. Splits unrelated work into separate PRs. Never resorts to --no-verify or destructive shortcuts to make a problem go away
  • This is a hybrid role that will require you to work from our Mountain View, CA office 3 days a week on Tuesday through Thursday

Bonus Points If You Have
  • GCP depth specifically: IAM, managed compute, identity-aware proxies
  • Bazel and remote build/cache internals; buildbarn or equivalent
  • Operating batch compute or job schedulers - HPC, Slurm, Nomad, Kubernetes batch, or similar
  • Working understanding of token-based auth and cloud identity flows
  • Rust or Python scripting for tooling (not product code)
  • EDA/semiconductor tool chain familiarity (Synopsys, Cadence)
  • Managing fleets at the OS level: policies, images, package distribution
  • You don't need to write RTL or understand hardware architect but this is a plus
  • You don't need to be a product-software engineer - but you should be able to read a build rule, a Rust error message, or a CI workflow and figure out what went wrong, and write small tools or services when the team needs one. This is another plus

Compensation

The US base salary for this full-time position is determined based on a variety of factors including role, experience, location, job related skills, and relevant education and training. Career length is only a guideline for compensation.
  • Early Career - $160,000 - $275,000 + equity
  • Mid Career - $175,000 - $400,000 + equity
  • Senior Career - $250,000 - $600,000 + equity


What We Offer
  • Time off: 4 weeks PTO (accrued) + 12 company Holidays + up to 3 weeks remote work
  • Health: Company-subsidized Medical (Kaiser or Anthem) for employees & dependents, Guardian Dental and Vision insurances for employee & dependents, and life insurance (employee only), plus HSA and FSA offerings via Lively.
  • Financial Wellbeing: Choose from Roth IRA or 401K (or both) retirement plans with up to 5% company contribution to 401K (even if you don't contribute). Also, 100% company-paid life insurance (up to $300K) and long-term disability insurances.
  • Professional Development: $1500 Professional Development Budget (per year)
  • Team Meals: MatX provides onsite team lunch & dinner Monday - Friday, with your choice of ordering via WeBox, Specialty's or via our reimbursement system
  • Commute on Us: Commute on our company Uber account, or reimburse your train rides. Either way, we pay 100% for your daily commute.
  • MatX E[x]tras: $50/mo to use on the perk you value most
  • Cell & Internet Reimbursement: $35/mo for cellular and $40/mo for wifi
  • Mental Wellbeing: 100% paid mental health benefit via SpringHealth and Guardian EAP.
  • Support to Parents: Up to 12 weeks paid parental leave regardless of path to parenthood, 10 weeks pregnancy disability leave, flexible return-to-work hours, and Benepass reproductive health & parental benefit.
  • AI Resources: Up to $20K/month plus a dedicated internal AI Tooling Team to support your productivity

Similar Jobs

More Jobs at MatX

More Enterprise Technology Jobs

Find similar Infrastructure Engineer jobs: