Full Job Description
Mattel Game Studios' mission is to harness the power of Mattel's iconic brands and IP to create engaging, high-quality digital games. Help Mattel build and operate the infrastructure behind reliable, scalable mobile games. As Senior Platform Engineer, Cloud Infrastructure & Live Services, you will own the cloud infrastructure, build and deployment pipelines, and monitoring systems that help our teams launch confidently and keep live games running well.
The Opportunity:
You will work with internal game production and technology teams, external development studios, and Mattel support teams from early planning through live operations. This is a hands-on role for an engineer who understands the demands of mobile games and enjoys solving the challenges of operating services at global scale.
Alongside infrastructure ownership, you will write automation, troubleshoot production issues, and help developers integrate and deploy their applications. Your development skills will also contribute to our publishing platform and live operations tools.
What Your Impact Will Be:
• Build and operate cloud infrastructure. Design and run GCP environments for games and backend services. Guide decisions on APIs, databases, and caching. Manage hosting and content delivery, and plan for traffic growth, capacity, and cloud costs.
• Automate infrastructure and workflows. Use Terraform, scripts, and reusable workflows to make environment setup, configuration, and maintenance consistent and repeatable. Apply version control and code review to reduce risk and manual effort.
• Own build and deployment pipelines. Manage CI/CD pipelines, build servers, and deployment infrastructure for games and services. Work with studios on build and packaging needs, with automated checks, secure credentials, release approvals, and rollback procedures.
• Own monitoring, alerting, and automated reporting. Define actionable alerts, routing, and escalation for production infrastructure. Automate reporting on availability, performance, incidents, capacity, and cost for studios and leadership. Work with game teams to test capacity before launches and major updates, and agree who monitors and responds.
• Improve service reliability. Troubleshoot incidents with studios, internal teams, and service providers. Document response procedures and follow fixes through to completion. Participate in agreed on-call and launch coverage, and test backup and recovery procedures under realistic failure conditions.
• Keep services secure. Work with Mattel Security and Privacy teams on access, credentials, data protection, and security fixes. Validate controls without disrupting legitimate players, and maintain records of production changes.
• Build full-stack integrations and tools. Connect game clients and web applications to backend APIs, authentication, telemetry, and content services. Contribute to the publishing platform, operational dashboards, live operations tools and interfaces, and game websites, including testing, hosting, and deployment.
• Guide planning with external studios. Bring infrastructure and deployment expertise from early scoping through launch. Review technical designs, define hosting, backend, and monitoring requirements, identify risks early, and agree operational handoffs so services can be deployed and supported reliably.
• Provide technical leadership. Set infrastructure priorities, explain technical tradeoffs, and align studios and internal teams on responsibilities. Resolve delivery blockers and maintain clear documentation so others can support the services you build.
Qualifications
What We're Looking For:
• 8+ years in infrastructure, platform engineering, DevOps, or site reliability engineering, with hands-on experience supporting production services through releases, growth, and incidents.
• Strong production experience with Google Cloud Platform (GCP), including networking, compute, storage, access management, and managed services. A clear understanding of how backend services, databases, and APIs work together.
• Practical Terraform experience and strong scripting or programming skills, with disciplined use of version control and code review for infrastructure changes.
• Experience building and maintaining CI/CD pipelines with Jenkins or similar tools, including build workers, automated testing, artifact storage, secrets, and release and rollback workflows.
• Experience using logs, metrics, and traces to diagnose problems and designing reliable alerting and automated reporting. The ability to test recovery and assess capacity, performance, reliability, and cost together.
• Working knowledge of web infrastructure, including CDNs, DNS, TLS, load balancing, and caching, with the ability to troubleshoot across applications and the services they depend on.
• Strong knowledge of full-stack development and deployment practices, including front-end and back-end application structure, API design, automated testing, environment configuration, containerized deployment, and safe rollout strategies.
• Experience working with external development partners or vendors on technical planning, infrastructure requirements, and engineering standards.
• Clear communication and documentation skills, with the ability to coordinate delivery across teams and move between infrastructure, troubleshooting, and programming as priorities change.
• A degree in computer science, engineering, or a related field, or equivalent practical experience.
Additional Experience We Value
• Proficiency in React development and/or Retool, ideally for internal tools, operational dashboards, or game websites.
• AWS experience for potential future infrastructure needs, plus familiarity with containers or serverless services.
• Experience with live games, game backends such as Nakama or PlayFab, and services such as Firebase.
• Experience with Linux and macOS build environments, Unity, Unreal, and mobile build and signing workflows.
• Experience with BigQuery or similar cloud data platforms and collaboration with data engineers.
What Success Looks Like
Studios can launch and update games with confidence. Deployments are repeatable, service health is visible, and teams can detect and recover from problems quickly. You will help reduce manual work and set practical targets for reliability, release quality, recovery, and cost with production and live operations teams.
Additional Information
How We Work:
We are a purpose driven company aiming to empower generations to explore the wonder of childhood and reach their full potential. We live up to our purpose employing the following behaviors:
• We collaborate: Being a part of Mattel means being part of one team with shared values and common goals. Every person counts and working closely together always brings better results. Partnership is our process and our collective capabilities is our superpower.
• We innovate: At Mattel we always aim to find new and better ways to create innovative products and experiences. No matter where you work in the organization, you can always make a difference and have real impact. We welcome new ideas and value new initiatives that challenge conventional thinking.
• We execute: We are a performance-driven company. We strive for excellence and are focused on pursuing best-in-class outcomes. We believe in accountability and ownership and know that our people are at their best when they are empowered to create and deliver results.
Our Approach to Flexible Work:
This role is an in-office role Monday-Thursday, and work from home on Fridays. We embrace a flexible work model designed to empower a culture of growth, optimism, and wellbeing, where every employee can reach their full potential. Combining purposeful in-person collaboration with flexibility, our focus is to optimize performance and drive connection for moments that matter.