Full Job Description
We are seeking a Senior Platform Engineer (SaaS Operations) to improve the reliability, security, operability, and continuous improvement of our SaaS platform. This hands-on senior engineering role designs and automates the infrastructure, delivery systems, and operational practices that enable reliable service delivery, safe software releases, and efficient engineering workflows.
The role combines AWS infrastructure ownership, production operations, incident response, observability, security controls, server lifecycle management, and platform enablement to improve stability, scalability, and supportability. You will partner closely with Engineering, QA, Technical Support, and Product to reduce operational risk, strengthen deployment and recovery practices, and improve service reliability across the platform. When platform capabilities intersect with customer-facing workflows, you will also troubleshoot and make targeted contributions to backend services and integrations.
The ideal candidate takes ownership of production platform health, leads the resolution of complex cross-layer issues, reduces recurring operational work through automation, and drives measurable improvements in reliability, performance, recovery readiness, and engineering efficiency.
Responsibilities:
• Lead complex platform and integration initiatives end-to-end, including solution design, implementation, testing, release, and production support.
• Own SaaS platform operations across application runtime, AWS infrastructure, and service-related networking.
• Administer and maintain production server environments, including patching, hardening, configuration management, backup and recovery validation, lifecycle planning, and health monitoring.
• Lead incident response, root-cause analysis, and follow-through for platform, infrastructure, and service-integration issues.
• Build and improve observability, monitoring, alerting, runbooks, and operational standards for consistent response quality.
• Monitor platform performance, capacity, and cloud resource utilization to proactively identify risks, optimize efficiency, and support reliable service delivery.
• Automate infrastructure provisioning, deployment, patching, and repeatable operational workflows.
• Partner with Technical Support to convert recurring customer issue patterns into durable platform and product improvements.
• Support SOC 2 and related compliance efforts by implementing technical controls and producing audit-ready evidence.
• Operate with high autonomy on complex technical issues and lead cross-functional technical execution for major incidents and platform initiatives.
• Contribute to code/design/architecture reviews and provide technical mentorship that improves engineering quality, consistency, and operational best practices.
Requirements
• Bachelor's degree in Computer Science, Engineering, Information Systems, or equivalent practical experience.
• 5+ years of progressive experience in platform engineering, cloud infrastructure, site reliability engineering, DevOps, software engineering, or related disciplines, including responsibility for the reliability and operation of production systems.
• Strong hands-on AWS experience in production SaaS environments, including infrastructure provisioning, operational security controls, monitoring/alerting, and reliability-focused incident response.
• Strong networking fundamentals (TCP/IP, DNS, VPNs, routing, and connectivity troubleshooting).
• Strong production SQL skills for operational troubleshooting, data integrity checks, and service support across customer-facing workflows.
• Experience with CI/CD, Infrastructure as Code, deployment automation, and operational scripting.
• Proven ownership of production incident handling and cross-functional technical coordination.
• Experience supporting, troubleshooting, and contributing to production backend services in C#/.NET environments.
• Excellent verbal and written communication skills.
• Strong troubleshooting, analytical, and problem-solving abilities across application, infrastructure, networking, and data domains.
Preferred:
• Experience managing cloud platform cost optimization and capacity planning.
• Experience supporting SOC 2, ISO 27001, or similar security/compliance frameworks.
• Familiarity supporting Azure-hosted services in production, including identity, networking, and integration workflows.
• Experience with observability tooling (Grafana, Betterstack, Datadog, or similar).
• Experience with vulnerability management and endpoint security tooling.
• Experience with Windows-based SaaS hosting, Microsoft SQL Server operations, ESRI/GIS systems, or ASP.NET (Web Forms, MVC, or Core).
• Familiarity with JavaScript, TypeScript, and React.
Benefits
We offer a competitive salary and benefits package.
We promote a strong work/life balance at MS2. We encourage our employees to pursue their professional interests and take ownership of projects from start to finish. You'll be working with big data and cloud-based solutions using the latest technologies as part of a fun and energetic team. We get along so well, we even have regular nights out and company sponsored dinners to celebrate our successes! It's a great place to work.
Salary:
$110,000 - $156,000 a year, to be determined based on candidate's individual skills and experience
Benefits:
• Participation in the company's annual bonus program
• 401(k) with matching
• Dental insurance
• Employee assistance program
• Flexible schedule
• Flexible spending account
• Health insurance
• Life insurance
• Paid time off
• Professional development assistance
• Referral program
• Disability insurance
• Vision insurance
Hybrid work
We require you to work in the office generally at least one day per week. This is not a fully remote position, but we may provide relocation assistance to successful candidates