Position SummaryWe are looking for a hands-on Senior Software/SRE Engineer who spends the majority of their time writing and reviewing production code. You will be a primary technical contributor on the team, designing systems, closing pull requests, debugging production incidents, and setting the technical bar through example. Roughly 80% of your time will be spent on individual technical contribution and architecture; the remaining 20% on the functional management responsibilities that keep a small, high-performing team running smoothly.
You will work across our Go backend services, Python APIs, TypeScript frontend, and Kubernetes infrastructure. Owning full-stack delivery in a HIPAA-regulated environment where correctness and reliability are non-negotiable.
Responsibilities- Write production-quality Go and Python backend services: APIs, data pipelines, business logic, and integrations.
- Build and maintain TypeScript/React frontend features with real attention to UI quality and client-side performance.
- Conduct deep, substantive code reviews that raise the technical quality of the entire codebase.
- Drive architectural decision records (ADRs) and hands-on prototypes to validate technical direction before committing the team.
- Champion HIPAA compliance and security best practices at the code level - in data handling, auth, encryption, and audit logging.
- Ensure the availability and reliability of applications and systems; proactively identify and address potential failure points before they affect patient-facing workflows.
- Implement and manage comprehensive monitoring, logging, and alerting using Prometheus, Kibana, the Elastic Stack, OpenTelemetry, or equivalent tools.
- Partner with other developers to design, implement, and maintain CI/CD pipelines (GitHub Actions, GitLab, or equivalent) that streamline development workflows and deployment safety.
- Drive SDLC improvements: branch strategies, automated testing gates, progressive delivery, and rollback procedures.
- Embed security and compliance checks into pipelines so HIPAA controls are verified automatically, not manually.
- Maintain cloud infrastructure on AWS and GCP - ensuring scalability, high availability, and cost efficiency.
- Manage infrastructure as code using Terraform/Terragrunt; maintain clean, version-controlled, peer-reviewed infrastructure changes.
- Manage Kubernetes clusters across AWS (EKS) and GCP (GKE); tune resource limits, manage rollouts, and operate services at scale.
- Implement and maintain deployment tooling - Helm, Kustomize, ArgoCD, Flux - to optimize workload management and enable GitOps workflows.
- Debug misbehaving workloads, manage cluster upgrades, and drive operational maturity across the platform.
Expected frequency and duration of travel: 2-3 times a year
Qualifications and ExperienceRequired:
- 7+ years of software engineering experience with a consistent record of hands-on technical contribution.
- Expert-level Go and/or Python backend development; you write idiomatic, production-grade code daily.
- Strong coding skills in Python and similar languages for automation, tooling, and operational scripting.
- Working knowledge of Kubernetes (EKS, GKE): debugging workloads, tuning resource limits, and managing rollouts. (Helm, Kustomize, ArgoCD, or Flux a plus.)
- Production TypeScript/React frontend experience - real client-side UI work, not just Node.js services
- Hands-on AWS and GCP experience: maintaining and debugging managed services, networking, and IAM; deploying and operating containerized workloads across both platforms in production.
- Experience with monitoring and observability tools: Prometheus, Elastic Cloud (or equivalent).
- Hands-on CI/CD experience with GitHub Actions, GitLab.
- Hands-on Terraform/Terragrunt experience: maintaining and extending existing infrastructure as code.
- Strong PostgreSQL skills: schema design, complex query optimization, indexing, partitioning, and migration management.
- Experience building in a regulated or high-stakes environment where data privacy and auditability matter
- Analytical problem-solver: you debug and resolve complex, cross-system issues without losing the thread from UI to database to infrastructure.
What We Offer- Comprehensive Medical, dental, and vision coverage
- 401(k) with company match
- Unlimited time off
- Major opportunity for career development to make significant impact at an exciting growth-stage company
- Collaborative and innovative work environment
- Chance to make a real impact on patient care
- Opportunity to work on cutting-edge AI technology in healthcare
Salary range: $120,000 -$150,000 per year