Senior Software Engineer, PlatformLocation: Redwood City, CA (3 days/week in office)
Job Type: Full-time • Hybrid
The RoleWe're looking for a
Senior Software Engineer, Platform to build the shared software foundations that let GridCARE turn grid intelligence into products: secure APIs, tenant-aware authorization, integrations with job orchestration systems, and reusable application services.
Working with our tech lead, product engineers, data owner, and power systems team, you'll take these capabilities
from architecture through production adoption-making them easy for developers and AI agents to discover, integrate, and operate. You'll independently resolve implementation choices, build on managed services, and own rollout and production behavior, with the tech lead guiding platform architecture. You'll raise the engineering bar through design reviews, code review, and mentoring.
Responsibilities- Build the API platform. Use managed gateway and identity services to support browser applications and machine clients. Own the integration code, trusted identity and tenant context, request validation, and clear, versioned API contracts.
- Partner on job orchestration. Work with data and power systems engineers to integrate job workflows with shared platform services. Build the API and access-control interfaces that connect products to these workflows, preserving tenant and study context through job submission, status, and result access.
- Make authorization and tenant isolation dependable. Implement shared access controls across APIs, services, jobs, and data interfaces. Enforce resource ownership and study boundaries, including for internal users authorized to work with multiple customers, and define how permission changes affect ongoing work.
- Make the platform easy to build on. Create reusable APIs, libraries, application templates, machine-readable contracts, and executable examples. Help developers and AI agents discover capabilities, compose workflows, and recover from errors.
- Make systems dependable in production. Instrument services, investigate failures and performance bottlenecks, and partner with SRE on deployment, observability, and recovery.
- Drive delivery and adoption. Turn ambiguous needs into incremental releases, make pragmatic build-versus-buy decisions, and help existing products adopt shared authentication, authorization, and execution patterns. Evolve interfaces safely and work with data and power systems engineers on integration boundaries.
- Build reliable usage metering. Capture durable API and job events, attribute usage to the right tenant and principal, and handle retries and deduplication so usage records remain accurate.
QualificationsRequired
- 5+ years of relevant software engineering experience, with a track record of building shared backend capabilities for a multi-tenant product and owning their rollout and operation in production.
- Strong Python engineering skills and experience building production APIs and services with clear interfaces, tests, and maintainable data models.
- Hands-on experience building shared APIs, workflow services, or identity and access capabilities on managed services, with ownership of integration code, policies, and production behavior.
- Strong distributed systems fundamentals: you can reason about concurrency, partial failures, retries, idempotency, consistency, and backpressure.
- Practical experience implementing multi-tenant authorization: resource ownership, fine-grained permissions, trusted context across service boundaries, and access controls that hold through asynchronous execution.
- Experience with relational databases, queues, cloud services, and diagnosing production behavior through logs, metrics, and traces.
- Sound technical judgment and the ability to independently turn an unclear requirement into a reliable system adopted by other engineers. You can work with a tech lead on architecture and coordinate delivery across product, data, and domain teams.
- Clear communication, thoughtful code review, and an interest in mentoring teammates. You can explain tradeoffs and work constructively across teams.
Required
- Experience with FastAPI, Pydantic, OpenAPI, PostgreSQL, or the wider Python service ecosystem.
- Experience with durable workflow or job orchestration systems such as Temporal, Prefect, or Airflow.
- Experience with OAuth2/OIDC, Auth0 or similar identity providers, service identities, and relationship-based authorization systems such as OpenFGA.
- Familiarity with AWS, Kubernetes, object storage, and OpenTelemetry; experience with durable event processing or usage metering.
- Experience supporting computationally intensive, geospatial, time-series, or AI workloads.
- Effective use of AI coding tools, with disciplined review and testing of generated code.
What We Offer- Competitive salary, performance bonus, and equity.
- Comprehensive health, dental, and vision coverage.
- Lunch provided three days a week in office.
- Hybrid schedule: 3 days in office for collaboration, 2 days remote for focused work.
- Access to leading academic, industry, and government partners in the AI-energy ecosystem.
- A mission-driven team focused on shaping the future of the energy transition.
Location- This is a hybrid, in-office role based in Redwood City, CA. Employees work onsite three days/week, Tu-Thurs.
Salary Range$170,000 - $195,000 Total
Join us in tackling one of the most important infrastructure challenges of our time - enabling the energy foundation for the age of AI.