The application window is expected to close on:
Job posting may be removed earlier if the position is filled or if a sufficient number of applications are received.
Role SummarySplunk, a Cisco company, helps organizations build digital resilience across security and observability. The Structured Store team builds and maintains critical data-storage infrastructure and services that power core Splunk workflows and applications. We help shape the future of structured storage at Splunk through reusable architectures, service APIs, and design patterns that set standards for how these systems are scaled, optimized, operated, and extended.
Our engineers own the full service lifecycle, from design and implementation through deployment, observability, on-call response, and continuous improvement. The team's portfolio spans graph, key-value, and related structured-storage services running on premises and across AWS, Azure, and GCP.
In this role, you will remain hands-on while leading end-to-end projects and medium-sized features, coordinating delivery across product, SRE, security, support, and engineering partners, and contributing to the team's defined roadmap.
What You'll Get to Do - Lead medium-sized features and end-to-end projects through scoping, design, delegation, implementation, review, rollout, and production validation; contribute evidence and technical judgment to the defined roadmap.
- Lead the design of versioned APIs and service capabilities for graph, key-value, and related structured-storage services, covering authentication, routing, placement, policy, caching, and database lifecycle operations; identify opportunities for related services to create greater customer value.
- Remain hands-on in Go and adjacent technologies, resolve complicated errors, identify larger-scope risks through code and design reviews, and contribute to feature threat modeling.
- Coordinate estimates, specifications, milestones, dependencies, and remediation plans across product, SRE, platform, security, support, and engineering partners.
- Use product adoption, reliability, scalability, security, performance, and cost signals to guide design tradeoffs and continuously improve customer outcomes.
- Act as an escalation point for production issues, lead cross-functional monitoring and recovery improvements, and lead postmortems, root-cause analysis, and durable corrective actions.
- Mentor peers and junior engineers through reviews, documentation, and technical guidance; identify and recommend improvements to software-development and operational practices.
Minimum Qualifications - Bachelor's degree and 7+ years of related experience OR master's degree and 4+ years of related experience OR PhD and 1+ year of related experience.
- Experience leading end-to-end software projects from design through production release and validation, including coordination across engineering teams or functions.
- Experience developing production backend services and versioned APIs in Go, C++, or another systems programming language.
- Experience designing and operating production database-backed services using at least one graph, relational, or NoSQL database, including schema evolution, transaction handling, backup or recovery, or performance tuning.
- Experience designing production distributed systems involving consistency, replication, partitioning, routing, or failure recovery.
- Experience delivering and operating containerized services through CI/CD on Kubernetes in a public-cloud or on-premises environment.
- Experience serving as a production escalation point and leading root-cause analyses or post-incident reviews.
- Experience mentoring engineers through code reviews, design reviews, or technical guidance.
Nice-to-Have Qualifications If you meet the core qualifications but not every preferred item, we still encourage you to apply.
- Deep Neo4j Enterprise and Cypher experience, including graph modeling, clustering, topology, Bolt or Query API usage, database provisioning, page cache, and transaction metrics.
- PostgreSQL, Amazon Aurora, or RDS experience involving connection security, IAM authentication, schema evolution, tuning, backup, restore, and regional recovery.
- Experience with Helm, Kubernetes controllers or Operators, GitLab CI/CD, Terraform, Puppet, or similar delivery and configuration tooling.
- Experience integrating services across public clouds and on-premises environments, including service mesh, network policy, workload identity, or Vault.
- Knowledge of database internals, storage engines, query planning, indexes, transaction logs, concurrency control, or large-scale capacity and performance engineering.