Detection and Response Platform Engineer

Cerebras Systems

$150K — $180K *
Information Technology
5 - 7 years of experience
Job Overview by Ladders

Qualifications

  • 5+ years of experience in cloud infrastructure, platform/SRE, software engineering, or security engineering with a strong infrastructure focus.
  • Strong proficiency in Python for writing clean, maintainable, and testable production code.
  • Hands-on experience in operating production systems within major cloud environments (AWS, Azure, GCP) including IAM and logging.
  • Experience with event/log pipelines and data workflows emphasizing reliability and data quality.
  • Proficiency in infrastructure as code and deployment automation, like Terraform or CI/CD practices.
  • Practical knowledge of detection and response concepts across various environments.
  • Strong fundamentals in operating systems and troubleshooting distributed systems.
  • Excellent written communication skills for creating documentation and runbooks.

Responsibilities

  • Own and evolve core infrastructure for detection and response workflows.
  • Design and maintain security telemetry pipelines for logs and events.
  • Manage the execution environment and tooling for automation processes.
  • Automate investigation workflows by integrating various tools.
  • Set engineering standards for tooling such as automated tests and clear documentation.
  • Ensure operational health of the platform through metrics and alerting.
  • Explore and apply emerging technologies, such as AI, to enhance security practices.
  • Maintain operational playbooks for quick diagnosis and recovery.

Benefits

  • Opportunity to build a breakthrough AI platform beyond GPU constraints.
  • Chance to publish and open source cutting-edge AI research.
  • Work on one of the world's fastest AI supercomputers.
  • Enjoy a stable job environment with the excitement of a startup.
  • Experience a simple, non-corporate work culture that values individual beliefs.
Full Job Description
We are seeking a senior Detection and Response Platform Engineer to own and evolve the systems and tooling that power our detection and response program. You will build and operate the services that move and shape security telemetry, execute detection logic reliably, and enable automation that makes investigation faster and more consistent. This is an engineering role on the Detection and Response team, focused on building durable internal systems that other engineers rely on, with strong expectations around reliability, operability, and developer experience.
Responsibilities
  • Own and evolve the core infrastructure that powers detection and response workflows at scale.
  • Design and maintain security telemetry pipelines, including collection, normalization, enrichment, and retention of logs and events.
  • Own and evolve the execution environment and supporting tooling for detections and response automation, including dependency and configuration management, release processes, and runtime health monitoring.
  • Automate investigation and response workflows by extending integrations across the tools we rely on.
  • Set and uphold engineering standards for our tooling: clear interfaces, automated tests, code review, and documentation that makes the system easy to extend.
  • Own operational health for the platform, including metrics, alerting, and on-call responsibilities focused on system reliability and availability.
  • Explore and apply emerging approaches, potentially leveraging AI, to reduce toil and strengthen our security posture.
  • Maintain operational playbooks and procedures that keep the platform healthy and make failures fast to diagnose and recover from.
Skills and Qualifications
  • 5+ years of experience in cloud infrastructure, platform/SRE, software engineering, or security engineering with a strong infrastructure focus.
  • Strong proficiency in Python, with the ability to write clean, maintainable, and testable code for production systems.
  • Hands-on experience operating production systems in a major cloud environment (AWS, Azure, or GCP), including IAM, networking, and logging/auditing.
  • Experience building and operating event/log pipelines and data workflows (streaming and/or batch), with strong fundamentals in reliability and data quality.
  • Experience with infrastructure as code and deployment automation (e.g., Terraform/Pulumi and CI/CD).
  • Practical knowledge of detection and response concepts across cloud, identity, and endpoint environments, with the ability to partner effectively with D&R engineers.
  • Strong fundamentals in operating systems, networking, and troubleshooting distributed systems.
  • Excellent written communication skills, with the ability to create clear documentation and runbooks.

Why Join Cerebras

People who are serious about software make their own hardware. At Cerebras, we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we've reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:
  1. Build a breakthrough AI platform beyond the constraints of the GPU.
  2. Publish and open source their cutting-edge AI research.
  3. Work on one of the fastest AI supercomputers in the world.
  4. Enjoy job stability with startup vitality.
  5. Our simple, non-corporate work culture that respects individual beliefs.

Find out more about what it's like to work at Cerebras here!

Apply today and become part of the forefront of groundbreaking advancements in AI!

Similar Jobs

More Jobs at Cerebras Systems

More Information Technology Jobs

Find similar Detection and Response Platform Engineer jobs: