Sr. Engineer II, EPICS, NG-SIEM (Hybrid)

CrowdStrike Holdings, Inc.$160K — $250K *
Information Technology
8 - 10 years of experience
Job Overview by Ladders

Qualifications

  • 10+ years of experience in software engineering, site reliability engineering, or platform engineering with large-scale distributed systems.
  • Strong proficiency in a systems programming language (Go, Java, Rust, C++) and a scripting language (Python, Bash).
  • Deep experience with end-to-end observability and building monitoring pipelines.
  • Experienced in diagnosing and resolving complex incidents across distributed components.
  • Hands-on experience with streaming platforms such as Kafka and knowledge of back-pressure and partition management.
  • Familiarity with infrastructure-as-code and CI/CD pipelines.
  • Strong written and verbal communication skills for incident leadership.

Responsibilities

  • Design, build, and maintain monitoring and synthetic test suites for deep visibility into the NG-SIEM pipeline.
  • Engineer coordinated scaling solutions to manage unified systems resources efficiently.
  • Serve as a subject matter expert in platform-wide incident responses, facilitating diagnosis and resolution.
  • Build forecasting models for end-to-end capacity and develop tooling for cost tracking.
  • Transform manual procedures into automated workflows for pipeline and infrastructure issues.
  • Collaborate with cell-level teams and stakeholders to manage SLO breaches and communication during incidents.
  • Identify and implement systemic improvements across the NG-SIEM platform.

Benefits

  • Market leader in compensation and equity awards.
  • Comprehensive physical and mental wellness programs.
  • Competitive vacation and holidays for recharge.
  • Paid parental and adoption leaves.
  • Professional development opportunities for all employees.
  • Employee Networks and volunteer opportunities to build connections.
  • Vibrant office culture with world-class amenities.
Full Job Description
About the Role:

Our mission is to make all of our customers' security-relevant data continuously available for automated detection and response, threat hunting, and other Falcon platform use cases. To enable this, the systems behind NG-SIEM (next-generation security information and event management) are growing to accommodate >100 PB of event and action data ingested every day, up to 10 years of retention, and dozens of millions of queries per hour across large sections of the data stored, for tens of thousands of customers. As a Senior Engineer II on the newly established NG-SIEM EPICS (End-to-End Performance, Incident-response, Cost, and Scaling) team, you will own the reliability and scalability of the security industry's largest SIEM platform — treating these as software engineering problems rather than purely operational ones.

The NG-SIEM platform comprises many decoupled components interacting across complex pipelines. As we scale, ensuring end-to-end health across ingest, search, and workflow execution requires deep cross-service expertise and coordinated action. You will be the engineer who builds the observability, automation, and scaling systems that keep the entire platform performing — not just individual components. You will join a distributed team of high-ownership technical leaders who share a strong passion for our mission: to stop breaches.

This is a hybrid opportunity, with the expectation to be in our Austin, TX office 2-3x a week.

What You'll Do:
  • End-to-end observability: Design, build, and maintain monitoring and synthetic test suites that provide deep visibility into the health of the entire NG-SIEM pipeline — from ingest through search and workflow execution — enabling rapid root cause analysis across component boundaries.

  • Coordinated scaling: Engineer orchestrated scaling solutions that treat the NG-SIEM pipeline as a unified system, proportionally increasing resources across all dependent components (Kafka, ingest pipelines, downstream services) to eliminate cascading bottleneck patterns.

  • Incident response engineering: Serve as a subject matter expert during platform-wide incidents (P2 and above), applying cross-service knowledge to diagnose and resolve multi-component failures. Partake in follow-the-sun on-call rotations, providing incident commander coordination for critical platform-wide events.

  • Capacity planning and cost management: Build and refine models for end-to-end capacity forecasting that account for all pipeline dimensions, including partner team dependencies (data services, GPS). Develop tooling to continuously track and surface cost drivers across the platform.

  • Automation and runbooks: Transform manual standard operating procedures into automated remediation workflows — including pipeline-wide scaling responses, CID rebalancing, and infrastructure healing — with the goal of resolving issues before customers are impacted.

  • Cross-team collaboration: Partner with cell-level teams, product engineering, GDI/3PI, and external stakeholders (e.g., CSM) to triage SLO breaches, drive problem management for large reliability efforts, and ensure consistent communication during incidents.

  • Platform improvements: Use your broad NG-SIEM knowledge to identify and drive systemic improvements across teams, contributing to the platform's long-term resilience and efficiency.


What You'll Need:
  • A passion for reliability engineering and curiosity about how large-scale running systems behave under pressure;

  • 10+ years of experience in software engineering, site reliability engineering, or platform engineering, with significant time spent on large-scale distributed systems, and the ability to make pragmatic tradeoffs between short-term delivery needs and long-term platform goals;

  • Strong proficiency in at least one systems programming language (Go, Java, Rust, or C++) and one scripting language (Python, Bash);

  • Deep experience with end-to-end observability — building monitoring pipelines, defining SLIs/SLOs, and creating dashboards that drive actionable insights across multi-service architectures;

  • Demonstrated ability to diagnose and resolve complex incidents spanning multiple distributed components operating 24/7;

  • Experience with coordinated capacity planning and scaling for systems with significant infrastructure footprints;

  • Hands-on experience with streaming platforms (Kafka or similar) and understanding of back pressure, partition management, and consumer group dynamics at scale;

  • Familiarity with infrastructure-as-code, CI/CD pipelines, and automated deployment practices;

  • A can-do attitude — you thrive collaborating in a team and are not afraid of taking on responsibilities;

  • Strong written and verbal communication skills — you will lead incident communications and produce post-incident analyses that drive lasting improvements;

  • Comfort working across time zones with globally distributed teams.


Bonus Points:
  • Experience in a similar reliability or platform engineering role at a hyperscaler (AWS, Azure, GCP) or large-scale SaaS provider;

  • Track record of building automated remediation and self-healing infrastructure;

  • Experience with cost modeling and unit economics for large compute and storage footprints;

  • Familiarity with cloud-native architectures and serverless computing paradigms;

  • Hands-on experience operating platforms processing over 1 trillion events per day or more than 10 PB of data per day;

  • Exposure to or experience with Log Management, cybersecurity products, or security operations workflows;

  • Experience with disaster recovery planning and execution for multi-region systems.

#LI-SS1

#HTF

Benefits of Working at CrowdStrike:

  • Market leader in compensation and equity awards

  • Comprehensive physical and mental wellness programs

  • Competitive vacation and holidays for recharge

  • Paid parental and adoption leaves

  • Professional development opportunities for all employees regardless of level or role

  • Employee Networks, geographic neighborhood groups, and volunteer opportunities to build connections

  • Vibrant office culture with world class amenities

  • Great Place to Work Certified™ across the globe

CrowdStrike, Inc. is committed to fair and equitable compensation practices. Placement within the pay range is dependent on a variety of factors including, but not limited to, relevant work experience, skills, certifications, job level, supervisory status, and location. The base salary range for this position for all U.S. candidates is $160,000 - $250,000 per year, with eligibility for bonuses, equity grants and a comprehensive benefits package that includes health insurance, 401k and paid time off.

For detailed information about the U.S. benefits package, please .

About CrowdStrike Holdings, Inc.

CrowdStrike Holdings, Inc. Careers

Joining CrowdStrike Holdings, Inc. presents an unparalleled opportunity to advance a career in the tech industry with a company at the forefront of digital security. As a leader in cybersecurity solutions, CrowdStrike Holdings, Inc. offers a range of job opportunities that cater to a variety of skills and experiences, from entry-level positions to senior leadership roles.

Explore Job Opportunities

CrowdStrike Holdings, Inc. is continuously seeking talented individuals who are passionate about protecting organizations against cyber threats. With a commitment to innovation and excellence, the company is hiring professionals who are eager to contribute to a team that values hard work and creative solutions.

Innovation and Professional Growth

At CrowdStrike Holdings, Inc., employees are encouraged to push the boundaries of technology and leadership. The company supports professional growth through robust training programs, including leadership development and diversity training, ensuring that every team member has the resources to thrive in their career.

Culture and Benefits

The culture at CrowdStrike Holdings, Inc. is dynamic and inclusive, fostering a workplace where diversity is celebrated and every voice is heard. Employees enjoy comprehensive benefits that support both their professional and personal lives, enhancing job satisfaction and team morale.

Internship Programs

For those starting their career, CrowdStrike Holdings, Inc. offers internship programs that provide a rich learning environment. Interns gain hands-on experience, working alongside seasoned professionals and participating in projects that deliver real-world solutions.

Networking and Career Advancement

CrowdStrike Holdings, Inc. emphasizes the importance of networking within the industry, offering numerous opportunities for employees to connect with thought leaders and innovators. These connections can lead to career advancement and a deeper understanding of the cybersecurity landscape.

Applying for a Position

To apply for a position at CrowdStrike Holdings, Inc., candidates should prepare a resume that highlights relevant experience and skills. The interview process is designed to assess not only professional qualifications but also a candidate's fit within the company culture and team.

Stay Connected with CrowdStrike Careers

Interested candidates can stay informed about new openings and company news by subscribing to job alert emails. This personalized service ensures that potential applicants are the first to know about new opportunities that match their career interests and skills.

Join the Team

CrowdStrike Holdings, Inc. is looking for curious, creative, and solution-driven team players. Explore the employment opportunities on the CrowdStrike Holdings, Inc. careers page to find a position that matches your skills and passions.

SEARCH CROWDSTRIKE JOBS

Keep Up to Date

Stay ahead with career tips, insider perspectives, and industry-leading insights you can put to use today—all from the professionals who work at CrowdStrike Holdings, Inc.

READ CAREERS BLOG

Job Alert Emails

Customize your subscription to receive job alerts, latest news, and insider tips tailored to your preferences. Discover the exciting and rewarding career opportunities waiting at CrowdStrike Holdings, Inc.
Learn more about CrowdStrike Holdings, Inc.

Similar Jobs

More Jobs at CrowdStrike Holdings, Inc.

More Information Technology Jobs

Find similar Sr. Engineer II, EPICS, NG-SIEM (Hybrid) jobs: