Site Reliability Engineer
As a global leader in cybersecurity, CrowdStrike protects the people, processes and technologies that drive modern organizations.
As a global leader in cybersecurity, CrowdStrike protects the people, processes and technologies that drive modern organizations. Since 2011, our mission hasn’t changed — we’re here to stop breaches, and we’ve redefined modern security with the world’s most advanced AI-native platform. We work on large scale distributed systems, processing almost 3 trillion events per day and this traffic is growing daily. Our customers span all industries, and they count on CrowdStrike to keep their businesses running, their communities safe and their lives moving forward. We're proud to work for a mission-driven company leveraging AI to transform the way we work. CrowdStrikers drive their careers through flexibility and autonomy while also being expected to contribute to a culture of responsible AI adoption, experimentation, and innovation. We use an AI-first mindset as a force multiplier to proactively and continuously accelerate execution, build expertise, uncover insights, and solve complex problems. We’re always looking to add talented CrowdStrikers to the team who have limitless passion, a relentless focus on innovation and a fanatical commitment to our customers, our community and each other. Ready to join a mission that matters? The future of cybersecurity starts with you.
About the Role:
At CrowdStrike, our engineering organization depends on shared infrastructure platforms that power critical product capabilities at global scale. These platforms require dedicated engineering ownership to operate reliably, scale safely, harden for security, and mature into self-service capabilities that teams across the organization can depend on.
As an SRE & DevOps Engineer, you will own production infrastructure spanning multiple cloud providers and regions, serving engineering teams across CrowdStrike. The work is equal parts reliability engineering and DevOps engineering - building automation, hardening security, establishing governance, and enabling consuming teams to adopt these platforms effectively.
You will work across a rich technology landscape including Kubernetes, Kafka, Cassandra, PostgreSQL, Apache Pinot, OpenSearch, and Apache Spark — operating and scaling microservices-based distributed systems that process millions of security events per second with zero tolerance for data loss or downtime.
What You'll Do:
Run production infrastructure - Deploy, upgrade, and maintain platform services across multiple clouds and regions on Kubernetes, including microservices-based distributed systems
Own delivery pipelines - Build and maintain scalable CI/CD pipelines using Jenkins, GitLab CI, and Bitbucket Pipelines with GitOps workflows via ArgoCD or Flux
Own capacity plannin g - Track usage, forecast growth, right-size clusters, and optimize infrastructure costs across multi-cloud environments
Posted July 29, 2026