CareerRiver

SRE/Dev Ops Engineer (Hybrid, Sunnyvale)

CrowdStrike · San Francisco Bay Area

📍 USA - Sunnyvale, CAvia workday
Apply on company site ↗
CareerRiver pulls this listing straight from the employer's hiring system — no recruiter middleman, no reposts. Applying takes you directly to CrowdStrike.
As a global leader in cybersecurity, CrowdStrike protects the people, processes and technologies that drive modern organizations. Since 2011, our mission hasn’t changed — we’re here to stop breaches, and we’ve redefined modern security with the world’s most advanced AI-native platform. We work on large scale distributed systems, processing almost 3 trillion events per day and this traffic is growing daily. Our customers span all industries, and they count on CrowdStrike to keep their businesses running, their communities safe and their lives moving forward. We're proud to work for a mission-driven company leveraging AI to transform the way we work. CrowdStrikers drive their careers through flexibility and autonomy while also being expected to contribute to a culture of responsible AI adoption, experimentation, and innovation. We use an AI-first mindset as a force multiplier to proactively and continuously accelerate execution, build expertise, uncover insights, and solve complex problems. We’re always looking to add talented CrowdStrikers to the team who have limitless passion, a relentless focus on innovation and a fanatical commitment to our customers, our community and each other. Ready to join a mission that matters? The future of cybersecurity starts with you. Sr. SRE  &  DevOps Engineer About the Role: At CrowdStrike, our engineering organization depends on shared infrastructure platforms that power critical product capabilities at global scale. These platforms require dedicated engineering ownership to operate reliably, scale safely, harden for security, and mature into self-service capabilities that teams across the organization can depend on. As an SRE & DevOps Engineer, you will own production infrastructure spanning multiple cloud providers and regions, serving engineering teams across CrowdStrike. The work is equal parts reliability engineering and DevOps engineering — building automation, hardening security, establishing governance, and enabling consuming teams to adopt these platforms effectively. You will work across a rich technology landscape including Kubernetes, Kafka, Cassandra, PostgreSQL, Apache Pinot, OpenSearch, and Apache Spark — operating and scaling microservices-based distributed systems that process millions of security events per second with zero tolerance for data loss or downtime. What You will Do Run production infrastructure -  Deploy, upgrade, and maintain platform services across multiple clouds and regions on Kubernetes, including microservices-based distributed systems Own delivery pipelines - Build and maintain scalable CI/CD pipelines using Jenkins, GitLab CI, and Bitbucket Pipelines with GitOps workflows via ArgoCD or Flux Own capacity plannin g - Track usage, forecast growth, right-size clusters, and optimize infrastructure costs across multi-cloud environments Build observability - Implement metrics, dashboards, alerts (Prometheus/Grafana), distributed tracing (Jaeger/OpenTelemetry), and actionable runbooks Own on-call and incidents - Participate in on-call rotation, lead incident resolution, write blameless postmortems, and automate repeat problems Drive reliability -  Apply SRE principles and AI-driven automation to move from reactive firefighting to proactive operations Harden security - Implement auth, encryption, secret rotation, and network policies. Own disaster recovery - Build and test backup strategies and failover mechanisms ensuring zero data loss Operate data infrastructure - Maintain reliability of Kafka, Cassandra, PostgreSQL, Apache Pinot, and OpenSearch in production Enable and collaborate - Support engineering teams with templates and patterns; partner with Infrastructure, SRE, and Data Services on shared operational problems Experience and Background 8+ years in SRE, Devops engineering, or infrastructure engineering Hands-on experience running stateful distributed systems and microservices architectures on Kubernetes in production Bachelor's degree in Computer Science or related field, or equivalent work experience Reliability Engineering Deep understanding of SRE principles - SLOs, SLIs, error budgets applied to large-scale distributed systems Strong incident management background - on-call ownership, blameless postmortems, and turning operational pain into automation Experience with chaos engineering and resilience validation for production systems Proven ability to build and maintain systems with zero tolerance for data loss or downtime Advanced observability experience including Prometheus, Grafana , distributed tracing ( Jaeger/OpenTelemetry ), and large-scale log aggregation ( ELK/Splunk ) with a focus on building custom SLO dashboards and reliability scorecards Programming & Automation Proficiency in Python and/or Golang for automation, tooling, and platform services Strong scripting and automation skills - if you do it by hand more than once, you automate it Platform and Delivery Engineering CI/CD pipeline experience - Building and owning scalable delivery pipelines using Jenkins, GitLab CI, Bitbucket Pipelines , Tekton, or equivalent Strong proficiency in Infrastructure as Code (IaC) — Terraform, Ansible, Pulumi , or equivalent Cloud and  Big Data Exposure Proficiency in at least one cloud environment (AWS, Azure, GCP) with emphasis on multi-region architecture, cloud-native reliability patterns, and security-first cloud design Strong experience with Kubernetes at scale - managing large cluster fleets, workload orchestration, and container lifecycle management Familiarity with distributed data systems including relational databases (PostgreSQL) , NoSQL ( Cassandra ), OLAP ( Pinot ), Indexing(OpenSearch) and real-time streaming platforms ( Kafka, Flink ) Exposure to Big Data and analytics technologies like Spark,Storm #LI-AP1 Benefits of Working at CrowdStrike: Market leader in compensation and equity awards C

More San Francisco Bay Area jobs

San Francisco Bay Area jobs · Browse all locations