CrowdStrike Logo

CrowdStrike

Engineer II, Site Reliability (Remote, GBR)

Posted 17 Days Ago
Be an Early Applicant
Remote or Hybrid
Hiring Remotely in United Kingdom
Senior level
Remote or Hybrid
Hiring Remotely in United Kingdom
Senior level
Responsible for Linux engineering and administration at scale, ensuring availability, latency, throughput, monitoring, incident response, capacity planning and performance tuning. Develop automation and tooling, lead incident analysis, participate in on-call rotation, troubleshoot hardware, and collaborate globally to improve reliability and observability across the commercial cloud platform.
The summary above was generated by AI

As a global leader in cybersecurity, CrowdStrike protects the people, processes and technologies that drive modern organizations. Since 2011, our mission hasn’t changed — we’re here to stop breaches, and we’ve redefined modern security with the world’s most advanced AI-native platform. We work on large scale distributed systems, processing almost 3 trillion events per day and this traffic is growing daily. Our customers span all industries, and they count on CrowdStrike to keep their businesses running, their communities safe and their lives moving forward. We're proud to work for a mission-driven company leveraging AI to transform the way we work. CrowdStrikers drive their careers through flexibility and autonomy while also being expected to contribute to a culture of responsible AI adoption, experimentation, and innovation. We use an AI-first mindset as a force multiplier to proactively and continuously accelerate execution, build expertise, uncover insights, and solve complex problems. We’re always looking to add talented CrowdStrikers to the team who have limitless passion, a relentless focus on innovation and a fanatical commitment to our customers, our community and each other. Ready to join a mission that matters? The future of cybersecurity starts with you.


About the Role:

CrowdStrike is looking to hire an Engineer II to the TechOps SRE team that will have a focus on our Commercial Cloud.  We’re looking for a deeply-technical, hands-on engineer, who loves to develop automation and tooling through software to ensure delivery of mission critical solutions and services for large-scale distributed systems.  

What You'll Do:

  • Have expertise with Linux engineering and administration for thousands of bare metal servers and virtual machines

  • Be responsible for all operational aspects of our platform -  Availability, Latency, Throughput, Monitoring, Issue Response (analysis, remediation, deployment) and Capacity Planning with respect to Latency and Throughput

  • Work in a team of highly motivated engineers distributed across the globe

  • On-call rotation with other team members

  • Troubleshoot server hardware issues

  • Use your passion for technology to ensure our platform operates flawlessly 24x7

  • Obsess about learning, and champion the newest technologies & tricks with others, raising the technical IQ of the team. We don’t expect you to know all the technology we use but you will be able to get up to speed on new technology quickly

  •  Have broad exposure to our entire architecture and become one of our experts in our overall process flow

  • Have an intrinsic drive to make things better

  • Bias towards small development projects and the occasional larger projects

  • Have experience with modern monitoring and telemetry stacks (ELK, Prometheus, Grafana, Zabbix)

  • Gather and analyze metrics from both operating systems and applications to assist in performance tuning and fault finding

  • Ability to lead incident analysis for incidents, champion incident response practices and assist in correlating incidents to systemic problems, and drive towards resolution.

What You'll Need:

  • Bachelor's degree and/or equivalent experience in Computer Science

  • A minimum of five years of experience working in a large scale production environment

  • A minimum of two years of experience in software engineering

  • A minimum of two years of experience in one or more of: C++, Java, Python, Go

  • Experience with storage technologies (Examples: SAN, NAS, NFS, Object Storage, FreeNAS, iSCSI)

  • Experience with Infrastructure technologies (Examples: Linux, Windows, VMware, Docker, Kubernetes, etc.)

  • Experience writing technical documentation

  • Configuration management experience with one or more tools such as Puppet, Chef, Ansible

  • Solid understanding of application design, including operational trade-offs of various designs

  • Analytical skills coupled with a strong sense of urgency, ownership, and drive

  • Ability to work with well in a diverse, team-focused environment with other SREs and Engineers

  • Ability to broadly communicate and present recommended conventions defined by the reliability team broadly

  • Proven experience utilizing AI technologies to enhance decision-making, streamline workflows and processes, improve efficiency and drive business outcomes.

#LI-GO1

#LI-Remote


Benefits of Working at CrowdStrike:

  • Market leader in compensation and equity awards
  • Comprehensive physical and mental wellness programs
  • Competitive vacation and holidays for recharge
  • Paid parental and adoption leaves
  • Professional development opportunities for all employees regardless of level or role
  • Employee Networks, geographic neighborhood groups, and volunteer opportunities to build connections
  • Vibrant office culture with world class amenities
  • Great Place to Work Certified™ across the globe

CrowdStrike is proud to be an equal opportunity employer. We are committed to fostering a culture of belonging where everyone is valued for who they are and empowered to succeed. We support veterans and individuals with disabilities through our affirmative action program.


CrowdStrike is committed to providing equal employment opportunity for all employees and applicants for employment. The Company does not discriminate in employment opportunities or practices on the basis of race, color, creed, ethnicity, religion, sex (including pregnancy or pregnancy-related medical conditions), sexual orientation, gender identity, marital or family status, veteran status, age, national origin, ancestry, physical disability (including HIV and AIDS), mental disability, medical condition, genetic information, membership or activity in a local human rights commission, status with regard to public assistance, or any other characteristic protected by law. We base all employment decisions--including recruitment, selection, training, compensation, benefits, discipline, promotions, transfers, lay-offs, return from lay-off, terminations and social/recreational programs--on valid job requirements.


If you need assistance accessing or reviewing the information on this website or need help submitting an application for employment or requesting an accommodation, please contact us at [email protected] for further assistance.


Similar Jobs at CrowdStrike

Yesterday
Remote or Hybrid
United Kingdom
Senior level
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Maintain and enhance a Python-based domain-specific language compiler and detection tooling. Design language and compiler features, resolve cross-platform integration issues, develop end-to-end tests, improve performance and user experience, gather user feedback, and shape tooling architecture and roadmap. Collaborate with distributed engineering teams and contribute to system design decisions. Security industry experience is not required, but strong software engineering, algorithms, architecture, communication, and compiler interest are expected.
Top Skills: AgdaCC++Compiler ToolingDistributed SystemsDomain-Specific LanguagesEvent Processing SystemsMessage Passing LanguagesPythonRustSwiftTla+
8 Days Ago
Remote or Hybrid
Senior level
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
The role involves developing and improving a Windows kernel mode sensor for cybersecurity, leading projects, and collaborating across platforms to enhance endpoint security features.
Top Skills: AgileC++GitWindows Os Kernel Development
10 Days Ago
Remote or Hybrid
United Kingdom
Mid level
Mid level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Perform malware and threat research across the attack lifecycle, validate and improve Falcon detections, collaborate with engineering and ML teams, automate analysis, support external security testing, and translate findings into product protections.
Top Skills: AICrowdstrike FalconElasticsearchKibanaMachine LearningSplunk

What you need to know about the Manchester Tech Scene

Home to a £5 billion digital ecosystem, including MediaCity, which consists of major players like the BBC, ITV and Ericsson, Manchester is one of the U.K.'s top digital tech hubs, at the forefront of advancements in film, television and emerging sectors like as e-sports, while also fostering a community of professionals dedicated to pushing creative and technological boundaries.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account