About this Manager, AI Benchmarking and Evaluation Research (Remote, ROU) role at CrowdStrike
As a global leader in cybersecurity, CrowdStrike protects the people, processes and technologies that drive modern organizations. Since 2011, our mission hasn’t changed — we’re here to stop breaches, and we’ve redefined modern security with the world’s most advanced AI-native platform. We work on large scale distributed systems, processing almost 3 trillion events per day and this traffic is growing daily. Our customers span all industries, and they count on CrowdStrike to keep their businesses running, their communities safe and their lives moving forward. We're proud to work for a mission-driven company leveraging AI to transform the way we work. CrowdStrikers drive their careers through flexibility and autonomy while also being expected to contribute to a culture of responsible AI adoption, experimentation, and innovation. We use an AI-first mindset as a force multiplier to proactively and continuously accelerate execution, build expertise, uncover insights, and solve complex problems. We’re always looking to add talented CrowdStrikers to the team who have limitless passion, a relentless focus on innovation and a fanatical commitment to our customers, our community and each other. Ready to join a mission that matters? The future of cybersecurity starts with you.
About the Role:
The CrowdStrike Data Science Team is looking for an experienced and driven leader to build and guide a team dedicated to designing and building evaluations for AI models that perform cybersecurity tasks. Our mission is to establish rigorous, reproducible standards for measuring how well AI and agentic systems support real-world security operations.
As the Evaluation Lead, you will bring deep, hands-on SOC expertise to the table, defining what "good" looks like for AI models operating in security workflows. You will build the datasets and methodologies that ground our models in the realities of frontline defense, and lead a team that turns findings into actionable insights for engineering and product stakeholders.
What You'll Do:
Lead, mentor and grow a team of researchers and engineers focused on evaluating AI models applied to cybersecurity tasks
Define the strategy, roadmap and success metrics for evaluating AI/LLM and agentic systems across security use cases (e.g. incident response, threat hunting, alert triage, investigation)
Design and standardize evaluation methodologies, benchmark datasets and reproducible testing pipelines rooted in real SOC workflows
Assess models for accuracy, robustness, reliability and operational effectiveness in security-analyst scenarios
Establish quantitative and qualitative metrics that reflect how AI models and agentic systems perform against genuine threat-detection and response challenges
Collaborate cross-functionally with engineering, product and threat-research teams to translate evaluation findings into scalable improvements
Communicate results, trade-offs and recommendations clearly to both technical and executive audiences
What You'll Need:
Hands-on SOC experience with a strong understanding of day-to-day security operations
At least 3 years in a management or team-leadership position, with a proven track record of mentoring and growing technical teams (formal manager experience is a plus)
Strong knowledge of incident response and threat hunting, including detection, investigation and remediation workflows
Broad knowledge of the cybersecurity landscape, including attack vectors, defense mechanisms, and the analyst workflows that AI aims to augment
Practical experience working with AI, including familiarity with AI/LLM capabilities and knowledge of agentic systems and how they operate
Ability to define, design and standardize evaluation methodologies and reproducible testing pipelines
Exceptional communication skills, including the ability to present complex technical findings clearly to both technical and non-technical audiences
Demonstrated track record of delivering results, supported by shareable projects or measurable outcomes
A genuine passion for working at the forefront of AI and cybersecurity, with a proactive mindset toward continuous learning and innovation in a rapidly evolving field
The following is required within every job description: "Proven experience utilizing AI technologies to enhance decision-making, streamline workflows and processes, improve efficiency and drive business outcomes.
Bonus Points:
Relevant security certifications (e.g., GCIA, GCIH, GCFA, OSCP, or equivalent)
Experience building or working with LLM/agentic evaluation frameworks and benchmarking pipelines
Programming experience in Python and familiarity with data-science tooling
Familiarity with SOC platforms, SIEM/SOAR tools, and detection engineering
Exposure to adversarial testing, red-teaming, or robustness evaluation of AI systems
Understanding of MITRE ATT&CK and how it maps to detection and response workflows
What We Offer:
Competitive salary
Stock options
Private Healthcare insurance
Life insurance
Training budget
Working with the latest technologies
Flexible time off
Team hangouts
#LI-JP1
#LI-GT1
#LI-Remote
Benefits of Working at CrowdStrike:
- Market leader in compensation and equity awards
- Comprehensive physical and mental wellness programs
- Competitive vacation and holidays for recharge
- Paid parental and adoption leaves
- Professional development opportunities for all employees regardless of level or role
- Employee Networks, geographic neighborhood groups, and volunteer opportunities to build connections
- Vibrant office culture with world class amenities
- Great Place to Work Certified™ across the globe
CrowdStrike is proud to be an equal opportunity employer. We are committed to fostering a culture of belonging where everyone is valued for who they are and empowered to succeed. We support veterans and individuals with disabilities through our affirmative action program.
CrowdStrike is committed to providing equal employment opportunity for all employees and applicants for employment. The Company does not discriminate in employment opportunities or practices on the basis of race, color, creed, ethnicity, religion, sex (including pregnancy or pregnancy-related medical conditions), sexual orientation, gender identity, marital or family status, veteran status, age, national origin, ancestry, physical disability (including HIV and AIDS), mental disability, medical condition, genetic information, membership or activity in a local human rights commission, status with regard to public assistance, or any other characteristic protected by law. We base all employment decisions--including recruitment, selection, training, compensation, benefits, discipline, promotions, transfers, lay-offs, return from lay-off, terminations and social/recreational programs--on valid job requirements.
If you need assistance accessing or reviewing the information on this website or need help submitting an application for employment or requesting an accommodation, please contact us at [email protected] for further assistance.