Red Team Specialist - Cyber
OpenAIGenerative AI company
San Francisco, United StatesMid
Microsoft
Nvidia
Thrive Capital
Sequoia Capital
SoftBank
Amazon
Software Engineering
About the role
TL;DR
Assess AI model cyber capabilities and safeguards against sophisticated attackers.
- •As a Red Team Specialist focused on cyber, you will help answer two practical questions: What cyber capabilities can our models provide to real-world attackers, and do our safeguards remain effective when those attackers use increasingly sophisticated techniques? The role combines scaled evaluation with expert-driven testing.
- •Key Responsibilities Design and run rigorous evaluations of model cyber capabilities and safeguards.
- •Conduct hands-on testing to understand what models can enable when used by experienced security practitioners.
- •Build and improve automated testing infrastructure that supports repeatable measurement.
- •Test novel abuse risks in agentic systems.
- •Translate findings into clear risk assessments and actionable recommendations.
- •Requirements Substantial depth in cybersecurity (application security, penetration testing, vulnerability research, adversary simulation, or red-team operations) OR AI model evaluation (designing/running evals, building agentic harnesses, automating adversarial testing).
- •Working literacy across both cybersecurity and model evaluation.
- •Ability to write code and build practical testing tools.
- •An attacker mindset and an interest in discovering failure modes.
- •Clear written and verbal communication.
Required skills
PythonBashPenetration TestingLLMsLangChainOWASPOAuthJWTSAMLDockerKubernetesCI/CD
Nice-to-have skills
JavaScriptAWSGoogle CloudAzure
Domain expertise
cybersecurityai
Benefits & perks
Relocation assistance
Tech stack
PythonJavaScriptBashGitLinuxDockerKubernetesCI/CDAWSGoogle CloudAzurePenetration TestingOWASPOAuthJWTSAMLLLMsLangChain