Find Jobs
Find Jobs Near You – Available Work in Your Location
Security Penetration Tester, Frontier AI Evaluation (Contract)
Choose a Location
This role is available in multiple locations. Pick one to apply.
Career Insights for Vulnerability Analyst / Penetration Tester
See where this job fits in the broader career landscape. Knowing your career path helps you see what's possible from here.
Scorecard
Based on California data
Review key factors to help you decide if this role fits your goals. How is this calculated?
What they do
A Vulnerability Analyst or Penetration Tester probes for and exploits security vulnerabilities in web-based applications, networks and systems. Penetration Tests are designed to achieve a specific, attacker-simulated goal and should be requested by customers who are already at their desired security posture. A typical goal could be to access the contents of the prized customer database on the internal network, or to modify a record in an HR system. Vulnerability Assessments are designed to yield a prioritized list of vulnerabilities and are generally for clients who already understand they are not where they want to be in terms of security. The customer already knows they have issues and simply need help identifying and prioritizing them.
$125,034 / year median in California
+2% projected growth
Job Description
Security Penetration Tester, Frontier AI Evaluation (Contract) at Cobalt Security Penetration Tester, Frontier AI Evaluation (Contract) at Cobalt in Piedmont, California Posted in 1 day ago.
Type:
contract About the role: Cobalt is seeking experienced penetration testers to contribute expert reasoning, technical problems, and evaluation data used to train and assess frontier AI models on security tasks. This opportunity is suited to people who test systems for a living or have done so: penetration testers, red team operators, and serious bug bounty hunters, whether from consultancies, internal security teams, or independent practice. You do not need prior experience in data annotation or AI research. What matters is that you can find and reason about real weaknesses in software and infrastructure unaided, and that you can document how you got there clearly enough for another practitioner to follow. All work is performed against sandboxed environments, purpose-built targets, and model endpoints supplied by us or by the lab. We do not accept work performed against systems you are not authorized to test, and we do not accept material covered by a client agreement or obtained without authorization. What you'll do: Depending on the project, you may: Produce written testing traces, capturing how you form and test hypotheses, what you rule out and why, and how you arrive at a working approach, rather than only the end result Author novel security problems, capture-the-flag style challenges, and lab environments with verifiable success criteria Evaluate model-generated security content and code, ranking responses, explaining what makes the stronger one stronger, and identifying the specific step at which the technical reasoning breaks down Assess whether stated findings are supported by the underlying evidence, and identify inconsistencies between reported results and what the target actually does Design rubrics and partial-credit criteria for scoring multistep testing and remediation tasks Projects follow their own guidelines, scope rules, and quality standards, and you will work with feedback from reviewers and lab research teams.
Required qualifications:
Demonstrable penetration testing experience, evidenced by professional engagements, published vulnerability research or CVEs, a substantive bug bounty record, competitive CTF results, or comparable work Strong hands-on coding ability in at least one of Python, C, C++, Go, Rust, or JavaScript, sufficient to read unfamiliar codebases and write your own tooling rather than only running existing tools Depth in at least one area, for example web and API security, cloud and container security, network and infrastructure testing, mobile security, or binary exploitation Ability to explain each step of your reasoning clearly in writing, and to produce documentation another practitioner could reproduce Willingness to work strictly within defined scope and authorization, and to sign a confidentiality agreement covering project materials Certifications such as OSCP, OSWE, OSEP, GPEN, or GXPN are useful but not required. Why join
Cobalt AI:
Advance frontier AI where it counts. Apply your security expertise to data that frontier labs cannot obtain any other way, where your judgment directly shapes how the next generation of models reasons about security. Grow professionally. Expand your influence through evaluation projects, advisory roles, and research collaborations, while developing a working understanding of how frontier models are trained and assessed. Work with a top-tier network. Collaborate with security engineers and researchers from leading organizations on high-impact, flexible work. Set your own schedule. Flexible 10 to 40 hour weeks that fit around your existing engagements and your life. Competitive pay. Rates vary by project and are determined by a number of factors, including scope, skillset, and experience.