Job description
Lead adversarial testing of frontier and open-weight models. Build automated red-team suites, track vulnerabilities, and coordinate disclosure.
Mistral AI is looking for a Lead-level professional who can own outcomes end-to-end. You will partner with product, research, and platform teams to ship reliable AI-powered features, instrumented with the evaluation harnesses this role is responsible for.
Day-to-day you will design systems around Red teaming, Prompt injection, Python, write production code, review pull requests, run evaluations, and contribute to a blameless postmortem culture. We expect strong written communication and comfort working in a fast-moving, evidence-driven environment.
This role is remote-friendly and reports into the Open-Weight LLMs practice. Compensation ranges from $160K–$230K plus equity and benefits. We are an equal-opportunity employer and actively seek candidates from non-traditional backgrounds who have built real AI systems.
Required competencies
- Safety · Capability C002
- Safety · Capability C010
- Safety · Capability C001
AI-Matched Skills
How your competency profile maps to this role, computed by Nextcraft's match engine.