Paris · London +3Remote listed$250k – $535k+ equity
Research Engineer - Cybersecurity (RL Environments)
What the posting asks for
- Names
- Kubernetes, Python
- Doctorate
- Not mentioned
Read out of the employer's own description. Absence means the posting does not say, not that the answer is no.
The role
About Mistral
Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems across high-stakes industries like finance, manufacturing, defense, healthcare, and the public sector, co-creating customized AI systems that they can run on their terms.
We are a dynamic, collaborative team passionate about AI and its potential to transform society. Our diverse workforce thrives in competitive environments and is committed to driving innovation. Our teams are distributed between Europe, North America, Asia and the Middle East. We are creative, low-ego and team-spirited.
The Role
Cybersecurity is one of the areas where AI could do the most good — and one where the capability must be built with particular care. As a Research Engineer on this track, you'll push our models' capabilities in vulnerability discovery and remediation, security analysis, incident response, and across the offensive-to-defensive spectrum of cybersecurity, and it will be your job to develop that ability deliberately and safely.
The work sits at the meeting point of research and engineering: you'll devise new approaches and be the one to implement them. Concretely, that means designing and building RL environments that reproduce realistic security scenarios, pipelines that generate synthetic training instances at scale, and experiments and evaluations that show what our models can actually do. Your results feed directly into the production training runs that shape our frontier models, in close collaboration with Mistral's researchers, engineers, and security specialists.
Security specialists and ML engineers are both welcome here — what matters is that you can hold your own in both worlds, and are eager to go deeper on the one you know less.
What You Will Do
Design and implement RL environments simulating cybersecurity scenarios.
Build the pipelines and infrastructure that generate synthetic cybersecurity training instances at scale — thousands of scenarios, not one hand-crafted exercise.
Conduct experiments and evaluations of model capabilities on these environments, from quick prototypes to controlled benchmark runs.
Own the infrastructure behind the environments: containers, sandboxes, VMs, cloud deployments, and orchestration (Kubernetes), all managed as code.
Partner with researchers and security specialists across Mistral, distilling their domain expertise into reproducible environments and datasets.
Write clear, efficient code in Python and enforce strong software-design practices: testing, code review, CI/CD.
What We're Looking For
Hands-on expertise in offensive and/or defensive cybersecurity: vulnerability analysis, web/network/cloud security, secure coding practices, and SOC.
Strong software engineering skills: clean, reliable, well-tested code — not just exploit scripts.
Fluency in Python plus comfort reading lower-level languages (C/C++) for vulnerability analysis.
DevOps experience: Docker, Kubernetes, cloud deployments, sandboxed or simulated environments.
A pragmatic research-to-engineering mindset: ship a working first version, then refine and scale it.
Working knowledge of RL techniques and LLM training methodologies, or strong motivation to develop it.
Self-starter, low-ego, collaborative — comfortable working across research and engineering.
Nice-to-haves
CTF, cyber-range, or bug-bounty experience — as a player, challenge author, or platform builder.
A research background in cybersecurity, academic or industrial, or in another experimental discipline.
Prior experience building RL environments or large-scale ML training infrastructure.
Relevant open-source contributions and projects.
What We Offer
We offer a comprehensive benefits package designed to support your well-being, growth, and work-life balance. Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location-specific perks.
For the most up-to-date details on benefits available in your location, please refer to our Benefits page.
Privacy Policy
Your privacy matters to us. You can learn more about how we handle your personal data in our Applicant Privacy Policy.
Published by Mistral AI on their own careers page and reproduced here unedited. Read it at Mistral AI.
Apply at Mistral AI → Applications go directly to Mistral AI. This board does not sit in between, take a fee from you, or see your application.
What this listing does not tell you
Listed 14 days, which is recent for this board. Of the 40 post-training and RL roles this board has watched from listing to removal, 5% were gone from their employer's careers page by day 14, and the median came down after 110 days. That is a description of other listings that have already ended, not a prediction about this one: this board records when a listing disappears, never why, and a posting still up is not on a clock it can see.
Pay is not confirmed for this role. Check the employer’s posting for current compensation. Explore published bands from other roles →
Mistral AI has 17 roles open on this board, 2 of them in post-training and RL.
Free. One email on Thursdays.
More roles like this
Matched by discipline, title, listed location and work arrangement.
Same discipline: Post-Training & RL · Shared listed location
London$250k – $535k+ equity
Same discipline: Post-Training & RL
United States - RemoteRemote listed
Same discipline: Post-Training & RL
San Francisco
Same discipline: Post-Training & RL
San Francisco$200k – $400k+ equity
Same discipline: Post-Training & RL