San Francisco
Research Scientist
The role
Our mission is to automate coding. The first step in our journey is to build the best tool for professional programmers, using a combination of inventive research, design, and engineering. Our organization is very flat, and our team is small and talent dense. We particularly like people who are truth-seeking, passionate, and creative. We enjoy spirited debate, crazy ideas, and shipping code.
Research Scientist
SpaceXAI is building the future of coding. We train frontier coding agents and scale RL on real user data to make them increasingly effective.
About the role
We’re looking for Research Scientists who can drive effective RL or mid-training research in a small-team setting. You’ll own ambiguous, hard research problems end-to-end: forming hypotheses, designing experiments, building the training/eval/data needed to test them, and pushing results into the next model. You should expect significantly more scope and autonomy than in other research labs.
What you’ll do
Improve our understanding of RL, what it takes to handle longer horizon tasks, and train with less compute
Train graders to improve performance on coding tasks with non-verifiable reward
Improve the quality and difficulty of datapoints we use for training our models
Realtime RL for coding agents
You may be a fit if
You have a deep background in RL and strong machine learning fundamentals
You’re an excellent programmer and software engineer
You can handle ambiguous research tasks with little guidance
You care a lot about data quality, and can dive into the data when appropriate
You are truth seeking, aiming to learn more about the science than proving your ideas are correct.
#LI-DNI
Published by Cursor on their own careers page and reproduced here unedited. Read it at Cursor.
Apply at Cursor → Applications go directly to Cursor. This board does not sit in between, take a fee from you, or see your application.
What this listing does not tell you
Listed 255 days, which is longer than most. Of the 114 evals and benchmarks roles this board has watched from listing to removal, 92% were gone from their employer's careers page by day 255, and the median came down after 56 days. That is a description of other listings that have already ended, not a prediction about this one: this board records when a listing disappears, never why, and a posting still up is not on a clock it can see.
Pay is not confirmed for this role. Check the employer’s posting for current compensation. Explore published bands from other roles →
Cursor has 6 roles open on this board, 4 of them in evals and benchmarks.
Free. One email on Thursdays.
More roles like this
Matched by discipline, title, listed location and work arrangement.
Same discipline: Evals & Benchmarks · Shared listed location
San Francisco$220k – $405k+ equity
Same discipline: Evals & Benchmarks · Shared listed location
San Francisco$244k – $305k
Same discipline: Evals & Benchmarks · Shared listed location
San Francisco$295k – $445k+ equity
Same discipline: Evals & Benchmarks · Shared listed location
San Francisco · RemoteRemote listed$150k – $300k+ equity
Same discipline: Evals & Benchmarks · Shared listed location