AI EvalsJobs
Hiring?
All roles / Perplexity / Member of Technical Staff (AI Researcher)

Member of Technical Staff (AI Researcher)

Post-Training & RL6w ago
Source verified: this exact posting was present on Perplexity’s careers feed on . View employer source →

What the posting asks for

Names
PyTorch, Python
Doctorate
Named as preferred, or equivalent experience
Experience
2+ years asked for

Read out of the employer's own description. Absence means the posting does not say, not that the answer is no.

The role

Perplexity is seeking top-tier AI Research Scientists and Engineers to advance our AI products and capabilities. We're building the future of AI-powered search and agent experiences through our Sonar models, Deep Research Agent, Comet Agent, and Search products. Join us in creating SOTA experiences that handle hundreds of millions of queries and continue to scale rapidly.

Team Structure

Depending on your interests and expertise, you'll work on one of three specialized teams:

1. Core Research Team (Horizontal)

Focus on generating and improving base models that power all our products. This team works on foundational model capabilities, post-training techniques, building RL infra and infrastructure that benefits the entire organization.

2. Agent Products Team (Vertical)

Concentrate on fine-tuning and optimizing models for our Deep Research Agent and Labs/Canvas products. This team bridges research and product, ensuring our agent capabilities deliver exceptional user experiences.

3. Comet Agent Team (Vertical)

Dedicated to developing and enhancing our Comet Agent product. This specialized team focuses on the unique requirements and optimizations needed for Comet's specific use cases.

Responsibilities

Research & Development

  • Post-train SOTA LLMs using the latest supervised and reinforcement learning techniques (SFT/DPO/GRPO)

  • Leverage our rich query/answer dataset to scale model performance across Sonar, Deep Research, Comet, and Search products

  • Stay current with the latest LLM research, especially in model training, optimization, and personalization techniques

  • Implement preference optimization and personalization capabilities to enhance user experience

  • Invent in-house improvements and optimizations to enhance SOTA models

  • Turn research ideas into algorithms and run experiments to launch new models

Infrastructure & Implementation

  • Own full-stack data, training, and evaluation pipelines required for model development

  • Build robust and effective training frameworks (on top of Megatron/PyTorch) for post-training LLMs

  • Implement necessary infrastructure and components to support cutting-edge model training at scale

  • Integrate models seamlessly into our product ecosystem

Collaboration

  • Work closely with engineering teams to integrate models into Perplexity's product suite

  • Collaborate across teams to ensure cohesive AI experiences throughout our platform

  • Partner with product teams to understand user needs and translate them into model improvements

Qualifications

Required

  • Proven experience with large-scale LLMs and Deep Learning systems

  • Strong programming skills in Python/PyTorch; versatility is a plus

  • Experience with post-training techniques and reinforcement learning

  • Self-starter with a willingness to take ownership of tasks

  • Passion for tackling challenging problems

  • Minimum 2-6 years of experience on relevant projects (depending on seniority level)

Nice-to-have

  • PhD in Machine Learning, AI, Systems, or related areas

  • Experience in post-training LLMs with SFT/DPO/GRPO

  • C++/CUDA programming skills

  • Experience building LLM training frameworks

  • Academic publications and research impact

  • Experience with agent systems and multi-step reasoning

  • Background in personalization and preference learning

Published by Perplexity on their own careers page and reproduced here unedited. Read it at Perplexity.

Apply at Perplexity → Applications go directly to Perplexity. This board does not sit in between, take a fee from you, or see your application.

What this listing does not tell you

Listed 45 days. Of the 40 post-training and RL roles this board has watched from listing to removal, 25% were gone from their employer's careers page by day 45, and the median came down after 110 days. That is a description of other listings that have already ended, not a prediction about this one: this board records when a listing disappears, never why, and a posting still up is not on a clock it can see.

Perplexity has 13 roles open on this board, 2 of them in post-training and RL.

Get the weekly AI Evals Jobs briefNew roles and board updates, with published pay where available. This is the general weekly brief. Or browse them all now.

More roles like this

Matched by discipline, title, listed location and work arrangement.

Same discipline: Post-Training & RL · Both have staff / lead titles · Shared listed location

Same discipline: Post-Training & RL · Both have staff / lead titles · Shared listed location

Same discipline: Post-Training & RL · Both have staff / lead titles · Shared listed location

Same discipline: Post-Training & RL · Both have staff / lead titles · Shared listed location

Same discipline: Post-Training & RL · Both have staff / lead titles · Shared listed location

Privacy · Terms