AI EvalsJobs
Hiring?
All roles / Letta / Software Engineer, Agent Harness

Software Engineer, Agent Harness

Evals & Benchmarks26mo ago
Source verified: this exact posting was present on Letta’s careers feed on . View employer source →

What the posting asks for

Names
Spark, SQL, Python, TypeScript, vLLM
Doctorate
Not mentioned

Read out of the employer's own description. Absence means the posting does not say, not that the answer is no.

The role

Solving Self-Improving Superintelligence

The human brain is a sponge. Today’s AI brains are brittle and rigid. At Letta, we’re building self-improving artificial intelligence: creating agents that continually learn from experience and adapt over time.

Founded by the creators of MemGPT from UC Berkeley’s Sky Computing Lab (the birthplace of Spark and Ray). Backed by Jeff Dean, Clem Delangue, and pioneers across AI infrastructure. Our agents already power production systems at companies like 11x and Bilt Rewards, learning and improving every day.

We’re assembling a world-class team of researchers and engineers to solve AI’s hardest problem: making machines that can reason, remember, and learn the way humans do.

Note that this role is in-person (no hybrid), 5 days a week in downtown San Francisco.

We are building the developer stack for agents that can run in production applications, not just demo notebooks. As a Software Engineer, you will bridge the gap bleeding-edge agent systems and production-grade software for running agents in real applications. You will help define developer APIs for agents, and lead development of our company’s OSS stack as well as the hosted service.

Responsibilities:

  • Develop Letta's OSS agents framework and cloud service

  • Design scaleable & resilient backend services and APIs to support Letta Cloud

  • Work with researchers and support bleeding-edge agent architectures in Letta

Required skills:

  • Strong proficiency with Python

  • Strong understanding of how to architect services for security, reliability, and performance

  • Ability to build scaleable backend services

  • Strong understanding of SQL databases (Postgres)

  • Familiarity with tooling across the AI stack, such as inference engines (e.g. vLLM, Ollama), and LLM provider APIs (OpenAI, Anthropic)

  • Bonus: proficient with TypeScript, React, Tailwind, etc. (the modern stack for web applications)

Our culture

Signs it could be a great fit:

  • You want to maximize your impact: you want work on a small, incredibly talented team where every individual plays a huge role in the team's success. You wonder what it would have been like to be at OpenAI when it was just a dozen people, or Google when it was just a couple of grad students in a garage.

  • You’re excited to go head-to-head with tech giants, frontier labs, and other startups that are many times our size in both headcount and funding.

  • You are fundamentally opposed to closed frontier AI that is controlled by a handful of billionaires and private tech companies.

Signs it’s a bad fit:

  • You like stability, and get stressed out when there’s nobody telling you exactly what to do. We look for people that thrive in ambiguity and can intuit the most important problems by talking to customers and dogfooding our product.

  • You want to work a 9-5, and value clear separation of work from life. The stakes are the highest they've ever been, and the only moat is execution velocity. Operating on a strict 9-5 guarantees failure.

  • You value titles or want to people-manage. Letta is a flat company where every researcher and engineer is an individual contributor.

Our Interview Process:

  • Initial screen (30 min)

  • Technical screen (1-1.5 hours)

  • Paid in-person work trial (2 days onsite in SF)

Published by Letta on their own careers page and reproduced here unedited. Read it at Letta.

Apply at Letta → Applications go directly to Letta. This board does not sit in between, take a fee from you, or see your application.

What this listing does not tell you

Listed 766 days, which is longer than most. Of the 114 evals and benchmarks roles this board has watched from listing to removal, 100% were gone from their employer's careers page by day 766, and the median came down after 56 days. That is a description of other listings that have already ended, not a prediction about this one: this board records when a listing disappears, never why, and a posting still up is not on a clock it can see.

Pay is not confirmed for this role. Check the employer’s posting for current compensation. Explore published bands from other roles →

Letta has 2 roles open on this board.

Get the weekly AI Evals Jobs briefNew roles and board updates, with published pay where available. This is the general weekly brief. Or browse them all now.

More roles like this

Matched by discipline, title, listed location and work arrangement.

Same discipline: Evals & Benchmarks

Same discipline: Evals & Benchmarks

Same discipline: Evals & Benchmarks · Shared listed location

Same discipline: Evals & Benchmarks

Same discipline: Evals & Benchmarks

Privacy · Terms