Member of Technical Staff - Research

at Polymath — Simulation environments to train & evaluate long-horizon AI agents

San Francisco, CA, USFull-timeAny (new grads ok)$150K - $300KYC-W26

What the role involves

About Polymath

Polymath is an applied research lab focused on advancing long-horizon agent capabilities through reinforcement learning. We design and scale simulation environments where agents learn to operate safely and autonomously. We work with the world’s leading model labs to push the frontier of agent capabilities. Polymath is backed by Base10, Founders Future, Y Combinator, and other incredible investors & angels. We've raised an $8M seed, and are growing out our founding team.

About the role

We’re hiring a Member of Technical Staff - Research to help advance the frontier of autonomous agents. You’ll work on core research problems in long-horizon evaluation, agent post-training, and environment design, with a focus on understanding where current models fail and how to improve them. As a member of the founding team, you should expect to wear multiple hats: building benchmarks, shaping environments, writing production code, and running rigorous experiments. We’re looking for people who are excited by hard open-ended problems and want to operate at the intersection of research and engineering. Examples of projects you could work on include:

Developing an advanced environment simulation engine for training & evaluating autonomous AI agents

Investigating failure modes of frontier models

Creating rigorous benchmarks that evaluate how well frontier agents perform on complex, realistic tasks requiring long-horizon reasoning and tool use in dynamic environments

Post-training agents in complex simulation environments

Publishing research

You’ll be a good fit if you:

Have strong engineering & research fundamentals and are a prolific user of AI tools

Have experience post-training frontier models

Have experience shipping reliable, production-quality code

Have a track record of publications

Perks

  • 🪷 Comprehensive health, dental, and vision insurance
  • 🏦 401(k)
  • 🌎 Unlimited PTO
  • 🍽 Free meals with the team
  • 🧘 Wellness stipend & learning stipend
  • 💻 Top of the line tech
  • 🌁 Frequent team activities and outings

Culture

  • Polymath is a team of researchers, engineers, and operators focused on advancing the frontier of safe, superintelligent AI agents.
  • We have a flat organizational structure. We believe that people do their best work when they’re self-motivated and driven by a desire to learn, contribute to the team’s goals, and advance scientific progress.
  • We’re looking for folks who ship fast, set high standards for themselves, and are great team players. You’ll be a member of our founding team: have fun, learn a lot, and do high-impact work alongside great people.

About Polymath

We’re heading towards a future where AI agents will be able to perform useful work over long horizons, with little or no human supervision. To increase the reliability, performance, and safety of autonomous agents, they must be trained in simulation environments that reflect the real world. Polymath builds simulated worlds for agents to practice and learn through experience. We're a team of researchers and engineers from UC Berkeley, Hume AI, Plaid, and Amazon. We have years of experience post-training frontier models in industry, and building large scale data systems. Polymath is backed by Y Combinator.

Full Polymath profile

Other roles at Polymath

Similar Engineering, Machine learning elsewhere

jobo is a browser extension. Open this on a computer to install it.