
What the role involves
About HUD
HUD is building infrastructure to create RL training data and evals for frontier AI agents, as well as a marketplace to sell these to frontier labs through the HUD marketplace. Our platform is used by frontier labs, Fortune 500 companies, and startups. We’ve raised $15M from top VCs and were YC W25.
About the role
HUD's product sells itself to the right person. Your challenge is finding and selling to them at scale.
The buyers of HUD’s data are ML engineers, AI researchers, and engineering leaders at frontier labs and enterprises. They respond not to traditional SaaS outreach, but to someone who understands what they're building, why their current eval pipeline is broken, and how HUD fixes it.
That’s why we’re looking for a GTM engineer who can both carry technical conversations and close - and do so at scale. We’re looking for systems builders, not just salespeople.
Responsibilities
Own HUD's outbound pipeline end-to-end - targeting, enrichment, outreach, follow-up, and meeting booking
Build and maintain GTM workflows (e.g., outbound email sequencing, CRM automation, lead scoring, custom tooling, inbound lead qualification)
Use AI tools aggressively to 10x your output (e.g., vibe code scrapers, automate workflows, build internal tools)
Develop and iterate on messaging that resonates with technical buyers, including ML engineers, AI researchers, and engineering leaders
Quickly learn new prospect businesses / industries and understand their pain points, workflows, and why HUD matters to them
Track pipeline metrics and optimize conversion at every stage
Work closely with founders to develop and refine our GTM playbook as we scale
Experience
You may be a good fit if you have
- Demonstrated ability to build and run a GTM motion with real, measurable results
- A strong track record of closing deals
- A CS or engineering background, or you vibe-code extensively on the side
- Deep familiarity with GTM tooling
- Strong written communication that technical people actually respond to
- Startup experience in early-stage technology companies with ability to work independently in fast-paced environments
Strong candidates may also have
- Experience selling developer tools, infrastructure, or ML/AI products
- History of building outbound systems using tools like Clay, Apollo, or custom automation
- Familiarity with the AI/ML ecosystem, e.g., you know what evals are, what RL is, and why researchers care
- Evidence of creative, high-leverage approaches to pipeline generation
- We prioritize technical aptitude and learning potential over years of experience. Motivated candidates are encouraged to apply even if they don't meet all criteria.
- Team & company details
Team Size
- ~15 people currently, mostly full-time in-person, but some remote.
- Our team: Our team includes 4 International Olympiad medalists (IOI, ILO, IPhO), serial AI startup founders, and researchers with publications at ICLR, NeurIPS, etc.
- Company stage: We have 8 figures in funding and high revenue growth. We’re scaling profitably and quickly to meet very strong demand.
Logistics
Employment : Full-time.
Location
- On-site in the San Francisco Bay Area.
- Visa Sponsorship : We provide support for relocation and visas for strong full-time candidates to the US.
- Timeline : Applications are rolling. The process is 2 technical interviews and a 1-week work trial.
What we offer
- Competitive compensation
- 100% covered top-of-the-line medical, dental, and vision from Blue Shield of CA
- Lunch and dinner when you’re in the office
- Company-wide holiday break (Christmas Eve to New Year’s Day) on top of PTO and paid holidays
- Other perks including an Equinox membership, 401k, and commuter benefits
- Unlimited* access to tokens for ChatGPT, Claude Code, Cursor, etc. *By unlimited, we mean no one on our token usage leaderboard has ever hit a limit. So we have no idea what the limit is.
Due to high volume, we may not actively respond to every application, but feel free to contact us if we missed your appl
About hud
HUD (YC W25) is developing agentic evals and RL environments for Computer Use Agents (CUAs) that browse the web for frontier AI labs. Our CUA Evals framework is the first comprehensive evaluation tool for CUAs. People don't actually know if AI agents are working reliably. To make AI agents work in the real world, we need detailed evals for a huge range of tasks. We're backed by Y Combinator, and work closely with frontier AI labs to provide agent evaluation and training infrastructure at scale.
Full hud profile