The roleBase: $200-275k •
Equity: 0.25-0.5% •
Location: San Francisco, 5 days in office
Own platform engineering, develop frontier long-horizon RL environments, and help build our engineering org from the ground up.
Responsibilities- Build our RL environment training and inference infrastructure to support customer usage at scale.
- Research and develop next-generation RL environments - increasingly realistic, long-horizon, and difficult for frontier models.
- Build software that 10-100x's the quality and throughput of RL environment creation.
- Innovate on synthetic data pipelines to create realistic problems.
- Build platform analytics for environment cost, hours, bottlenecks, and SME management.
- Craft domain-specific verifiers (verifiable rewards for decks, Excel modeling, quantitative trading, and beyond).
- Establish engineering culture and practices from the ground up.
You should have- Familiarity building evaluations, benchmarks, or RL for AI agents.
- Startup speed: iterate quick, ask quick, respond quick, make mistakes quick.
- Strong product and user ownership; ability to prioritize across a large roadmap.
- Client-facing comfort - you'll talk to users, customers, and SMEs.
Nice to have- Research background (training models, publishing papers).
- Former founders or experience at early-stage startups.
What we offer- Healthcare - Fully covered medical, dental, and vision for you. Subsidized for dependents.
- Relocation - $10k+ relocation stipend to help you get to San Francisco.
- 401(k) - Up to 4% employer match.
- Food - Free breakfast, lunch, and dinner while working.
- Health - Free gym membership.
- Transportation - Free Ubers and Waymo rides to and from the office.
- Visa sponsorship - We sponsor visas, including H-1B.