About the Role
Every game on Pax Historia depends on our ability to send huge numbers of requests across many AI providers and get the right response back quickly, reliably, and cheaply. You will work on provider reliability, structured outputs, caching, routing, and monitoring to ensure that the 37+ models on our site reliably handle our 30 billion+ monthly tokens and reduce costs and latency for our 60k+ daily active users.
Responsibilities
- Own the system for sending requests across AI providers and getting responses.
- Work on provider reliability, structured outputs, caching, routing, and monitoring.
- Ensure that the 37+ models on our site reliably handle our 30 billion+ monthly tokens.
- Reduce costs and latency for our 60k+ daily active users wherever possible.
Requirements
- Obsessed with detail, data oriented, and have a strong tendency to verify everything.
- Fast learner with extremely strong fundamentals.
- Not in charge of model-training.
- Do not run model inference ourselves.
Location
- San Francisco
Work Type
- In-person
- 5+ days/week
Experience Level
- Founding engineer
Benefits
- Meaningful equity
About the Company
- Pax Historia is creating the future of interactive entertainment.
- We’re building the platform for worldbuilders, storytellers, and game devs to create and publish any ‘what if’ they can think of.
- Most of the 25k+ user published presets on our site are historical, but fantasy and sci fi are the fastest growing categories.
- We have seen over 92 million rounds played and raised a 10M+ seed round in March.
- We’re backed by Y Combinator (W26), Bessemer, Pace Capital, Z Fellows, and more.
