P

Founding Engineer - AI Infrastructure and Systems

pax historia United State
Visa Sponsorship
Apply
AI Summary

Build and optimize the core infrastructure that powers Pax Historia’s AI-driven interactive storytelling platform. Own the system responsible for routing, caching, and monitoring requests across 37+ AI models to serve 60k+ daily active users. Focus on reliability, cost reduction, and latency optimization without model training or inference.

Key Highlights
Own the AI request routing, caching, and monitoring infrastructure for a platform serving 30B+ monthly tokens
Optimize reliability, cost, and latency for 37+ AI models across 60k+ daily active users
Work in-person in San Francisco as the 7th team member with meaningful equity compensation
Key Responsibilities
Own the system responsible for routing AI requests across 37+ models to ensure reliable, fast, and cost-effective responses
Optimize caching, monitoring, and provider reliability to handle 30B+ monthly tokens for 60k+ daily active users
Reduce latency and costs while maintaining high availability and performance for the interactive storytelling platform
Technical Skills Required
Distributed Systems API Routing Caching
Benefits & Perks
Meaningful equity compensation
Visa sponsorship

Job Description


Company

Pax Historia is creating the future of interactive entertainment.

We’re building the platform for worldbuilders, storytellers, and game devs to create and publish any ‘what if’ they can think of. So far, most of the 25k+ user published presets on our site are historical (eg, ‘what if Rome never fell?’), but fantasy and sci fi are the fastest growing categories on the site.

We have seen over 92 million rounds played and raised a 10M+ seed round in March. We’re backed by Y Combinator (W26), Bessemer, Pace Capital, Z Fellows, and more.

Role

Every game on Pax Historia depends on our ability to send huge numbers of requests across many AI providers and get the right response back quickly, reliably, and cheaply.

We’re Hiring An Engineer To Own That System. You’ll Work On Provider Reliability, Structured Outputs, Caching, Routing, And Monitoring. At The End Of The Day, You Will Need To

  • Ensure that the 37+ models on our site reliably handle our 30 billion+ monthly tokens.
  • Reduce costs and latency for our 60k+ daily active users wherever possible.

You will not be in charge of model-training. We do not run model inference ourselves.

If you’re obsessed with detail, data oriented, and have a strong tendency to verify everything, then this might be a great fit for you. We don’t expect you to have done this exact job before, but we are expecting you to be a fast learner with extremely strong fundamentals.

Logistics

You’ll be our 7th team member. We work in-person in San Francisco, 5+ days/week. We are open to sponsoring visas.

Comp will include meaningful equity to reflect your responsibility as a founding engineer.

Similar Jobs

Explore other opportunities that match your interests

Fuel Systems Test Engineering Manager

Programming
49m ago

Premium Job

Sign up is free! Login or Sign up to view full details.

•••••• •••••• ••••••
Job Type ••••••
Experience Level ••••••

Caterpillar Inc.

United State

Applied AI Field Researcher

Programming
4h ago

Premium Job

Sign up is free! Login or Sign up to view full details.

•••••• •••••• ••••••
Job Type ••••••
Experience Level ••••••

anthropic

United State
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Associate

torentify

United State

Subscribe our newsletter

New Things Will Always Update Regularly