As the owner of the infrastructure tasked with handling a substantial number of requests from various AI providers, you will ensure optimal reliability, minimal latency, and cost-effectiveness. This responsibility includes overseeing structured outputs, implementing caching strategies, managing routing, and monitoring processes for over 30 billion tokens on a monthly basis. A solid foundation in engineering principles and a data-driven approach with meticulous attention to detail and verification is essential. While direct experience in this specific position is not necessary, the ideal candidate should be a quick learner adept at navigating complex AI infrastructure.
Founding Engineer – LLM Infra & Platform
San Francisco, California, 94102
Full Time
Onsite
$130000 - $220000/yr
supports..pdf, .doc, .docx, or .txt file


