Member of Technical Staff, Infrastructure ($120K – $250K + Equity) at The Token Company
Jack & Jill · San Francisco, CA · Yesterday
On-siteEngineering$120k–$250k/yrFull-time
About the role
You will own the end-to-end cloud infrastructure for a high-performance compression API. This role involves building global, low-latency GPU ML inference systems that sit directly in the critical path of customer traffic. As a core member of a 5-person team, you will drive reliability, scalability, and cost-efficiency for a revolutionary AI product.
Why this role is remarkable
- Join an elite 5-person team backed by $12M from top-tier investors like First Round Capital and founders of Dropbox, Slack, and Supercell.
- Solve critical infrastructure challenges for LLM efficiency, building the low-latency systems that enable enterprises to scale AI cost-effectively.
- Benefit from an unparalleled San Francisco living experience with provided housing, food, laundry, and cleaning services, plus full visa sponsorship.
What you will do
- Architect and manage global, high-throughput GPU inference infrastructure using AWS, Terraform, and Docker to ensure maximum performance and uptime.
- Own the entire infrastructure stack from initial deployment and CI/CD pipelines to long-term reliability and cost-efficiency research.
- Collaborate closely with research and product teams to integrate and scale compression models into a high-traffic production API.
The ideal candidate
- Demonstrated experience building and operating production-grade cloud infrastructure at scale using tools like AWS and Terraform.
- Strong background in managing containerized environments with Docker and establishing robust CI/CD pipelines for high-reliability systems.
- Proven ability to take end-to-end ownership of an infrastructure stack, ideally within a fast-paced startup or high-growth engineering environment.
Pay
$120K – $250K + Equity
Location
San Francisco, USA