Infrastructure Engineer, Kubernetes Specialist
About TypeSafe
TypeSafe is a frontier model lab building reliable, general AI systems for real-world automation.
About the role
Build and operate the infrastructure behind TypeSafe’s products at global scale. You’ll own systems serving users across regions, from provisioning Kubernetes clusters across clouds to optimizing networking for low-latency AI inference.
What you’ll do
– Design, deploy, and operate Kubernetes clusters across regions and clouds
– Build infrastructure for global LLM inference workloads
– Own networking, GPU infrastructure, autoscaling, and observability
– Maintain infrastructure as code with Pulumi or Terraform
What we’re looking for
– Deep production Kubernetes experience
– Strong AWS and Linux networking fundamentals
– Infrastructure-as-code experience
– Programming fluency, ideally Python
– Experience with high-traffic ML systems is a plus
This is a full-time, five-days-per-week in-person role in San Francisco.
Compensation: $180,000-$280,000 plus equity, health insurance, daily meals, visa sponsorship, and a 401(k).