You are the infrastructure expert who enables our rapid product development and guarantees 99.9%+ stability and performance of our clinical AI platform for major health systems. Your focus on operational excellence is directly tied to a patient's access to life-saving treatment.
Key Responsibilities
Infrastructure Ownership: Design, implement, and maintain the production environment, having previously handled 500+ machine deployments.
Kubernetes Mastery: Own our containerized infrastructure, leveraging deep expertise in Kubernetes and Helm to manage deployment, scaling, and operational health.
CI/CD & Deployment Optimization: Optimize and streamline both the TypeScript and Python/ML deployment pipelines to support high-velocity feature release while maintaining the highest reliability.
DevX Support: Support Developer Experience (DevX) work to streamline developer workflows, enhance tool proficiency, and improve CI/CD systems.
Infrastructure as Code (IaC): Manage and maintain infrastructure definitions using Terraform.
Required Qualifications
Tool Proficiency: Highly proficient with command line and keyboard shortcuts.
Ownership: Proven track record of scaling mission-critical deployments.
Automation Drive: Passion for automating processes and defining standards for operational excellence.
Problem Solver: Willingness to jump into whatever needs to get done without waiting for others.
Technical Qualifications: Deep experience with Kubernetes, Helm, Terraform, PostgreSQL, Redis, and Kafka.
Core Team Member: Excitement about working five days per week in our San Francisco office.