Volta builds and operates large scale GPU compute infrastructure for AI workloads. Our platform is Kubernetes-native, spans multiple regions, and delivers virtual machines, storage, and networking through a fully automated infrastructure stack built on custom Kubernetes operators. We are building out several platform engineering teams that together own the full stack, from managed bare metal and IaaS through to higher-order platform services.
🏢 About Volta
Volta is the category-defining, fully vertically integrated AI infrastructure platform – from capital to clusters to software, under a founder-led enterprise. Our mission is The Utility of Compute™: AI infrastructure as dependable and available as electricity, for every organization that needs it. Launched with a $10B strategic partnership with one of the leading frontier AI labs, a Series A led by Andreessen Horowitz, and a $5B AI Infrastructure Fund, Volta is building the infrastructure layer of the AI era from the ground up. We are 100+ people across London, Palo Alto, and New York, with rapid growth expectations to hundreds.
🎯 The Role
Platform Engineers work at the intersection of infrastructure and software development. Across every team, you will translate three key inputs into durable platform capabilities: product roadmap requirements from the product team, operational learnings from the bring-up teams, and security guidance from the security engineering team. The output of this role is production platform code, not configuration, not runbooks.
✅ Key Responsibilities
Design and implement Kubernetes operators and controllers that manage the lifecycle of platform resources.
Work closely with the product team to turn roadmap requirements into the platform capabilities that support them.
Collaborate with the bring-up teams to identify operational pain points and turn them into scalable platform features.
Integrate security guidance from the security engineering team into platform-level controls, and remediate findings at the platform layer.
Treat observability as a platform concern: instrument services, define meaningful metrics, and build tooling that gives the team visibility into platform health.
Own the services you build in production, including participation in an on-call rotation, incident response, and the follow-up work that closes structural gaps rather than only the immediate issue.
Hold to clean interface and versioning practice on anything other teams or customers depend on, including disciplined handling of breaking changes.
Participate in code review, technical design discussions, and cross-team collaboration in an Agile (Kanban or Scrum) environment.
📌 Required Qualifications
3+ years of software engineering experience, with a meaningful portion spent on infrastructure or platform systems.
Strong backend or systems programming experience in a production environment. Our working languages are Python, Go, and Rust; we welcome strong engineers from other compiled or object-oriented languages (for example C++, C#, or Java) who are ready to work across our stack as it evolves.
Solid understanding of Kubernetes internals: the control loop model, CRDs, controllers and operators, and reliable reconciliation logic.
Comfortable working close to the infrastructure layer: Linux, networking fundamentals, and distributed systems behavior.
Experience designing, building, and versioning production-grade APIs or service interfaces that other teams depend on, including disciplined handling of breaking changes and backward compatibility.
Experience operating what you build: debugging production systems, and taking part in on-call or incident response.
Strong engineering fundamentals: clean code, testing, version control, code review, and CI/CD practices.
⭐ Desirable Experience
Fluency with AI-assisted development: agentic CLI tools, IDE assistants, and orchestrating multiple coding agents through MCP, skills, or APIs to amplify delivery.
Depth in Go or Rust beyond working proficiency.
Familiarity with confidential computing technologies: TEEs, AMD SEV, Intel TDX, or Confidential Containers (CoCo).
Experience with large-scale distributed systems and cloud infrastructure.
Experience with container orchestration and management tools beyond Kubernetes.
Understanding of networking protocols and infrastructure design.
Experience with infrastructure as code tools like Terraform or Pulumi.
The posted range reflects the typical compensation for this role. Actual compensation is based on market rate, factoring in qualifications, experience, location, and the value you bring - while ensuring fairness and consistency across our organization. Eligible employees also receive a total rewards package, including a discretionary bonus, equity awards, and comprehensive benefits.
🛂 Visa & Eligibility
None of these are required. Several map to specific teams, so strength in one or more helps us match you to the right one.