Beamery’s AI talent platform empowers companies to understand the skills and capabilities they have, build more agile workforce plans, and attract, retain, upskill and redeploy their workforce. With …
Job Overview
Beamery is seeking a Principal Platform Engineer (SRE/Cloud) to join our Engineering team in London, UK. This is a full-time position that offers the opportunity to solve the toughest reliability, scalability, and infrastructure problems with the highest impact.
Key Responsibilities
Design solutions, produce architectures, and write RFCs for platform, reliability, and infrastructure initiatives which have impact across Engineering and beyond
Maintain hands-on credibility by helping teams ship scalable, reliable, and cost-effective services
Partner with Product and Engineering Directors on planning by scoping and sizing large new platform initiatives that span multiple teams
Work with other Principal Engineers to create and update the architectural vision that will deliver on Beamery's strategy
Engage with other Principal Engineers in setting and advocating company-wide standards for operational excellence, observability, reliability, and incident response
Take a whole-company view of major incidents, identifying recurring themes and turning them into company-level investments
Own and evolve the platform and infrastructure tech radar, making data-backed cases for the adoption and deprecation of technologies and services
Set the standard for operational ownership, supporting teams to own their services end-to-end
Coaching and mentoring Engineers all the way up to Staff level from teams across the organization
Collaborate on customer solutions with non-technical stakeholders including Sales, Customer Success, and our high profile, Fortune 500 customers
Stay connected to production through on-call, and use that vantage point to raise the bar for incident response, observability, and reliability across the whole engineering organization
About Beamery
With the rapid expansion of AI and automation, the future of work has never been more challenging for organizations. Beamery’s unique jobs, skills, and tasks data platform helps organizations navigate these challenges and make more informed decisions across Talent Lifecycle Management. Our solutions power recruitment, mobility, upskilling, diversity, work architecture, and workforce planning for some of the world’s most forward-thinking companies.
We believe that where you work is much more than just a job. Millions of people are being left behind every day in their careers, and we’re on a mission to fix this by creating equal access to meaningful work, skills, and careers for all. We are an equal opportunity employer committed to building a representative Beamery, creating an equitable, inclusive, and engaging environment for our people.
Required Qualifications
A proven track record of designing and delivering scalable, reliable cloud-based infrastructure and platform services
Extensive experience running and supporting services in production environments, including on-call leadership and incident command at scale
Previous experience as an individual contributor in an Engineering leadership position (Staff+, Principal, Architect, etc.)
Deep expertise managing Kubernetes production clusters at scale — cluster lifecycle and upgrades, resource optimization (autoscaling, quotas), security (RBAC, Network Policies), high availability, and troubleshooting complex networking or scheduling issues
Expertise with Infrastructure as Code (IaC), particularly Terraform, and modern GitOps practices
Strong software engineering foundations with the ability to build and maintain production-grade tooling. Experience with Go would be preferable. NodeJS is a plus
A strong understanding of observability, SLOs, alerting, and cost management for large-scale systems
Operational experience with our key infrastructure components: Kafka, MongoDB, PostgreSQL, Elasticsearch, and Istio
Familiarity with LLMOps is a plus; helping shape how we operate LLM-powered systems at scale (model gateways and routing with LiteLLM, experiment and model tracking with MLflow) would be a welcome bonus
Please let Beamery know that you found this role at devopsprojectshq.com as a way to support us, so we can keep providing you with awesome DevOps jobs.
Never miss a job
Join 2,000+ DevOps developers getting weekly alerts for remote and US/EU roles, Kubernetes, AWS, Terraform, filtered for your stack.
🔒 Need an IP to whitelist?
Get a dedicated static EU outbound IP for Banks, payments, EHRs, APIs, AI.