Omnidian, Inc. is a fast-growing Series C tech-enabled service company revolutionizing performance assurance for the distributed solar and energy storage industries. Omnidian is building a more sustainable future for the planet through our passionate teams, our innovative technology, and by creating an amazing customer experience. We are a certified B Corp, headquartered in Seattle, WA.
🏢 About Omnidian
Omnidian, Inc. is a fast-growing Series C tech-enabled service company revolutionizing performance assurance for the distributed solar and energy storage industries. Omnidian is building a more sustainable future for the planet through our passionate teams, our innovative technology, and by creating an amazing customer experience. We are a certified B Corp, headquartered in Seattle, WA.
🎯 The Role
As the Principal SRE / DevOps Engineer, you will play a foundational role in architecting the future of our platform reliability and operational ecosystem, serving as a technical lead and strategist as we build a robust, scalable, and highly supportable infrastructure to support our clients and products.
This is a hands-on technical position with some project management and leadership responsibilities. In this role, you will work closely with our Software Product, Engineering, and Operations teams to document and evangelize a vision for platform reliability, observability, and automation. You will guide the team to break the high-level vision into well-defined milestones. You will assist in establishing and will champion and participate in best practices in workload and reliability management, including providing visibility to stakeholders and executives for progress toward our vision.
✅ Key Responsibilities
Formulate the multi-year technical roadmap for our platform reliability, infrastructure, and operational excellence, identifying where intelligent automation can significantly remove engineering and operational bottlenecks.
Obtain executive and stakeholder buy-in for a documented vision for our SRE / DevOps deliverables.
Translate the technical vision to actionable and trackable work plans, including meaningful milestones.
Translate business and reliability requirements to right-sized technical specifications.
Maintain a security-first mindset, ensuring all infrastructure, automation frameworks, and CI/CD pipelines are robustly defended.
Provide strong technical leadership and mentorship across teams.
Develop an appropriate reliability and testing strategy.
Design, implement, and maintain production-grade Kubernetes platforms, cluster management, and related orchestration.
Build and evolve Infrastructure as Code and configuration management using Ansible and Terraform.
Own and continuously improve CI/CD pipelines, deployment strategies, and release automation.
Implement and refine observability stacks centered on Grafana.
Create and maintain automation for operational toil reduction, self-healing systems, and infrastructure provisioning.
Use AI coding and ops assistants daily for scripting, infrastructure code, refactoring, pipeline improvements, and reliability testing.
Follow established agile ceremonies, including best practice metrics.
Create and deliver on visible project plans, break down milestones into distinct work, plan and assign work, and ensure timely and accurate delivery.
Clearly and proactively communicate progress against the plan including status, blockers, dependencies, and risks.
Clear blockers to timely or accurate delivery; escalate to leadership as appropriate.
Identify and proactively communicate work needed from outside the team, obtain commitment, and follow through to ensure dependencies will be delivered to plan.
Set appropriate documentation expectations for the team.
Identify, align stakeholders on, and implement improvements to our processes, codebases, and architecture.
Please let Omnidian know that you found this role at devopsprojectshq.com as a way to support us, so we can keep providing you with awesome DevOps jobs.
Never miss a job
Join 2,000+ DevOps developers getting weekly alerts for remote and US/EU roles, Kubernetes, AWS, Terraform, filtered for your stack.
🔒 Need an IP to whitelist?
Get a dedicated static EU outbound IP for Banks, payments, EHRs, APIs, AI.