Trase Systems is AI, Uncomplicated. Trase empowers enterprise leaders to harness the full potential of AI without the associated complexity and risks. We are an end-to-end solution for deploying, managing, and optimizing AI in the enterprise. Our platform specializes in bridging the “last mile” of AI adoption, unlocking AI's full potential while driving efficiency and significant cost savings.
🏢 About Trase
Co-founded in 2023 by Joe Laws and Grant Verstandig, Trase Systems is AI, Uncomplicated. Trase empowers enterprise leaders to harness the full potential of AI without the associated complexity and risks. We are an end-to-end solution for deploying, managing, and optimizing AI in the enterprise. Our platform specializes in bridging the “last mile” of AI adoption, unlocking AI's full potential while driving efficiency and significant cost savings. Trase is at the forefront of AI Agent innovation, topping the Hugging Face GAIA Leaderboard for Generalized AI Assistants, ahead of industry giants such as Google, Meta, Microsoft, and OpenAI. We are leveraging our cutting-edge technologies to develop mission-critical agentic applications in complex industries such as Healthcare, Oil & Gas, and National Security.
🎯 The Role
As a Senior or Staff DevOps Engineer on Platform Infrastructure, you will design and operate the infrastructure that supports Trase OS across Trase-hosted and customer-controlled environments. You will build secure, repeatable deployment patterns spanning AWS, Microsoft Azure, Google Cloud Platform, private cloud, hybrid cloud, and on-premises infrastructure. This is a hands-on engineering role with broad ownership. You will work across application packaging, Kubernetes, infrastructure as code (IaC), networking, security, observability, release engineering, and production reliability. You will also partner directly with engineering and customer-facing teams to turn deployment requirements into systems that can be installed, upgraded, operated, and supported consistently. The level will reflect your experience and demonstrated scope. Staff-level candidates will be expected to lead architecture across teams, establish engineering standards, and mentor other engineers.
✅ Key Responsibilities
Architect, build, and operate secure infrastructure across AWS, Microsoft Azure, and Google Cloud Platform, as well as private-cloud, hybrid-cloud, on-premises, and customer-controlled environments.
Containerize and package Trase OS application services using Docker, Kubernetes, Helm, Kustomize, or equivalent tools.
Create reusable infrastructure-as-code (IaC) modules and deployment workflows using Terraform, Pulumi, or comparable tooling.
Design deployment patterns that account for customer-specific requirements such as restricted networks, limited or no egress, approved registries, data residency, and cloud account ownership.
Identify, replace, or abstract hard dependencies on managed cloud services when they prevent portability across deployment environments.
Troubleshoot complex issues across applications, Kubernetes clusters, cloud services, networks, and infrastructure rather than treating platform work as CI/CD scripting alone.
Build and maintain CI/CD and GitOps workflows, release orchestration, environment promotion, upgrade paths, rollback procedures, and version compatibility controls.
Build and operate observability for Trase OS across metrics, logs, traces, dashboards, alerting, and SLOs so teams can diagnose failures and support customer deployments.
📌 Required Qualifications
10+ years of software, platform, infrastructure, SRE, or DevOps engineering experience, including ownership of production systems.
Deep hands-on experience with Linux, Docker, Kubernetes, and production cluster operations.
Strong experience with Helm and infrastructure as code, such as Terraform or Pulumi, including reusable modules, state management, testing, and change review.
Hands-on infrastructure experience with at least two of AWS, Azure, and GCP, with working knowledge of core compute, networking, storage, identity, and managed-service patterns across all three.
Experience deploying and operating software across customer-controlled, private-cloud, hybrid-cloud, or on-premises environments, including adapting cloud-native SaaS products for these deployment models.
Strong understanding of networking, DNS, ingress, load balancing, certificates, IAM, secrets management, persistent storage, and service-to-service security.
Experience building and operating CI/CD or GitOps systems, including production releases, upgrades, and rollbacks, with metrics, logs, traces, SLOs, and alerts for incident response.
Strong software engineering and automation skills in Python, Go, TypeScript, or a similar language, with the ability to work across application and infrastructure layers.
Experience designing secure, portable production infrastructure, including hardening systems and replacing or abstracting cloud-specific managed services when needed.
Experience converting a cloud-native SaaS product into a customer-hosted, private-cloud, or on-premises deployment model.
Demonstrated experience using AI-assisted coding and engineering tools to accelerate development, infrastructure automation, troubleshooting, operational analysis, or incident investigation.
⭐ Desirable Experience
Experience operating in restricted-network, disconnected, regulated, or security-sensitive environments.
Familiarity with compliance and security frameworks such as HIPAA, SOC 2, NIST, FedRAMP, or related government requirements.
Experience with service mesh, policy as code, admission controls, software supply-chain security, artifact signing, or software bills of materials.
Hands-on experience with Crossplane or meaningful contributions to CNCF projects, such as code, documentation, design, or community maintenance.
Experience supporting long-running, stateful, data-intensive, AI/ML, or GPU-enabled workloads.
Customer-facing engineering, solutions architecture, or forward-deployed engineering experience.
🎁 Benefits
100% employer-paid, comprehensive health care including medical, dental, and vision for you and your family.
Paid maternity and paternity for 14 weeks at employees' normal pay.
Unlimited PTO, with management approval.
Opportunities for professional development and continued learning with educational reimbursements.
Optional 401K, FSA, and equity incentives available.
Please let Red Cell know that you found this role at devopsprojectshq.com as a way to support us, so we can keep providing you with awesome DevOps jobs.
Never miss a job
Join 2,000+ DevOps developers getting weekly alerts for remote and US/EU roles, Kubernetes, AWS, Terraform, filtered for your stack.
🔒 Need an IP to whitelist?
Get a dedicated static EU outbound IP for Banks, payments, EHRs, APIs, AI.