InfrastructureAWSPlatformCI/CDKubernetesTerraformCloudFormationJenkinsIaCGitlab CI
Job Overview
We are building a robust, scalable trading platform to serve high-traffic, latency-sensitive applications. Our infrastructure leverages state-of-the-art technologies to support real-time trading while providing unparalleled reliability and performance. Join us to shape the future of our platform and engineering culture.
Key Responsibilities
Platform Engineering: Architect and implement scalable infrastructure to support the deployment and management of our trading platform.
Developer Tooling: Build and maintain internal tools to streamline developer workflows, including advanced CI/CD pipelines.
Infrastructure as Code (IaC): Champion IaC practices using Terraform, CloudFormation, or Pulumi.
Core Services Management: Manage and optimize platform-critical services such as NATS Cluster, RabbitMQ, AWS RDS PostgreSQL, Redis Cluster.
DevOps Automation and CI/CD: Automate and optimize deployment processes to ensure seamless continuous integration and delivery.
Container Orchestration: Manage and scale containerized workloads using Kubernetes and Docker.
Cloud Optimization: Monitor and optimize cloud resource usage for performance and cost efficiency.
Site Reliability Engineering (SRE): Define and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs).
Monitoring & Observability: Implement observability tools and dashboards (e.g., Prometheus, Datadog, Grafana) for real-time system monitoring.
Incident Management: Lead incident response efforts, conduct root cause analysis, and implement actionable postmortem reviews.
Infrastructure Management: Architect and manage cloud-based systems to handle high-traffic, latency-sensitive applications.
Disaster Recovery: Implement robust disaster recovery and business continuity strategies, including backups and multi-region failover.
Security Practices: Collaborate with security teams to enforce best practices for IAM, encryption, and compliance.
Collaboration & Leadership: Partner with software engineers to design infrastructure solutions tailored to their application needs.
Culture Building: Help shape the engineering culture, promoting a philosophy of security, velocity, and reliability.
Mentorship: Mentor junior engineers and document best practices to drive knowledge sharing and operational excellence.
Long-Term Tech Evolution: Contribute to evolving our backend microservices (currently NodeJS, with some Python and C#) towards Go and Rust.
Third-Party Integration: Evaluate and integrate critical third-party software and infrastructure, such as payment gateways and mobility stacks.
Required Skills & Experience
Technical Expertise: 5-8+ years of hands-on experience with cloud platforms, particularly AWS, including services like EC2, RDS, S3, Lambda, and VPC.
Containerization: Proficiency with Docker and Kubernetes (EKS) or ECS.
Infrastructure as Code (IaC): Strong experience with Terraform, CloudFormation, or Pulumi.
Programming Skills: Proficiency in at least one programming language (e.g., Python, Go, TypeScript/JavaScript, Ruby, Java).
DevOps & SRE: Expertise in building and maintaining CI/CD workflows using tools like GitLab CI, Jenkins, or GitHub Actions.
Monitoring Tools: Experience with observability platforms (e.g., Prometheus, Datadog, Grafana).
Incident Management: Proven ability to handle incident response, root cause analysis, and postmortem reviews.
Soft Skills: Ability to research, design, and deliver solutions to complex infrastructure challenges.
Collaboration: Experience working directly with product engineers to improve workflows incrementally.
Leadership: Ownership mindset with the ability to mentor team members and advocate for best practices.
Preferred Skills (Nice-to-Have)
Familiarity with backend languages like Go or Rust.
Experience with networking concepts (e.g., load balancers, DNS, VPNs) and traffic optimization.
Knowledge of emerging CNCF technologies and CI/CD trends.
What We Offer
Competitive salary with future equity options Opportunities to work with cutting-edge technologies and evolve our platform. Flexible working hours and a remote-friendly environment. Professional growth through certifications, conferences, and internal training.
Please let OnHires know that you found this role at devopsprojectshq.com as a way to support us, so we can keep providing you with awesome DevOps jobs.
Never miss a job
Join 2,000+ DevOps developers getting weekly alerts for remote and US/EU roles, Kubernetes, AWS, Terraform, filtered for your stack.
🔒 Need an IP to whitelist?
Get a dedicated static EU outbound IP for Banks, payments, EHRs, APIs, AI.