At Focused, we move quickly to deliver quality software that achieves client outcomes and meets their customer’s needs. We strategically partner with our clients to leverage our expertise in design and software, while our clients bring their own domain expertise. We work with a variety of clients from different industries, collaborating as we get new products to market, modernizing legacy systems, or helping teams learn the skills they need to be successful.
Our values:
Listen first • We are experts in product practices but life long learners in the domain of our customers. We research, collaborate, and understand.
Learn why • We ask questions and talk to users to understand problem spaces, objectives, and goals, which allows us to deeply invest and drive towards the outcomes of our clients.
Love your craft • We love diving into a variety of domains and solving problems. We take pride in delivering value, in communicating progress, and guiding our clients to success.
We are seeking an experienced Observability Consultant with deep expertise in OpenTelemetry and strong Platform Engineering capabilities to help organizations implement, optimize, and scale their observability infrastructure. This role requires a specialist who can design comprehensive telemetry strategies, implement distributed tracing solutions, establish robust monitoring practices, and interface closely with clients on the observability journey.
Key Responsibilities:
OpenTelemetry & Observability
Design and implement end-to-end OpenTelemetry solutions across diverse technology stacks
Configure and deploy OpenTelemetry Collectors for efficient data collection, processing, sampling, and routing
Establish telemetry pipelines for metrics, traces, and logs across microservices architectures
Optimize collector configurations for performance, reliability, and cost-effectiveness
Platform Engineering & Infrastructure
Augment existing infrastructure with integrated observability solutions
Implement Infrastructure as Code (IaC) solutions using Terraform, Pulumi, CloudFormation, etc.
Architect and manage Kubernetes clusters with comprehensive monitoring and logging
Build CI/CD pipelines with embedded observability and automated testing
Site Reliability Engineering (SRE)
Establish and maintain Service Level Indicators (SLIs), Objectives (SLOs), and Agreements (SLAs)
Implement error budgets, toil reduction strategies, and capacity planning
Support incident response procedures and post-mortem processes
Cloud & DevOps Engineering
Deploy and manage observability infrastructure across AWS, GCP, and Azure
Establish security, compliance, and governance frameworks for telemetry data
Experience automating Agent Evaluations in CI/CD pipelines and observability backends.
Required Qualifications:
Core Observability & OpenTelemetry
3-5 years of experience in observability, monitoring, and distributed systems
Deep hands-on experience with OpenTelemetry ecosystem, including SDKs, APIs, and specifications
Proficiency with OpenTelemetry Collector configuration, processors, exporters, and receivers
Strong understanding of telemetry data models, semantic conventions, and instrumentation best practices
Platform Engineering & DevOps
5+ years of Platform Engineering or DevOps experience with focus on site reliability, observability, and incident response
Proficiency with Infrastructure as Code tools (Terraform, Pulumi, CloudFormation, CDK)
Please let Focused know that you found this role at devopsprojectshq.com as a way to support us, so we can keep providing you with awesome DevOps jobs.
Never miss a job
Join 2,000+ DevOps developers getting weekly alerts for remote and US/EU roles, Kubernetes, AWS, Terraform, filtered for your stack.
🔒 Need an IP to whitelist?
Get a dedicated static EU outbound IP for Banks, payments, EHRs, APIs, AI.