GT was founded in 2019 by a former Apple, Nest, and Google executive. GT’s mission is to connect the world’s best talent with product careers offered by high-growth companies in the UK, USA, Canada, Germany, and the Netherlands. On behalf of Feeld, GT is looking for a Site Reliability Engineer (SRE) to join a fast-growing consumer mobile product in the online dating space.
🏢 About Feeld
Founded in 2014 as a dating app, Feeld gathered millions of users in one place to create a safer and more inclusive space online for everyone open to experiencing people and relationships in a new way. Their mission is to elevate the human experience of sexuality and relationships and create a world where everyone is more intimately connected to each other and themselves.
🎯 The Role
We are looking for an experienced Site Reliability Engineer with a strong backend engineering background in Node.js and TypeScript. The ideal profile is someone who started in backend/software engineering and has moved into SRE or reliability-focused work, combining a strong understanding of application code with hands-on experience in observability, monitoring and production incident management. You will be embedded within a product squad and take ownership of the reliability and observability of critical user journeys.
✅ Key Responsibilities
Own observability for critical product and user journeys within your squad.
Define, build and maintain meaningful metrics, dashboards and alerts.
Define and maintain SLIs/SLOs for key services and product-level metrics.
Improve monitoring, logging, tracing and alerting across the squad’s systems.
Act as the first responder for critical P0/P1 production incidents, including out-of-hours incidents.
Investigate production signals, identify potential root causes and begin mitigating issues independently.
Coordinate with other engineers when broader support or escalation is required.
Participate in incident triage, mitigation and postmortems.
Identify recurring reliability issues and drive improvements to infrastructure, tooling and incident-response processes.
Work closely with backend and product engineers in a distributed, autonomous squad.
📌 Required Qualifications
Strong previous experience as a Backend / Software Engineer, with senior-level hands-on experience in Node.js and TypeScript.
Hands-on experience working in an SRE, Production Engineering or similar reliability-focused role.
Strong production experience with AWS.
Experience with monitoring and observability across metrics, logging, tracing and alerting.
Practical experience responding to production incidents, including triage, mitigation and postmortems.
Ability to interpret monitoring signals and independently investigate and begin resolving production issues.
Understanding of both the application and infrastructure layers rather than infrastructure-only experience.
Strong communication skills and the ability to work autonomously within a distributed engineering team.
Comfortable participating in out-of-hours incident response as part of the team’s coverage model.
⭐ Desirable Experience
Experience with Cloudflare and CloudWatch.
Experience with observability tools such as Sentry.
Experience defining SLIs and SLOs for product-level metrics.
Experience with React Native or exposure to mobile application environments.
Previous experience with consumer mobile products or high-traffic B2C systems.