EBS Sp. z o.o. is a dynamically growing Polish technology company with over 35 years of experience, specializing in the development of smart security solutions. We are part of the American corporation Alarm.com (Nasdaq: ALRM), a leading global provider of intelligent security solutions.
Alarm.com values working together and collaborating in person. We are building a new team in Krakow. Our employees work fully from the office with the possibility to work remotely occasionally.
🏢 About Alarm.com
Alarm.com is the leading cloud-based platform for smart security and the Internet of Things. More than 7.6 million home and business owners depend on our solutions every day to make their properties safer, smarter, and more efficient. And every day, we're innovating new technologies in rapidly evolving spaces including AI, video analytics, facial recognition, machine learning, energy analytics, and more. We're seeking those who are passionate about creating change through technology and who want to make a lasting impact on the world around them.
🎯 The Role
Alarm.com is seeking a versatile Staff Software Engineer (SRE) to join our Video team. Our platform powers live streaming, recorded video, and real-time object detection and identification notifications for millions of IoT cameras worldwide. In this role, you will help keep the platform reliable at scale, advance observability, reduce incident response time, and build for long-term growth.
You will join a collaborative team that cares deeply about reliability, customer experience, and engineering quality. This is a hands-on role for someone who enjoys solving complex distributed systems problems, improving how teams detect and respond to issues, and building operational patterns that scale. You will work across software, infrastructure, and data flows, partnering with product and engineering teams to turn reliability goals into practical outcomes.
✅ Key Responsibilities
Own reliability for core production systems, including deployments, on-call operations, and incident response.
Drive capacity planning across cloud and physical infrastructure.
Define and improve reliability goals and operational outcomes (for example availability and MTTR).
Lead observability improvements across metrics, logs, traces, alert quality, and runbook maturity.
Build automation and tooling that improve reliability and engineering velocity.
Partner with engineering, product, and business stakeholders to align reliability priorities with roadmap goals.
Shape database and data-processing reliability, including performance and operational safety.
Apply sound engineering judgment to balance delivery speed with long-term resilience.
Other duties as assigned
📌 Required Qualifications
Bachelor’s degree in Computer Science, Computer Engineering, a related field, or equivalent practical experience.
10+ years of professional software engineering experience, including production operations and on-call ownership.
Proven track record of improving measurable reliability outcomes in distributed systems.
Deep expertise in observability and production support practices.
Strong networking fundamentals for distributed systems, including TCP/IP, DNS, TLS, HTTP, and L4/L7 behavior.
Strong experience with distributed systems technologies such as Kubernetes, Kafka, and Redis.
Experience with cloud infrastructure and operations at scale; Azure experience is a plus.
Strong software engineering fundamentals in object-oriented design and software development lifecycle practices; C# and .NET experience is a plus.
Strong analytical, communication, and decision-making skills under production pressure.
⭐ Desirable Experience
Experience leading high-severity incident response and driving durable post-incident improvements is a plus.
Strong database and data-processing experience, including performance tuning, schema evolution, and rollback-safe deployment practices, is a plus.
Deep experience with advanced traffic management and troubleshooting in production (for example Envoy, ingress controllers, service mesh, global load balancing, and cross-region routing) is a plus.
Experience with video streaming systems, WebRTC, IoT device ecosystems, OpenVPN, and MQTT is a plus.
🎁 Benefits
Collaborate with outstanding people: We have a strong focus on teamwork, and we work to create a collaborative and welcoming environment that enables our teams to excel.
Make an immediate impact: You can expect to be given real responsibility for bringing new technologies to the marketplace. You will be empowered to perform as soon as you join the team!
Be Empowered: We don't want to micro-manage you. We want you to own stuff and bring your experience to make those products the best in class.
Long-term employment based on a permanent employment contract (CoE).
Attractive benefits package: including medical care, life insurance, sports package, annual budget for professional development ($2,000).
Please let Alarm.com know that you found this role at devopsprojectshq.com as a way to support us, so we can keep providing you with awesome DevOps jobs.
Never miss a job
Join 2,000+ DevOps developers getting weekly alerts for remote and US/EU roles, Kubernetes, AWS, Terraform, filtered for your stack.
🔒 Need an IP to whitelist?
Get a dedicated static EU outbound IP for Banks, payments, EHRs, APIs, AI.