Software Development๐ฅ 1,001-5,000 employees๐ New York, NY (US)Est. 2007
Taboola empowers businesses to grow through performance advertising technology that goes beyond search and social and delivers measurable outcomes at scale.
Job Overview
Join our team as a Site Reliability Engineer at Taboola, a leader in performance-driven advertising technology. This engineering position is ideal for those passionate about working with advanced technologies in a dynamic, remote work environment. As a Site Reliability Engineer, you will enhance service reliability and tackle complex infrastructure challenges using our sophisticated technology stack.
Key Responsibilities
Ensure Reliability & Scalability: Design, implement, and manage highly reliable and scalable distributed systems across our on-premise, cloud, and AI/ML environments.
Drive Automation: Automate repetitive tasks, infrastructure provisioning, configuration, and deployments using IaC and scripting languages such as Python, Go, Rust.
Develop Observability & Capacity: Implement comprehensive monitoring and alerting systems to ensure system health using tools like Prometheus, Grafana, and ELK.
Maintain Security & Compliance: Integrate security best practices and ensure compliance with industry standards.
Lead Incident Management: Participate in on-call rotations, lead incident responses, and conduct root cause analysis to minimize downtime.
Foster Collaboration & Improvement: Work closely with development, operations, and security teams to drive shared responsibility and continuous improvement in SRE practices.
Required Qualifications
7+ years of experience as a Site Reliability Engineer, DevOps Engineer, or System Administrator in a large distributed environment with a focus on Linux operating systems.
Proven ability to support, troubleshoot, and scale large distributed systems in production.
Deep understanding of HTTP protocol, including HTTP/1.1, HTTP/2, caching semantics, TLS, and gRPC delivery.
Experience configuring and operating CDN services such as Akamai, Fastly, Cloudflare, AWS CloudFront.
Expertise in Linux system internals and system performance tuning.
Proficiency with Configuration Management Tools like Puppet, Ansible, Chef, Terraform.
Programming skills in languages such as Python, Golang, Rust, Ruby, C++, Java.
Familiarity with monitoring and metrics collection systems like Prometheus, Grafana, ELK.
Experience with cloud providers and platforms including AWS, Azure, GCP, Alibaba.
Knowledge of containerization technologies such as Kubernetes, Docker.
Deep understanding of networking principles including TCP/IP, DNS, and load balancing.
Preferred Qualifications
Advanced certifications in cloud technologies, security, or network management are highly desirable.
About Taboola
At Taboola, we empower our employees to realize their full potential in an environment that fosters growth, collaboration, and learning with talented and smart colleagues. Our unique company culture is the backbone of our success in the digital advertising space.
Benefits & Perks
Enjoy comprehensive benefits including health, dental, and vision, along with perks like a fully stocked kitchen and location-specific benefits such as gym partnerships and parking.
Please let Taboola know that you found this role at devopsprojectshq.com as a way to support us, so we can keep providing you with awesome DevOps jobs.
Never miss a job
Join 2,000+ DevOps developers getting weekly alerts for remote and US/EU roles, Kubernetes, AWS, Terraform, filtered for your stack.
🇪🇺 Secure Your EU Traffic
Ensure digital sovereignty for your infrastructure. Get EU static IPs with full data residency for compliance and peace of mind.