WATI
· Tsim Sha Tsui, Kowloon, Hong Kong
· Full time
LocationTsim Sha Tsui, Kowloon, Hong Kong
EmploymentFull time
Role typeEngineering
PostedApr 08, 2024
W
WATI
IT Services and IT Consulting👥 201 employees📍 Manhattan Beach, California, USEst. 1998
WATI (West Advanced Technologies, Inc.) is a California-headquartered technology solutions provider with significant presence in USA & India. Since 1998, WATI is delivering 100% success for clients f…
Description
WATI is an early-stage, fast-growing SaaS platform that is revolutionizing how companies communicate with their customers. Through our cutting-edge customer engagement software built on Whatsapp’s Business API, businesses are now able to have personalized conversations, be easily accessible, and engage with their customers in real-time - at scale! We live in the on-demand economy, where customers expect fast, simple, and easy service and that’s exactly what our platform empowers companies to do.
This is made possible through WATI’s easy-to-use platform which can be made up and running in no time. As a result, small and medium businesses have embraced the platform rapidly, and thousands of customers across 54 countries are now using WATI within just a year of launch.
We are seeking a highly motivated and experienced Site Reliability Engineer (SRE) to join our team. As an SRE, you will be responsible for ensuring the reliability, performance, and availability of our production systems. You will work closely with our development team to design, implement, and maintain infrastructure and tools that support the delivery and operation of our software. This role is to complete our 7x24 global technical support service.
Key Responsibilities:
Collaborate with the development team, DevOps team, and QA team to identify and address reliability issues in the software development lifecycle
Implement and maintain monitoring, alerting, and reporting systems to ensure the health and availability of our production systems
Diagnose and resolve production issues in a timely and effective manner
Conduct post-mortem analysis of production incidents and identify areas for improvement
Contribute to the development of best practices and guidelines for reliability and infrastructure management
Requirements
Bachelor's degree in Computer Science or a related field, or equivalent work experience
3+ years of experience as an SRE or in a related field such as DevOps or systems engineering Site Reliability Engineer
Strong knowledge of public cloud infrastructure. Preferably GCP.
Experience of managing and querying cloud database provider like Mongo Atlas.
Experience with configuration management and automation tools such as Terraform, Ansible, or Puppet Experience with containerization technologies such as Docker and Kubernetes
Experience with monitoring and alerting tools such as Prometheus, Grafana, and Datadog
Experience with programming languages such as NodeJS, Python, C#, or Bash
Strong problem-solving and communication skills
Ability to work in a fast-paced and dynamic environment