We are seeking a skilled DevOps Engineer with 3 to 6 years of experience in DevOps, Site Reliability Engineering (SRE), cloud engineering, or infrastructure management. The ideal candidate will have hands-on expertise with major cloud platforms such as AWS, Tencent Cloud, or Azure, and a solid understanding of core cloud services. This role requires strong technical skills in CI/CD pipelines, containerization, infrastructure automation, and security best practices, along with experience supporting high-traffic OTT or streaming platforms.
Key Responsibilities
- Design, build, and maintain CI/CD pipelines for web, mobile, backend, and data services to ensure smooth and efficient software delivery.
- Manage cloud infrastructure primarily on AWS, with additional exposure to Azure environments as needed.
- Automate infrastructure provisioning and configuration using Infrastructure as Code tools such as Terraform.
- Containerize applications and manage them using Docker and Kubernetes to support scalable and reliable deployments.
- Implement comprehensive monitoring, logging, alerting, and observability practices to maintain system health and performance.
- Ensure secure, reliable, and repeatable production deployments by following operational best practices.
- Collaborate closely with Engineering, QA, Security, and Product teams to align infrastructure needs with business goals.
- Investigate and resolve incidents, performance issues, and infrastructure bottlenecks through thorough root-cause analysis.
- Support backup, disaster recovery, scalability, and availability strategies to maintain high service uptime.
- Continuously improve deployment speed, operational efficiency, and infrastructure cost visibility.
Required Qualifications
- 3 to 6 years of professional experience in DevOps, SRE, cloud engineering, or infrastructure roles.
- Proven hands-on experience with cloud platforms such as AWS, Tencent Cloud, or Azure, including a strong understanding of core cloud services.
- Expertise in CI/CD tools like Jenkins, GitHub Actions, GitLab CI, or equivalent.
- Solid experience with containerization technologies including Docker and Kubernetes.
- Strong Linux system administration skills, along with networking and monitoring knowledge.
- Familiarity with Infrastructure as Code tools, particularly Terraform.
- Good understanding of security principles, Identity and Access Management (IAM), secrets management, and operational best practices.
- Experience supporting high-traffic OTT or streaming platforms.
- Knowledge of autoscaling, Content Delivery Networks (CDN), caching mechanisms, databases, and distributed systems.
- Relevant certifications in AWS, Azure, Kubernetes, or Terraform are highly desirable.
- A mindset focused on automation and reliability, with strong incident response and root-cause analysis capabilities.
- Security-conscious approach to infrastructure management, emphasizing clear documentation and effective technical communication.
Preferred Qualifications and Benefits
While not explicitly listed, candidates with additional certifications and experience in cloud security, advanced monitoring tools, or large-scale distributed systems will be well-positioned for success. The role offers the opportunity to work in a collaborative, fast-paced environment where innovation and continuous improvement are highly valued. Candidates can expect to contribute to critical infrastructure that supports high-availability streaming services, gaining exposure to cutting-edge cloud technologies and best practices.
This position is ideal for professionals who thrive in dynamic environments and are passionate about building scalable, secure, and efficient cloud infrastructure.