Astronomer, a leader in unified DataOps platforms powered by Apache Airflow®, is looking for a Customer Reliability Engineer (CRE) to join their Hyderabad team on a full-time, hybrid basis. The CRE team is essential in ensuring customers’ success with Astronomer’s managed Airflow service by operating, monitoring, and maintaining the platform to achieve optimal availability and reliability. This role focuses on maintaining the reliability of cloud infrastructure and Kubernetes clusters, handling incident response, conducting root cause analysis, and continuously improving the observability platform. The position involves direct interaction with customers across various industries and cloud environments, offering a unique opportunity to influence product development and enhance the overall customer experience. The hybrid work model requires a minimum of three days per week in the Hyderabad office, fostering a balance of digital and in-person collaboration.

Key Responsibilities
- Deliver effective solutions to customers to ensure their success with Astronomer products.
- Troubleshoot customer environments and actively triage issues alongside customers.
- Provide actionable feedback to product development teams based on customer needs and pain points.
- Develop and maintain monitoring and alerting systems to improve platform reliability.
- Build automation tools to streamline daily operational tasks and enhance efficiency.
- Contribute to product architecture decisions and improvements.
- Own the customer experience by prioritizing and resolving issues, meeting SLAs, and offering expert guidance toward production readiness.
- Collaborate remotely within a fully distributed team environment.
- Enhance and enrich customer-facing documentation.
- Work on a sophisticated, cloud-native product integrating with numerous external systems.
- Support 24x7 coverage through a designated 6-hour pager rotation during the workday and participate in paid weekend on-call rotations.

Required Qualifications
- Minimum of 5 years’ experience managing large, complex cloud infrastructures at scale.
- At least 3 years of hands-on experience with Kubernetes.
- Proven experience managing production distributed systems on major cloud platforms such as AWS, GCP, or Azure.
- Strong networking knowledge within one or more major cloud environments.
- Solid Linux system administration skills.
- Expertise in operating and monitoring distributed systems.
- Familiarity with observability tools and practices.
- Demonstrated experience handling customer issues, both internal and external.
- Excellent communication skills to collaborate effectively with customers and internal teams.
- Background in DevOps or CI/CD pipelines.
- Proficiency in Python scripting.
- Strong troubleshooting and problem-solving capabilities.

Preferred Qualifications and Benefits
- Experience working as a Site Reliability Engineer.
- Knowledge of Kubernetes Custom Resources.
- In-depth expertise with Microsoft Azure.
- Familiarity with Airflow or big data orchestration tools.
- Experience with Infrastructure as Code (IaC) methodologies.

Astronomer is committed to fostering a diverse and inclusive workplace and is an equal opportunity employer. The company does not discriminate based on race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. This role offers the chance to work with cutting-edge technology in a dynamic environment that values innovation and customer success.

Job Details

Total Positions:
1 Post
Job Shift:
First Shift (Day)
Job Type:
Job Location:
Gender:
No Preference
Age:
18 - 65 Years
Career Level:
Manager
Experience:
3 Years - 5 Years
Apply Before:
Sep 09, 2026
Posting Date:
Sep 03, 2026

Astronomer

· 11-50 employees - Hyderabad

What is your Competitive Advantage?

Get quick competitive analysis and professional insights about yourself
Talk to our expert team of counsellors to improve your CV!
Try Rozee Premium
I found a job on Rozee!