Contour Software, a subsidiary of Constellation Software Inc., is seeking a skilled DevOps Engineer to oversee and enhance the reliability, security, and maintenance of RazorSync’s cloud infrastructure. This role requires hands-on management of production systems, troubleshooting across application and infrastructure layers, and balancing urgent operational demands with long-term platform stability. The position focuses primarily on AWS and Windows-based environments, including Amazon RDS, EC2 Windows instances, SQL Server, Windows Server, domain controller services, messaging components, DNS, load balancing, AMIs, and server patching. The successful candidate will work closely with engineering, QA, security, and business teams to ensure dependable releases, improve automation, and strengthen operational readiness.
Key Responsibilities
- Manage daily operations, support, and continuous improvement of AWS infrastructure across production and non-production environments.
- Administer and troubleshoot Amazon RDS environments supporting Microsoft SQL Server workloads, ensuring availability, connectivity, backups, maintenance, and performance.
- Provision, configure, maintain, and troubleshoot Amazon EC2 Windows instances and Windows Server operating systems.
- Manage Windows domain controller services, including Active Directory, Group Policy, DNS, access control, and overall domain health.
- Maintain Windows SQL Server environments in collaboration with engineering and database teams.
- Oversee AWS AMIs and image lifecycle processes to ensure consistent and secure server provisioning and recovery.
- Plan, test, execute, and document server patching for Windows servers and related platform components.
- Configure and support Amazon SQS and SNS integrations for application messaging and notifications.
- Administer Route 53 DNS records, hosted zones, and routing configurations.
- Configure, monitor, and troubleshoot Application Load Balancers, target groups, health checks, certificates, and traffic routing.
- Support secure networking and connectivity across AWS services, including VPCs, subnets, security groups, routing, and private service access.
- Monitor system health, performance, capacity, and availability; proactively identify trends and mitigate risks.
- Participate in incident response, root cause analysis, remediation planning, and validation.
- Develop and maintain operational runbooks, architecture documentation, recovery procedures, and support guides.
- Collaborate with software engineers and QA to improve deployment reliability, environment consistency, and release readiness.
- Identify opportunities to automate repetitive operational tasks and reduce configuration drift.
- Support backup, restore, disaster recovery, vulnerability remediation, and infrastructure security initiatives.
- Assist with cloud cost visibility and responsible resource utilization.
Required Qualifications
- Minimum 5 years of hands-on experience in DevOps, cloud infrastructure, systems engineering, or site reliability engineering.
- At least 4 years administering and supporting production AWS environments.
- Deep practical knowledge of Amazon RDS, EC2, SQS, SNS, Route 53, Application Load Balancers, VPC networking, security groups, and AMI lifecycle management.
- Over 5 years of Windows Server administration experience, including patching, troubleshooting, services, permissions, performance tuning, event logs, and system recovery.
- Strong expertise in Windows Active Directory, Group Policy, domain controllers, DNS, and identity management in enterprise settings.
- Minimum 3 years supporting Microsoft SQL Server infrastructure, including backups, maintenance, performance troubleshooting, and high availability.
- Experience supporting high-availability SaaS platforms with responsibility for uptime, scalability, security, and operational excellence.
- Advanced troubleshooting skills across infrastructure, OS, databases, networking, application layers, and cloud services.
- Proven experience with monitoring, alerting, incident response, change management, root cause analysis, and implementing long-term corrective actions.
- Proficiency in designing and implementing automation solutions using scripting, infrastructure-as-code, or configuration management tools.
- Experience leading production incidents, coordinating cross-functional responses, and driving post-incident improvements.
- Knowledge of cloud security, disaster recovery, backup, and business continuity strategies.
- Ability to create clear technical documentation, operational runbooks, architectural diagrams, and implementation plans.
- Strong communication skills and ability to collaborate effectively across teams.
- Demonstrated independence, prioritization skills, sound technical judgment, and ownership through resolution.
- Experience mentoring engineers and driving continuous improvement is highly preferred.
Preferred Qualifications and Benefits
Preferred experience includes managing PCI DSS-compliant environments, AWS Well-Architected Framework reviews, infrastructure monitoring and observability, acquisition-related migrations or enterprise integrations, Azure DevOps Pipeline administration, and AWS Solutions Architect certification.
Contour offers a competitive salary and comprehensive benefits including medical coverage for employees and dependents, provident fund, performance-based bonuses, home internet subsidy, conveyance allowance, profit-sharing for tenured employees, life benefits, childcare facilities, company-provided meals, professional development budgets, recreational areas, occasional on-shore training, a friendly work environment, and leave encashment.
Contour Software is committed to fostering an inclusive workplace that respects diverse perspectives and experiences. Reasonable accommodations are provided to qualified individuals with special needs, and all applicants are encouraged to apply.