About this role
<h3>Role Summary</h3><p>Ensure 24/7 monitoring, safe deployments, backup management, and incident response while improving performance and reliability to meet monthly availability targets.</p><h3>Key Responsibilities</h3><h3>Cloud Operations (GCC AWS)</h3><ul><li>Manage AWS infrastructure operations, configuration, and service reliability.</li></ul><h3>Monitoring, Alerting & Performance</h3><ul><li>Implement and operate 24/7 monitoring/alerting and performance monitoring; improve observability.</li></ul><h3>CI/CD & Deployment</h3><ul><li>Own pipeline management and deployment processes; enable safe releases and rollback readiness.</li></ul><h3>Backup, DR & Ops Readiness</h3><ul><li>Own backup management, operational documentation updates, and readiness practices.</li></ul><h3>Incident Response</h3><ul><li>Participate in incident response and ensure SLA response times and status updates are met.</li></ul><h3>Security Support</h3><ul><li>Support security patching and vulnerability remediation in collaboration with engineers; help close VAPT/audit findings.</li></ul><h3>Required Experience / Skills</h3><ul>
<li>7+ years DevOps/SRE experience in AWS environments (GCC exposure is a plus).</li>
<li>Strong CI/CD, infra operations, monitoring/alerting, incident response.</li>
<li>Comfortable with production governance and audit/compliance processes.</li>
<li>Able to work with offshore delivery model where applicable.</li>
</ul><p>Originally posted on <a href="https://himalayas.app">Himalayas</a></p>
devops-engineersite-reliability-engineeringcloud-operationsaws-engineeringinfrastructure-engineeringgcp-devops-engineerdevops-software-engineercloud-devops-engineerdevops-release-engineerprodops-engineerazure-devops-engineer