Loading open roles
Loading open roles
Loading role

UpMan Placements Private Limited · posted 5 months ago
Role of Cloud Operations / Sustenance Engineer
Position Title | Cloud Operations / Sustenance |
Role | The Cloud Operations / Sustenance role is responsible for the day‑to‑day operations, stability, security posture, and continuous availability of enterprise cloud platforms and workloads. The role ensures that cloud environments remain reliable, performant, secure, compliant, and cost‑efficient throughout their operational lifecycle, post implementation. This role works closely with Cloud Engineering, Application Teams, FinOps, Information Security, and IT Infrastructure teams to ensure smooth production operations and adherence to SLAs, governance standards, and regulatory requirements. |
Reporting to | R1- VP - IT, Infra Core R2- EVP- IT, Infra Core, Cloud & CITSO |
Key Responsibilities | 1. Cloud Platform Operations
2. Monitoring, Incident & Problem Management · Implement and manage cloud monitoring, logging, and alerting solutions · Perform 24x7 monitoring of cloud resources, applications, and services · Respond to incidents, outages, and performance degradation · Conduct root cause analysis (RCA) and problem management · Track and improve SLAs, SLOs, MTTR, and incident trends · Coordinate incident resolution with application, network, and security teams
3. Change, Patch & Release Management · Execute approved changes in line with ITIL / ITSM processes · Apply OS, platform, middleware, and cloud service patches · Ensure configuration consistency and prevent configuration drift · Support release activities and maintenance windows · Maintain rollback and recovery readiness for changes
4. Security Operations & Compliance · Enforce cloud security controls defined by Information Security teams · Manage IAM operations including access changes, role reviews, and credential hygiene · Monitor security alerts, logs, and compliance dashboards · Support vulnerability remediation and security incident response · Ensure ongoing compliance with internal policies and regulatory standards · Support audits by providing operational evidence and reports
5. Backup, DR & Business Continuity · Ensure backups are configured, running, and tested regularly · Monitor backup success and conduct restore validations · Support disaster recovery planning and execution · Participate in DR drills and audit walkthroughs · Ensure RPO and RTO targets are consistently met
· Work with FinOps teams to identify idle, under‑utilized, and orphan resources · Implement cost optimization actions such as scheduling, shutdowns, and cleanup · Enforce tagging standards and cost allocation practices · Monitor unexpected spikes or anomalies in cloud usage · Ensure operational decisions align with cost efficiency goals
7. Automation & Continuous Improvement · Automate routine operational tasks wherever possible · Improve operational efficiency through scripts, runbooks, and tooling · Enhance platform stability and reliability through continuous improvements · Maintain operational documentation, SOPs, and runbooks · Support standardization across environments and platforms
8. Vendor & Platform Support
|
· Participate in operational governance meetings and reviews · Provide operational inputs to cloud roadmap and capacity planning · Mentor junior operations engineers (if applicable) · Ensure operational readiness for new application go‑lives · Maintain clear handover between engineering and operations teams | |
Key Performance Indicators (KPIs)
| · Cloud platform availability and uptime · Incident response and resolution times (MTTR) · Number of repeat incidents and problem closure rate · Patch and backup compliance percentage · Security and audit findings related to cloud operations · Operational cost optimization improvements · Change success and rollback rates |
Qualifications & Skills | · Hands‑on experience with AWS / GCP · Strong understanding of cloud compute, storage, networking, and IAM · Experience with monitoring and observability tools · Knowledge of backup, DR, and high‑availability architectures · Basic scripting skills (PowerShell, Bash, Python – preferred) · Familiarity with CI/CD and Infrastructure as Code concepts
|
Key Interactions | · IT Development Team · IT Infra Teams · Head – IT Infrastructure · Application, Infrastructure & Vendors Team Leaders within IT · Risk Department · Information Security Team |
Experience | · 6–10 years of overall IT experience · Minimum 3–5 years of hands-on experience in cloud engineering and implementation · Experience in enterprise-scale or regulated environments preferred |
Education Qualifications | · Bachelor’s degree in Computer Science, Engineering or equivalent · Cloud certifications (GCP/AWS) – preferred |
Location | Mumbai HO-CPC(Seawoods, Navi Mumbai) |