Key Responsibilities:
Cloud Operations & Reliability
Drive 24×7 cloud operations ensuring high availability of critical applications
Own SLA adherence, monitoring, incident management, and performance optimization
Lead capacity planning for growing workloads and business demand
Infrastructure & Cost Optimization
Monitor and optimize cloud utilization and spend
Drive efficiency improvements and cost control initiatives
Manage infra performance trends, forecasting, and budgeting
Cloud Governance & Security
Ensure compliance with cloud governance, ISMS, and security standards
Oversee patching, hardening, and vulnerability management
Drive risk assessment across public, private, and hybrid cloud setups
Cloud Provisioning & Architecture
Evaluate and implement new cloud solutions and services
Define SOWs, SLAs, and vendor alignment
Support architecture decisions for integrating enterprise applications
Stakeholder & Service Management
Partner with business stakeholders for service delivery and improvements
Conduct regular service reviews and performance tracking
Drive continuous improvement and customer satisfaction
What We’re Looking For:
Must-Have Skills
Strong experience in multi-cloud environments (AWS / Azure / GCP / OCI)
Deep understanding of cloud operations, monitoring, and incident management
Experience managing large-scale infrastructure (1000+ servers)
Knowledge of cloud security, governance, and compliance frameworks
Strong stakeholder management and communication skills
Good to Have
Exposure to cloud architecture & solution design
Experience in cost optimization / FinOps
Familiarity with DR setups, high availability, and zero data loss systems
Qualifications
Bachelor’s in Computer Science / Engineering (or related field)
7+ years experience (5–6 yrs in Cloud Ops + ops/project management exposure)
Certifications: ITIL / AWS / Azure (preferred)
Who should Apply?
You enjoy running large-scale, business-critical cloud environments
You think in terms of uptime, SLAs, performance, and cost trade-offs
You’ve handled multi-cloud ecosystems and complex infra setups
You take ownership of operations, reliability, and stakeholder outcomes
You thrive in high-scale, high-impact enterprise environments