Manager - Core Infrastructure Operations
Sonata Software
Job Description
Job Description
Job Title: Manager โ Core Infrastructure Operations
\nLocation: (Bangalore / Onsite(5-Days WFO)
\nExperience: 10+ Years
\nEmployment Type: Full-Time
\nNP_ Immediate joiners
\nAbout the Role
\nWe are looking for an experienced Manager โ Core Infrastructure Operations to lead and manage the day-to-day operations of enterprise infrastructure services. This is a hands-on leadership role responsible for ensuring the reliability, security, availability, and continuous improvement of core infrastructure environments across data centers, virtualization, compute, storage, backup & recovery, network, and server platforms.
\nThe ideal candidate should have strong operational leadership experience, excellent incident management skills, and the ability to drive infrastructure stability while leading a high-performing technical team.
\nKey Responsibilities
\nInfrastructure Operations & Service Management
\n- \n
- Lead daily operations of enterprise infrastructure services across: \n
- Data Centers & Server Rooms \n
- Compute & Virtualization Platforms (VMware/Hyper-V) \n
- Storage Infrastructure \n
- Backup & Recovery Systems \n
- Network Infrastructure \n
- Windows/Linux Server Operating Systems \n
- Manage incidents, service requests, changes, maintenance activities, and lifecycle upgrades. \n
- Ensure high availability, performance, security, and operational stability of infrastructure services. \n
- Define priorities across operational support, risk reduction initiatives, and project activities. \n
- Maintain operational documentation, runbooks, and standard operating procedures. \n
- Manage 24x7 on-call support coverage and escalation processes. \n
Incident & Problem Management
\n- \n
- Act as the primary escalation point for critical infrastructure incidents. \n
- Lead major incident response and coordinate cross-functional technical teams. \n
- Drive Root Cause Analysis (RCA) and implement corrective/preventive actions. \n
- Reduce recurring incidents through automation, monitoring enhancements, and process improvements. \n
- Partner with Service Management teams for incident communication and stakeholder updates. \n
Change & Production Readiness
\n- \n
- Ensure infrastructure changes are properly planned, tested, approved, implemented, and documented. \n
- Support high-risk and emergency changes with rollback and recovery planning. \n
- Validate production readiness for new infrastructure services and changes. \n
- Ensure compliance with change management, audit, and operational standards. \n
Team Leadership
\n- \n
- Lead, mentor, and develop infrastructure operations engineers. \n
- Drive accountability, performance management, and technical capability development. \n
- Establish clear ownership, escalation paths, and operational governance. \n
- Foster a culture of continuous improvement and operational excellence. \n
Required Skills & Experience
\nMust Have
\n- \n
- 12+ years of Infrastructure Operations experience with at least 5+ years in Team Management. \n
- Strong experience managing enterprise infrastructure environments. \n
- Hands-on knowledge of: \n
- Data Center Operations \n
- Compute & Virtualization (VMware, Hyper-V) \n
- Storage Technologies (SAN/NAS) \n
- Backup & Recovery Solutions \n
- Windows & Linux Server Administration \n
- Network Infrastructure Fundamentals \n
- Experience managing: \n
- Incident Management \n
- Problem Management \n
- Change Management \n
- Production Support Operations \n
- Strong understanding of: \n
- Disaster Recovery (DR) \n
- Business Continuity \n
- Backup & Restore Processes \n
- RTO/RPO Concepts \n
- Operational Risk & Change Control \n
- Experience leading teams during major service-impacting incidents. \n
- Excellent stakeholder management and communication skills. \n
Good to Have
\n- \n
- Exposure to Hybrid Cloud environments (Azure/AWS). \n
- Infrastructure Automation experience (PowerShell, Ansible, Terraform, etc.). \n
- ITIL Foundation or equivalent certification. \n
- Experience with monitoring and observability platforms. \n