Job description
Job Title: Team Lead – NOC Location: Bangalore Experience: 10+ Years Industry Experience | 6+ Years Relevant Experience Employment Type: Full Time Job Summary The Team Lead – NOC will be responsible for managing and overseeing 24x7 Network Operations Center (NOC) activities to ensure the stability, availability, and performance of critical infrastructure and network services.
The role involves leading a team responsible for network monitoring, incident response, service availability, and operational support across cloud, telecom, and data center environments. The ideal candidate will have strong expertise in network operations, incident management, infrastructure monitoring, and team leadership.
Key Responsibilities
NOC Operations & Monitoring Lead daily operations of the Network Operations Center (NOC) ensuring 24x7 monitoring of infrastructure and services. Monitor network, servers, cloud platforms, and application environments to ensure high availability. Ensure timely detection, escalation, and resolution of infrastructure and network incidents.
Incident & Escalation Management Act as the first level escalation point for major incidents and service outages. Coordinate with engineering, infrastructure, and SRE teams to resolve critical issues. Ensure incidents are handled according to defined SLA and incident management processes. Network & Infrastructure Monitoring Oversee monitoring of enterprise networks, ISP infrastructure, and cloud environments.
Ensure proper functioning of monitoring and alerting systems. Proactively identify risks impacting system performance and availability. Team Leadership Lead, mentor, and manage NOC engineers and support teams. Ensure team adherence to operational procedures, escalation processes, and SLAs. Conduct regular team reviews, shift planning, and performance monitoring.
Service Reliability & Operations Ensure compliance with service availability targets and operational KPIs. Work with infrastructure and engineering teams to improve service reliability and operational efficiency. Support operational readiness for new infrastructure deployments and upgrades. Drive continuous improvement initiatives to improve operational efficiency.
Reporting & Documentation Maintain incident reports, operational dashboards, and service reports. Provide regular updates on service performance, outages, and operational metrics to management. Participate in root cause analysis (RCA) and service improvement initiatives.
Required Skills
& Technical Expertise Strong understanding of network operations and NOC monitoring environments. Knowledge of networking technologies including BGP, MPLS, VXLAN, EVPN, and SD-WAN. Experience with network monitoring tools such as SolarWinds, Nagios, Zabbix, PRTG, or similar platforms. Familiarity with cloud infrastructure environments including AWS, Azure, or Google Cloud.
Understanding of data center infrastructure including servers, storage, virtualization, and networking. Experience working with ISP or telecom network infrastructure environments. Strong knowledge of incident management and escalation processes in a 24x7 operations environment. Excellent leadership, communication, analytical, and problem-solving skills.
Required Qualifications 10+ years of industry experience in network operations, infrastructure operations, or telecom environments. 6+ years of relevant experience in NOC operations or infrastructure monitoring roles. Experience managing 24x7 NOC teams and operational processes. Strong understanding of incident management, service monitoring, and operational support.
Experience managing cloud-based infrastructure in AWS, Azure, or Google Cloud environments.
Preferred Qualifications
Experience working in ISP, telecom, or large-scale infrastructure environments. Certifications such as CCNA, CCNP, ITIL, AWS, Azure, or Google Cloud certifications. Exposure to automation tools, observability platforms, and DevOps practices. About Zybisys Zybisys is a fast-growing technology company delivering enterprise infrastructure, data center, and cybersecurity solutions to clients.
With a strong focus on reliability, security, and operational excellence, we support mission-critical environments for enterprise customers. Our teams combine technical expertise with strong service delivery to ensure high availability and performance across complex IT environments.