Download of Record of Default (RoD) is disabled for creditors - To request for RoD, please call RoD Hotline No 88673 92123 or Email at rod@nesl.co.in

Senior Infrastructure Monitoring Engineer

Take your next career step at NeSL with an amazing team that is energizing the transformation of Financial Institutions.

Background

National E-Governance Services Ltd. (NeSL) is an Information Utility (IU), registered with Insolvency & Bankruptcy Board of India (IBBI) under IBC, 2016, for its Operations Department is looking for candidates to work in a completely computerised environment, with knowledge and experience in IBC, 2016 & allied Regulations and Banking domain, especially loans/advances & in Information Technology.

Job Summary :

We are seeking an experienced Senior Infrastructure Monitoring Engineer to monitor, maintain, and optimize enterprise server infrastructure across mission-critical production environments. The candidate should have hands-on experience in server monitoring, infrastructure health management, incident response, performance analysis, capacity planning, and ensuring high availability of Windows, Linux, and RHEL server environments, virtualization platforms, and associated infrastructure.

The role requires a proactive professional with strong technical ownership, excellent communication and written documentation skills, and the ability to learn quickly, implement effectively, and support continuous improvement in infrastructure monitoring and operations.

Key Responsibilities :

  • Infrastructure Monitoring and Availability
  • Monitor enterprise server infrastructure to ensure 24×7 availability, operational health, and service reliability.
  • Continuously monitor Windows, Linux, and RHEL servers using enterprise monitoring tools such as Zabbix, Grafana, Prometheus, SolarWinds, PRTG, or ManageEngine.
  • Analyze server health, resource utilization, application availability, service performance, and infrastructure trends.
  • Monitor CPU, memory, disk, network, storage, backup jobs, and virtual machine performance to prevent service degradation.
  • Incident Response and Operational Support
  • Investigate monitoring alerts, identify probable root causes, and coordinate timely resolution of infrastructure incidents.
  • Coordinate with infrastructure, network, storage, database, security, and application teams during incident resolution and service restoration.
  • Support operating system patch verification, firmware monitoring, server health validation, and post-change checks.
  • Performance, Capacity and Continuous Improvement
  • Coordinate with MSP and perform proactive health checks and recommend preventive maintenance and performance optimization actions.
  • Monitor virtualization environments including VMware vSphere, Microsoft Hyper-V, and Nutanix AHV clusters.
  • Monitor enterprise storage infrastructure including SAN, NAS, RAID, and Fibre Channel environments.
  • Perform infrastructure capacity analysis and provide recommendations for scalability, performance improvement, and service reliability.
  • Reporting, Documentation and Governance
  • Create and maintain dashboards, alerts, reports, and operational views for infrastructure availability and performance tracking.
  • Maintain monitoring documentation, operational runbooks, SOPs, RCA inputs, knowledge articles, and audit-supporting records.
  • Coordinate with OEM vendors and service providers for issue resolution, support cases, and hardware replacement where required.

Required Technical Skills :

  • Strong understanding of enterprise server infrastructure, production operations, and high-availability environments.
  • Hands-on experience with server monitoring tools including Zabbix, Grafana, Prometheus, SolarWinds, PRTG, and ManageEngine or equivalent platforms.
  • Experience monitoring Windows Server, Linux, and Red Hat Enterprise Linux (RHEL) environments, including system health, services, logs, resource utilization, and patch/firmware validation.
  • Good understanding of RHEL administration fundamentals, including common system commands, service management, log review, filesystem usage, user/service checks, and basic troubleshooting.
  • Working knowledge of virtualization platforms including VMware vSphere, Microsoft Hyper-V, and Nutanix AHV.
  • Knowledge of enterprise storage technologies including SAN, NAS, RAID, and Fibre Channel.
  • Basic networking knowledge covering TCP/IP, VLANs, switches, firewalls, and load balancers.
  • Experience monitoring enterprise applications, services, infrastructure availability, and performance trends.
  • Understanding of backup and disaster recovery processes.
  • Familiarity with log analysis, alert correlation, performance troubleshooting, and root cause analysis.

Communication and Professional Skills:

  • Good verbal communication skills with the ability to coordinate clearly with technical teams, vendors, service providers, and business stakeholders during operational situations.
  • Good written communication skills for preparing incident updates, operational reports, SOPs, RCA inputs, and management summaries.
  • Ability to communicate technical issues in a clear, structured, and professional manner.
  • Fast learner with the ability to quickly understand new tools, technologies, operational processes, and infrastructure environments.
  • Strong implementor with the ability to convert recommendations, monitoring requirements, and process improvements into practical operational actions.

Preferred Qualifications :

  • Bachelor’s degree in Computer Science, Information Technology, Electronics, or a related field.
  • VMware VCP certification.
  • Microsoft Certified: Windows Server / Azure Administrator or equivalent certification.
  • Red Hat Certified System Administrator (RHCSA) or equivalent RHEL/Linux certification will be preferred.
  • Zabbix, Grafana, or equivalent monitoring platform certification will be preferred.
  • Knowledge of Azure, AWS, or Google Cloud monitoring services will be an added advantage.

Experience Requirements :

  • 5-10 years of experience in enterprise infrastructure monitoring, server operations, or production support.
  • Experience supporting mission-critical production environments with defined incident, change, and service management controls.
  • Experience with enterprise monitoring platforms, proactive infrastructure management, and operational reporting.
  • Hands-on exposure to Windows, Linux, and RHEL server monitoring and operational troubleshooting is preferred.
  • Exposure to cloud infrastructure monitoring is an added advantage.

Key Competencies:

  • Infrastructure Monitoring
    Strong ability to monitor availability, health, alerts, and performance across infrastructure platforms.
  • RHEL / Linux Operational Skills
    Hands-on familiarity with RHEL/Linux monitoring, basic administration checks, logs, services, and troubleshooting.
  • Incident & Problem Management
    Ability to investigate alerts, support RCA, coordinate resolution, and follow ITIL processes.
  • Documentation & Reporting
    Strong written communication for SOPs, runbooks, dashboards, reports, incident updates, and handovers.
  • Communication & Coordination
    Clear verbal communication with internal teams, vendors, and stakeholders during incidents and reviews.
  • Fast Learner & Implementor
    Ability to learn new tools and quickly implement monitoring, process, and operational improvements.
  • Capacity & Performance Analysis
    Ability to review trends and recommend preventive or optimization actions.
  • Cross-functional Collaboration
    Ability to work effectively with infrastructure, network, storage, database, and application teams.

Working Conditions :

  • Willingness to work extended hours as per operational needs.
  • Participation in critical production incident response activities.
  • Ability to work in a fast-paced production operations environment while ensuring high service availability and operational discipline.

 Work Location: Bengaluru

What we offer

An exciting work experience in a class apart company and industry leader whose mission is to provide a fully digital and superior experience in documentation, authentication, storage and access of financial and operational debt obligations in the country, held as evidence, through its entire journey starting from contract execution and to be the most complete and current information repository in the country for loans and debt obligations.

More About us

NeSL is India’s first Information Utility and is registered with the Insolvency and Bankruptcy Board of India (IBBI) under the aegis of the Insolvency and Bankruptcy Code, 2016 (IBC). The company has been set up by leading banks and public institutions. The primary role of NeSL is to serve as a repository of legal evidence holding the information pertaining to any debt/claim, as submitted by the financial or operational creditor and verified and authenticated by the parties to the debt.

We Value Diversity

At NeSL, we have the clear goal of driving diversity and inclusion across all dimensions: gender, abilities, ethnicity and generations. We welcome applications for employment from all qualified candidates, regardless of race, colour, gender, ethnic origin, age, disability, religion, gender identity or any other status projected by applicable law. We comply with all applicable laws in every jurisdiction in where we operate.

Last date of application: 09th August,2026

To Apply
Email your resume to hr@nesl.co.in with Subject line – “Senior Infrastructure Monitoring Engineer”

Apply for this position

Allowed Type(s): .pdf