System Manager
Job Summary
We are looking for an experienced System Manager to manage 24x7 mission-critical infrastructure, leading system/application engineers and ensuring availability, cybersecurity, incident recovery, and operational continuity.
Mandatory Skill-set
- Degree/Diploma in Computer Science, IT, Computer Engineering, or related field;
- Must have 6 years of experience in enterprise infrastructure, systems administration, or mission-critical environments, include 3 years of leading a team;
- Must have 3 years of experience managing Operations & Maintenance (O&M) activities for critical systems;
- Strong hands-on experience with Windows Server 2019+/Windows 11, Active Directory, GPO, Microsoft CA, PKI, and SSL/TLS certificate management;
- Must have VMware experience, including ESXi, vCenter, vSAN, and vSAN Stretch Cluster;
- Hands-on experience in patch management, including WSUS Offline and manual WSUS catalogue imports;
- Hands-on experience with Microsoft SQL Server, endpoint security (Trellix/SEP), cybersecurity technologies, and enterprise backup/restore solutions;
- Experience troubleshooting 24x7 mission-critical, highly available, on-premise environments, including critical incidents and service recovery, is essential;
- Willingness to support after-hours deployments, maintenance windows, and on-call duties.
Desired Skill-set
- Good to have experience in Splunk Enterprise / Syslog / Kiwi Syslog VMware NSX.
- Aviation systems experience in airport operations, air traffic management, surveillance, or aviation support systems.
Responsibilities
- Lead day-to-day operations of a 24x7 mission-critical system with a team of system engineers;
- Manage system availability, performance, security, and operational readiness;
- Work closely with the Service Delivery Manager (SDM) to plan and execute maintenance, patching, upgrades, and service improvement initiatives;
- Lead incident response, root cause analysis, service recovery, and operational reporting activities with the SDM;
- Review and approve deployment plans, impact assessments, risk evaluations, and rollback strategies;
- Manage infrastructure lifecycle activities, including server, workstation, virtualization, network, and COTS software upgrades;
- Ensure compliance with cybersecurity, backup, disaster recovery, and configuration management requirements;
- Identify operational risks, technology gaps, and potential single points of failure, and implement effective mitigation measures;
- Provide technical leadership, mentoring, and guidance to the operations team;
- Engage stakeholders, customers, and vendors on operational matters and service delivery performance.
Should you be interested in this career opportunity, please send in your updated resume to [email protected] at the earliest.
When you apply, you voluntarily consent to the disclosure, collection and use of your personal data for employment/recruitment and related purposes in accordance with the SCIENTE Group Privacy Policy, a copy of which is published at SCIENTE’s
website(https://www.sciente.com/privacy-policy).
Confidentiality is assured, and only shortlisted candidates will be notified for interviews.
EA Licence No. 07C5639