Network Operations (NOC) Engineer
| Company: | Media Stream AI Limited |
|---|---|
| Salary: | £41,000 - £47,000 |
| Hours: | Full-time |
| Location: | Dundee, DD2 1UR |
| Working pattern: | On-site |
| Job type: | Permanent |
| Posting date: | 19 Aug 2026 |
| Closing date: | 18 Sept 2026 |
Summary
The Network Operations (NOC) Engineer will support the 24/7 operation and monitoring of the MSAI Scotland data centre network and GPU infrastructure, ensuring network availability, performance and reliability across the campus.
The role will monitor the compute fabric, InfiniBand/Ethernet networks, transit and management networks, provide first-line response to network and infrastructure alerts, and support customer connectivity into GPU clusters. The engineer will work as part of a 24/7 NOC function, following established runbooks and escalating incidents to the appropriate engineering teams when required.
Duties
- Monitor the 24/7 network operations environment across the data centre and GPU clusters.
- Monitor the compute fabric, management network, transit network and client connectivity.
- Provide first-line response to network, compute and infrastructure alerts.
- Identify, investigate and resolve network incidents within agreed procedures and escalation paths.
- Monitor network performance, availability, latency, packet loss and utilisation.
- Support the operation of high-performance GPU cluster networking, including Ethernet and InfiniBand environments.
- Monitor and troubleshoot switches, routers, firewalls, network interfaces and connectivity services.
- Support client cluster connectivity, including provisioning, troubleshooting and fault resolution.
- Assist with the deployment and configuration of network equipment and services.
- Perform basic network diagnostics using appropriate tools and command-line utilities.
- Support network maintenance, upgrades and planned changes in accordance with approved change-control procedures.
- Maintain and follow NOC runbooks, Standard Operating Procedures (SOPs) and escalation procedures.
- Ensure incidents and service requests are accurately logged, updated and closed within agreed SLAs.
- Escalate complex network, hardware or infrastructure incidents to senior network, systems or infrastructure engineers.
- Work closely with the Data Centre, GPU/Compute, Systems and Critical Facilities teams during incidents and planned works.
- Monitor customer environments and support the technical operations required to maintain service availability.
- Maintain accurate network documentation, including network diagrams, asset records, configurations and troubleshooting guides.
- Identify recurring incidents and contribute to root-cause analysis and preventative actions.
- Support capacity and performance monitoring across the network infrastructure.
- Assist with network security monitoring and escalate suspicious or abnormal activity in accordance with site procedures.
- Participate in incident reviews and contribute to continuous improvement of NOC processes.
- Maintain and improve runbooks, knowledge articles and operational procedures.
- Participate in a 24/7 shift rota covering nights, weekends and public holidays.
Related jobs
GPU / Compute Systems Engineer
£54,000 to £60,000 per year
Media Stream AI Limited
Dundee
On-sitePermanentFull timeData Centre Technician
£33,000 to £40,000 per year
Media Stream AI Limited
Dundee
On-sitePermanentFull timeSecurity Officer (24/7 Gatehouse / Patrol)
£25,500 to £29,500 per year
Media Stream AI Limited
Dundee
On-sitePermanentFull time