This is an experimental service. Give feedback (opens in new tab)

English |

Network Operations (NOC) Engineer

Company:Media Stream AI Limited
Salary:£41,000 - £47,000
Hours:Full-time
Location:Dundee, DD2 1UR
Working pattern:On-site
Job type:Permanent
Posting date:19 Aug 2026
Closing date:18 Sept 2026
Apply for this job

Summary

The Network Operations (NOC) Engineer will support the 24/7 operation and monitoring of the MSAI Scotland data centre network and GPU infrastructure, ensuring network availability, performance and reliability across the campus.

The role will monitor the compute fabric, InfiniBand/Ethernet networks, transit and management networks, provide first-line response to network and infrastructure alerts, and support customer connectivity into GPU clusters. The engineer will work as part of a 24/7 NOC function, following established runbooks and escalating incidents to the appropriate engineering teams when required.

Duties

- Monitor the 24/7 network operations environment across the data centre and GPU clusters.

- Monitor the compute fabric, management network, transit network and client connectivity.

- Provide first-line response to network, compute and infrastructure alerts.

- Identify, investigate and resolve network incidents within agreed procedures and escalation paths.

- Monitor network performance, availability, latency, packet loss and utilisation.

- Support the operation of high-performance GPU cluster networking, including Ethernet and InfiniBand environments.

- Monitor and troubleshoot switches, routers, firewalls, network interfaces and connectivity services.

- Support client cluster connectivity, including provisioning, troubleshooting and fault resolution.

- Assist with the deployment and configuration of network equipment and services.

- Perform basic network diagnostics using appropriate tools and command-line utilities.

- Support network maintenance, upgrades and planned changes in accordance with approved change-control procedures.

- Maintain and follow NOC runbooks, Standard Operating Procedures (SOPs) and escalation procedures.

- Ensure incidents and service requests are accurately logged, updated and closed within agreed SLAs.

- Escalate complex network, hardware or infrastructure incidents to senior network, systems or infrastructure engineers.

- Work closely with the Data Centre, GPU/Compute, Systems and Critical Facilities teams during incidents and planned works.

- Monitor customer environments and support the technical operations required to maintain service availability.

- Maintain accurate network documentation, including network diagrams, asset records, configurations and troubleshooting guides.

- Identify recurring incidents and contribute to root-cause analysis and preventative actions.

- Support capacity and performance monitoring across the network infrastructure.

- Assist with network security monitoring and escalate suspicious or abnormal activity in accordance with site procedures.

- Participate in incident reviews and contribute to continuous improvement of NOC processes.

- Maintain and improve runbooks, knowledge articles and operational procedures.

- Participate in a 24/7 shift rota covering nights, weekends and public holidays.

Apply for this job

Related jobs

GPU / Compute Systems Engineer

£54,000 to £60,000 per year

Media Stream AI Limited

Dundee

On-sitePermanentFull time

Data Centre Technician

£33,000 to £40,000 per year

Media Stream AI Limited

Dundee

On-sitePermanentFull time

Security Officer (24/7 Gatehouse / Patrol)

£25,500 to £29,500 per year

Media Stream AI Limited

Dundee

On-sitePermanentFull time
Browse more jobs