About this Operations Tech Lead role at Megaport
About Latitude.sh
Latitude.sh global computing platform was launched in 2019, enabling businesses to programmatically deploy single-tenant Bare Metal instances in different parts of the world. We are a team of passionate individuals about hardware, software, and network infrastructure looking to build the fastest, easiest-to-use, developer-centric single-tenant Cloud infrastructure. If you share this passion, join our growing team of talented people and help build the future of the Internet.
Why Latitude.sh?
We're a lean, agile team of passionate professionals who believe in the power of innovation and creative problem-solving. As part of our team, you won't be lost in the crowd – you'll be an essential contributor, making a real impact from day one. Our values at Latitude.sh guide us in all our work and partnerships. We're proud to be an inclusive company, and we welcome all applicants for our open positions, regardless of their background, religion, sexual orientation, gender identity, age, nationality, or disability. If these values speak to you, we'd love for you to become a part of our team.
The Role
The Operations Tech Lead at Latitude.sh is responsible for leading a team of Senior Operations Analysts, ensuring the reliable and efficient functioning of our global data center infrastructure. This role combines deep hands-on technical expertise with a people-first leadership approach.
As the primary escalation point for the analysts team, the Tech Lead ensures operational continuity, drives process improvements, and maintains high standards of execution across all DCOps workflows. This position reports directly to the Operations Manager.
What You'll Be Doing
Team Leadership & Shift Management
Lead and supervise the Senior Operations Analysts team, ensuring full operational coverage and adherence to standards;
Serve as the primary escalation point for complex operational issues, coordinating resolution across internal teams and remote data center techs;
Conduct shift handover briefings, communicating ongoing incidents, pending tasks, and critical updates;
Monitor team performance, provide real-time coaching, and identify development opportunities for analysts;
Foster a culture of ownership, accountability, and continuous improvement within the team.
Infrastructure Operations & Incident Management
Oversee and participate in the execution of operational tasks involving servers, storage, networking equipment, and platform software;
Lead root cause analysis for complex incidents related to hardware failures, RAID rebuilds, storage allocation, and disk lifecycle management;
Ensure rack and stack activities, hardware installation, and in-cabinet maintenance are executed to the highest standards;
Validate and improve out-of-band management configurations (IPMI/BMC) across the infrastructure fleet;
Coordinate with remote data center techs to ensure operational standards and SLAs are consistently met;
Install, configure, and troubleshoot Linux and Windows operating systems, supporting analysts with complex cases;
Documentation & Process Improvement
Maintain and improve internal operational documentation, runbooks, workflows, and escalation procedures;
Identify process gaps and propose improvements to the Operations Manager;
Contribute to the development and review of standard operating procedures (SOPs) for the DCOps team.
What We're Looking For
Advanced English communication skills (B2 or higher);
Proven experience leading or mentoring technical teams in a data center or infrastructure operations environment;
Deep expertise with server hardware, including diagnostics, replacements, upgrades, and fleet management;
Strong understanding of data center standards, rack layouts, cabling, power distribution (PDU), and remote hands coordination;
Advanced knowledge of Linux system administration (CentOS, Ubuntu, or similar), including scripting, automation, and performance tuning;
Solid knowledge of Windows Server environments;
Strong familiarity with RAID technologies, disk provisioning, and storage troubleshooting;
Strong incident command skills with demonstrated ability to lead resolution during critical events;
Proficiency with monitoring platforms, ticketing systems, and incident handling workflows;
Bachelor's degree in Computer Science, Information Systems, or equivalent working experience;
Prior experience in a team lead or senior role within a data center operations environment is strongly preferred.
Nice to Have:
Experience with automation and scripting tools (Python, Bash, Ansible) applied to infrastructure operations;
Familiarity with DCIM or asset management platforms;
Experience coordinating with multiple data centers or PoPs simultaneously;
Background in ISP or cloud infrastructure environments.