Jobs Companies Equinix Staff Engineer, Kubernetes Platform & Networking

Über diese Staff Engineer, Kubernetes Platform & Networking Stelle bei Equinix

Equinix · Vor Ort · Bangalore Office BLS2

Who are we?

Equinix is the world’s digital infrastructure company®, shortening the path to connectivity to enable the innovations that enrich our work, life and planet. 
 

A place where bold ideas are welcomed, human connection is valued, and everyone has the opportunity to shape their future.

Help us challenge assumptions, uncover bias, and remove barriers—because progress starts with fresh ideas. You’ll find belonging, purpose, and a team that welcomes you—because when you feel valued, you’re empowered to do your best work.

Job Summary

We are looking for a Staff Engineer to join the Reliability Engineering team that operates a bare-metal Kubernetes platform across multiple metro locations. The platform's architecture is defined and the build is underway you will provision and operate it day to day: bringing up nodes and clusters, running upgrades, keeping the fleet consistent through GitOps, and owning your shift of the on-call rotation.

This is a deliberately broad role. We are not looking for someone who works one layer and hands everything below it across a boundary. You will work alongside the teams that own the physical hardware, the underlay network, and the data plane software, and you need working knowledge of all three enough to isolate where a problem lives before escalating it.

You will join the Digital Interconnection Engineering organization. The team owns the full application stack for interconnection and the Kubernetes platform for the software-defined data plane across multiple metros. We operate with an automation-first, GitOps-driven culture and a strong commitment to operational rigor, documented runbooks, and blameless postmortems.

Responsibilities

OS Provisioning & Bare-Metal Infrastructure

  • Provision bare-metal servers across multiple metros using the team's PXE/iPXE imaging and cloud-init pipeline, and harden the OS prior to cluster bootstrap

  • Extend and maintain that provisioning automation as new hardware, metros and OS versions are introduced, keeping node bring-up free of manual per-node steps

  • Work with the teams that own the hardware on BIOS/firmware baselines, NIC configuration and disk layout standards, and use BMC/IPMI and platform sensors to isolate hardware faults before handing them over

Kubernetes Cluster Operations

  • Bring up and operate Kubernetes clusters following the team's established control-plane topology, including etcd members, kube-apiserver HA behind the L4 load balancer or VIP, and defined failure-domain placement

  • Provision worker nodes, configure CNI, and run end-to-end cluster validation before workloads are onboarded

  • Perform routine cluster operations: node drains and replacements, certificate rotation, capacity checks, and remediation of cluster-level alerts

GitOps & Fleet Configuration

  • Make cluster and fleet configuration changes through GitOps, keeping clusters consistent, auditable and drift-free

  • Perform cluster registration, RBAC and policy changes in the management plane, respecting the isolation between production and nonproduction planes

  • Extend the deployment pipeline for the software-defined data plane adding validation coverage and test cases working from the pipeline design and gates owned by the senior engineers on the team

Upgrades & Patching

  • Execute rolling OS and Kubernetes upgrades across the fleet from tested runbooks, with no unplanned downtime, and improve those runbooks as you find gaps

  • Take etcd snapshots ahead of significant changes, and run scheduled restore drills to prove the recovery path works

  • Apply the platform's patching cadence across the fleet, tracking CVEs affecting the OS, Kubernetes and cluster components

Observability & On-Call

  • Own your shift of a follow-the-sun on-call rotation, and act as first responder for control-plane, node, and cluster-level incidents

  • Build and tune dashboards and alerts for the systems you operate, and reduce alert noise where it is costing the rotation time

  • Write up incidents you respond to, contribute to blameless postmortems, and turn recurring toil into automation

  • Author and maintain runbooks for the components you operate

What Success Looks Like First 9-12 Months

  • Provisioning and bringing up nodes and clusters independently, without per-node manual steps

  • Running OS and Kubernetes upgrades on cadence from runbooks, and improving those runbooks where they fall short

  • Owning your on-call shift unassisted, as first responder for cluster and node-level incidents

  • Restore drills run on schedule, with the recovery path demonstrated rather than assumed

  • Measurable reduction in toil or alert noise in the areas you operate

  • Runbooks and dashboards in place for the components you own

Qualifications

Required

  • Linux on bare metal: Solid hands-on Linux/Ubuntu administration and provisioning; familiarity with PXE/iPXE imaging or equivalent, cloudinit, and OS hardening

  • Hardware working knowledge: Comfortable with BMC/IPMI consoles, SMART data and platform sensors; able to identify a failing disk, NIC, DIMM or power supply and follow a vendor replacement process

  • Kubernetes: Production experience operating clusters ideally RKE2 or kubeadm on bare metal including etcd basics, node lifecycle, CNI configuration and cluster-level troubleshooting

  • Infrastructure as code: Day-to-day use of Ansible, Terraform, ArgoCD, Fleet or similar; comfortable making infrastructure changes through a pipeline rather than by hand

  • Networking: Working knowledge of VLANs, BGP, bonded NICs and SR-IOV enough to read a configuration, interpret what you see, and isolate whether a problem is above or below the network layer

  • Upgrades: Experience executing OS and Kubernetes upgrades in production against an established runbook

  • On-call: Experience in an on-call operations role, responding to production incidents and writing them up afterward

  • Automation mindset: A reflex to automate repeated manual work, and to leave the runbook better than you found it

Preferred

  • Experience with Rancher, Rancher Prime or a comparable multi-cluster management plane

  • Familiarity with DPDK-based data planes such as 6WIND VSR, or with SR-IOV NIC tuning

  • Exposure to SDN or private-connectivity platforms (Equinix Fabric, VyOS, or equivalent)

  • Familiarity with LLM provider APIs and AI/agent gateway proxying patterns, since inference traffic from multiple providers transits this platform

  • Experience operating infrastructure across geographically distributed sites

Equinix is committed to ensuring that our employment process is open to all individuals, including those with a disability.  If you are a qualified candidate and need assistance or an accommodation, please let us know by completing this form.

Equinix is an Equal Employment Opportunity and, in the U.S., an Affirmative Action employer.  All qualified applicants will receive consideration for employment without regard to unlawful consideration of race, color, religion, creed, national or ethnic origin, ancestry, place of birth, citizenship, sex, pregnancy / childbirth or related medical conditions, sexual orientation, gender identity or expression, marital or domestic partnership status, age, veteran or military status, physical or mental disability, medical condition, genetic information, political / organizational affiliation, status as a victim or family member of a victim of crime or abuse, or any other status protected by applicable law. 

We use artificial intelligence in our hiring process. Learn more here.

This posting is a new position within our organization.
Bereit, sich bei Equinix zu bewerben?
Bei Equinix bewerben

Über Equinix

Equal Employment Opportunity : Equinix is an Equal Employment Opportunity and, in the U.S., an Affirmative Action employer. All qualified applicants will receive consideration for employment without regard to unlawful consideration of race, color, religion, creed, national or ethnic origin, ancestry, place of birth, citizenship, sex, pregnancy/ childbirth or related medical conditions, sexual orientation, gender identity or expression, marital or domestic partnership status, age, veteran or military status, physical or mental disability, medical condition, genetic information, political/ organizational affiliation, status as a victim or family member of a victim of crime or abuse, or any oth

Alle Jobs bei Equinix ansehen →

Ähnliche Jobs

Equinix
Staff Software Engineer (Python, AI)
Equinix
⚡ Früh bewerben Bangalore Office BLS2 Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 1 Wo.
Equinix
Staff SDET (Playwright, Typescript)
Equinix
⚡ Früh bewerben Bangalore Office BLS2 Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 1 Wo.
Equinix
Sr Staff Agentic AI Engineer
Equinix
⚡ Früh bewerben Bangalore Office BLS2 Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 1 Wo.
Ironclad
Staff IAM Engineer
Ironclad
⚡ Früh bewerben San Francisco Hybrid $160,000–$190,000
● Neu 👁 Gesehen ✓ Beworben vor 4 Std.
Redwood Materials
Staff Mechanical Design Engineer, Big Machines
Redwood Materials
⚡ Früh bewerben McCarran, NV; San Francisco, C... Vor Ort $147,500–$235,000
● Neu 👁 Gesehen ✓ Beworben vor 4 Std.
Redwood Materials
Staff Cathode Engineer
Redwood Materials
⚡ Früh bewerben McCarran, NV Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 5 Std.
Redwood Materials
Staff Chemical Engineer
Redwood Materials
⚡ Früh bewerben McCarran, NV Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 5 Std.
Coupang
Staff Software Engineer, Ads
Coupang
⚡ Früh bewerben Seattle, USA Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 6 Std.
CI
Staff Software Engineer, Ads
Coupang Internal
⚡ Früh bewerben Seattle, USA Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 6 Std.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Equinix

Alle Jobs bei Equinix ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos