Jobs Companies Sysco Senior Site Reliability Engineer

Sobre este puesto de Senior Site Reliability Engineer en Sysco

Sysco · Híbrido · Sysco LABS - Sri Lanka

JOB DESCRIPTION

Senior Site Reliability Engineer

About the Role

Join the Sysco Commercial Technology (CT) Site Reliability Engineering team as a Senior Software Site Reliability Engineer, where you will improve the reliability, scalability, performance, security, and operational excellence of Sysco's Commercial Technology ecosystem. Our team is responsible for delivering end-to-end reliability across the technology stack—from cloud infrastructure and platform services to APIs, applications, and customer-facing digital experiences—supporting enterprise API platforms, digital commerce, B2B integrations, and other business-critical systems across the CT landscape.

This is a Software SRE role focused on applying software engineering principles to solve reliability challenges at scale. You will design and build automation, develop internal tools and platforms, enhance observability, reduce operational toil, and deliver engineering solutions that strengthen the reliability and resilience of distributed systems.

As a Senior Software Site Reliability Engineer, you will combine software engineering, systems thinking, and production reliability expertise to identify and address reliability gaps, improve incident response, and drive long-term reliability initiatives. Working closely with product, platform, and infrastructure engineering teams, you will help build highly available, scalable, and resilient services while embedding reliability throughout the software development lifecycle.

What You'll Do

  • Own the reliability, scalability, performance, and operational excellence of critical platforms and services across the Commercial Technology ecosystem.
  • Apply software engineering principles to design and build systems that improve reliability, resilience, and engineering productivity.
  • Design, develop, and maintain automation, internal platforms, self-service capabilities, and reliability engineering tools that eliminate operational toil.
  • Define and evolve service reliability through SLIs, SLOs, error budgets, observability standards, and actionable operational metrics.
  • Partner with product, platform, and infrastructure engineering teams to design reliable, scalable, and operable systems from inception through production.
  • Build engineering solutions that improve deployment safety, release automation, progressive delivery, and production readiness.
  • Investigate complex production issues using application code, distributed systems knowledge, logs, metrics, traces, and infrastructure telemetry to identify systemic failures and drive permanent solutions.
  • Lead incident reviews and postmortems, ensuring root causes are understood and translated into engineering improvements rather than operational workarounds.
  • Drive continuous reliability improvements by identifying recurring operational patterns and solving them through software, automation, or platform capabilities.
  • Improve the resilience of distributed systems through capacity planning, performance engineering, disaster recovery, and resilience testing.
  • Participate in architecture and design reviews to ensure systems are reliable, scalable, observable, secure, and operationally efficient.
  • Contribute to shared engineering libraries, frameworks, and platform capabilities that enable development teams to build and operate reliable services.
  • Participate in an on-call rotation, using operational experience to continuously improve system reliability, automate manual work, and reduce future operational burden.
  • Mentor engineers and champion Software SRE principles, engineering excellence, and a culture of reliability across the organization.

What We're Looking For?

  • 4+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Production Engineering, or a related engineering role.
  • Strong software engineering skills with experience designing, building, testing, and operating reliable production systems.
  • Proficiency in one or more programming languages such as Java, Go, Python, or JavaScript/TypeScript, with a focus on writing maintainable, production-quality code.
  • Solid understanding of distributed systems, cloud-native architectures, microservices, APIs, databases, messaging systems, caching, networking, and system design.
  • Experience applying Site Reliability Engineering principles, including SLIs, SLOs, error budgets, incident management, postmortems, toil reduction, and automation.
  • Experience building automation, internal tools, developer platforms, or operational tooling to improve reliability and engineering productivity.
  • Hands-on experience with observability practices, including metrics, logs, traces, profiling, and modern observability platforms such as Datadog, Prometheus, Grafana, OpenTelemetry, Splunk, or ELK.
  • Experience debugging complex production issues across applications, distributed systems, APIs, infrastructure, and cloud platforms.
  • Experience with Kubernetes, containers, CI/CD, Infrastructure as Code, and public cloud platforms.
  • Ability to read, understand, and debug application code to identify systemic issues and implement long-term engineering solutions.
  • Strong analytical and systems thinking skills, with the ability to translate operational challenges into scalable engineering improvements.
  • Excellent collaboration and communication skills, with experience working across product, software, platform, and infrastructure engineering teams.
  • Demonstrated ownership, curiosity, and a continuous improvement mindset, with the ability to drive reliability initiatives independently.

Nice to Have

  • Experience with large-scale distributed systems, enterprise SaaS, digital commerce, or high-traffic platforms.
  • Experience building internal developer platforms, automation frameworks, or reliability engineering tools.
  • Experience with AWS, GCP, or Azure and cloud-native technologies such as Kubernetes.
  • Experience with Terraform, Helm, Argo CD, Jenkins, GitHub Actions, or similar platform engineering tools.
  • Experience with service mesh, API gateways, messaging systems, caching, database reliability, or resilience engineering.
  • Experience with AIOps, AI-assisted operations, or modern observability platforms.

Why Join Us?

  • Build software that improves the reliability of business-critical platforms across Sysco's Commercial Technology ecosystem.
  • Solve large-scale distributed systems challenges using software engineering, automation, and platform engineering.
  • Shape the future of Software SRE by driving automation, observability, and engineering best practices.
  • Work across the entire technology stack—from cloud infrastructure and Kubernetes to APIs and applications.
  • Collaborate with talented engineers while making a measurable impact on reliability, scalability, and operational excellence.

Benefits 

  

  • Performance-based annual bonus  

  • Performance rewards and recognition  

  • Agile Benefits - special allowances for Health, Wellness & Academic purposes  

  • Paid birthday leave Team engagement allowance  

  • Comprehensive Health & Life Insurance Cover - extendable to parents and in-laws  

  • Hybrid work arrangement  

  

  

Sysco LABS is an Equal Opportunity Employer.

¿Listo para postularte en Sysco?
Postúlate en Sysco

Sobre Sysco

Sysco is the global leader in selling, marketing and distributing food products to restaurants, healthcare and educational facilities, lodging establishments and other customers who prepare meals away from home. Its family of products also includes equipment and supplies for the foodservice and hospitality industries. With more than 71,000 colleagues, the company operates 333 distribution facilities worldwide and serves approximately 700,000 customer locations. For fiscal year 2022 that ended July 2, 2022, the company generated sales of more than $68 billion. Information about our Sustainability program, including Sysco’s 2022 Sustainability Report and 2022 Diversity, Equity & Inclusion Repo

Ver todos los empleos en Sysco →

Empleos similares

Anduril Industries
Senior Infrastructure Reliability Engineer
Anduril Industries
⚡ Postúlate pronto Costa Mesa, California, United... Presencial $166,000–$220,000
● Nuevo 👁 Visto ✓ Postulado hace 4h
Anduril Industries
Site Reliability Engineer
Anduril Industries
⚡ Postúlate pronto Waltham, Massachusetts, United... Presencial $166,000–$220,000
● Nuevo 👁 Visto ✓ Postulado hace 4h
Anduril Industries
Site Reliability Engineer, Intelligence Systems
Anduril Industries
⚡ Postúlate pronto Reston, Virginia, United State... Presencial $146,000–$194,000
● Nuevo 👁 Visto ✓ Postulado hace 5h
Roku
Senior Software Engineer, MLOps/SRE
Roku
⚡ Postúlate pronto Bengaluru, India Presencial
● Nuevo 👁 Visto ✓ Postulado hace 6h
Roku
Senior Software Engineer,  SRE
Roku
⚡ Postúlate pronto Bengaluru, India Presencial
● Nuevo 👁 Visto ✓ Postulado hace 6h
Roblox
Senior Site Reliability Engineer, Compute
Roblox
⚡ Postúlate pronto San Mateo, CA, United States Presencial $243,290–$295,250
● Nuevo 👁 Visto ✓ Postulado hace 7h
Genesys
Senior Operations Reliability Engineer - IAM
Genesys
⚡ Postúlate pronto Ontario, Canada Presencial
● Nuevo 👁 Visto ✓ Postulado hace 21h
LSEG
Senior Manager, Site Reliability Engineering
LSEG
⚡ Postúlate pronto IND-BLR-Divyasree Technopolis Presencial
● Nuevo 👁 Visto ✓ Postulado hace 21h
NL
Senior Cloud Platform & Site Reliability Engineering Lead
National Life Insurance Company
⚡ Postúlate pronto Addison, TX; Montpelier, VT Presencial $136,875–$200,750
● Nuevo 👁 Visto ✓ Postulado hace 22h

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en Sysco

Ver todos los empleos en Sysco →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis