Jobs Companies Sysco Senior Site Reliability Engineer

À propos de ce poste Senior Site Reliability Engineer chez Sysco

Sysco · Hybride · Sysco LABS - Sri Lanka

JOB DESCRIPTION

Senior Site Reliability Engineer

About the Role

Join the Sysco Commercial Technology (CT) Site Reliability Engineering team as a Senior Software Site Reliability Engineer, where you will improve the reliability, scalability, performance, security, and operational excellence of Sysco's Commercial Technology ecosystem. Our team is responsible for delivering end-to-end reliability across the technology stack—from cloud infrastructure and platform services to APIs, applications, and customer-facing digital experiences—supporting enterprise API platforms, digital commerce, B2B integrations, and other business-critical systems across the CT landscape.

This is a Software SRE role focused on applying software engineering principles to solve reliability challenges at scale. You will design and build automation, develop internal tools and platforms, enhance observability, reduce operational toil, and deliver engineering solutions that strengthen the reliability and resilience of distributed systems.

As a Senior Software Site Reliability Engineer, you will combine software engineering, systems thinking, and production reliability expertise to identify and address reliability gaps, improve incident response, and drive long-term reliability initiatives. Working closely with product, platform, and infrastructure engineering teams, you will help build highly available, scalable, and resilient services while embedding reliability throughout the software development lifecycle.

What You'll Do

  • Own the reliability, scalability, performance, and operational excellence of critical platforms and services across the Commercial Technology ecosystem.
  • Apply software engineering principles to design and build systems that improve reliability, resilience, and engineering productivity.
  • Design, develop, and maintain automation, internal platforms, self-service capabilities, and reliability engineering tools that eliminate operational toil.
  • Define and evolve service reliability through SLIs, SLOs, error budgets, observability standards, and actionable operational metrics.
  • Partner with product, platform, and infrastructure engineering teams to design reliable, scalable, and operable systems from inception through production.
  • Build engineering solutions that improve deployment safety, release automation, progressive delivery, and production readiness.
  • Investigate complex production issues using application code, distributed systems knowledge, logs, metrics, traces, and infrastructure telemetry to identify systemic failures and drive permanent solutions.
  • Lead incident reviews and postmortems, ensuring root causes are understood and translated into engineering improvements rather than operational workarounds.
  • Drive continuous reliability improvements by identifying recurring operational patterns and solving them through software, automation, or platform capabilities.
  • Improve the resilience of distributed systems through capacity planning, performance engineering, disaster recovery, and resilience testing.
  • Participate in architecture and design reviews to ensure systems are reliable, scalable, observable, secure, and operationally efficient.
  • Contribute to shared engineering libraries, frameworks, and platform capabilities that enable development teams to build and operate reliable services.
  • Participate in an on-call rotation, using operational experience to continuously improve system reliability, automate manual work, and reduce future operational burden.
  • Mentor engineers and champion Software SRE principles, engineering excellence, and a culture of reliability across the organization.

What We're Looking For?

  • 4+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Production Engineering, or a related engineering role.
  • Strong software engineering skills with experience designing, building, testing, and operating reliable production systems.
  • Proficiency in one or more programming languages such as Java, Go, Python, or JavaScript/TypeScript, with a focus on writing maintainable, production-quality code.
  • Solid understanding of distributed systems, cloud-native architectures, microservices, APIs, databases, messaging systems, caching, networking, and system design.
  • Experience applying Site Reliability Engineering principles, including SLIs, SLOs, error budgets, incident management, postmortems, toil reduction, and automation.
  • Experience building automation, internal tools, developer platforms, or operational tooling to improve reliability and engineering productivity.
  • Hands-on experience with observability practices, including metrics, logs, traces, profiling, and modern observability platforms such as Datadog, Prometheus, Grafana, OpenTelemetry, Splunk, or ELK.
  • Experience debugging complex production issues across applications, distributed systems, APIs, infrastructure, and cloud platforms.
  • Experience with Kubernetes, containers, CI/CD, Infrastructure as Code, and public cloud platforms.
  • Ability to read, understand, and debug application code to identify systemic issues and implement long-term engineering solutions.
  • Strong analytical and systems thinking skills, with the ability to translate operational challenges into scalable engineering improvements.
  • Excellent collaboration and communication skills, with experience working across product, software, platform, and infrastructure engineering teams.
  • Demonstrated ownership, curiosity, and a continuous improvement mindset, with the ability to drive reliability initiatives independently.

Nice to Have

  • Experience with large-scale distributed systems, enterprise SaaS, digital commerce, or high-traffic platforms.
  • Experience building internal developer platforms, automation frameworks, or reliability engineering tools.
  • Experience with AWS, GCP, or Azure and cloud-native technologies such as Kubernetes.
  • Experience with Terraform, Helm, Argo CD, Jenkins, GitHub Actions, or similar platform engineering tools.
  • Experience with service mesh, API gateways, messaging systems, caching, database reliability, or resilience engineering.
  • Experience with AIOps, AI-assisted operations, or modern observability platforms.

Why Join Us?

  • Build software that improves the reliability of business-critical platforms across Sysco's Commercial Technology ecosystem.
  • Solve large-scale distributed systems challenges using software engineering, automation, and platform engineering.
  • Shape the future of Software SRE by driving automation, observability, and engineering best practices.
  • Work across the entire technology stack—from cloud infrastructure and Kubernetes to APIs and applications.
  • Collaborate with talented engineers while making a measurable impact on reliability, scalability, and operational excellence.

Benefits 

  

  • Performance-based annual bonus  

  • Performance rewards and recognition  

  • Agile Benefits - special allowances for Health, Wellness & Academic purposes  

  • Paid birthday leave Team engagement allowance  

  • Comprehensive Health & Life Insurance Cover - extendable to parents and in-laws  

  • Hybrid work arrangement  

  

  

Sysco LABS is an Equal Opportunity Employer.

Prêt à postuler chez Sysco ?
Postuler chez Sysco

À propos de Sysco

Sysco is the global leader in selling, marketing and distributing food products to restaurants, healthcare and educational facilities, lodging establishments and other customers who prepare meals away from home. Its family of products also includes equipment and supplies for the foodservice and hospitality industries. With more than 71,000 colleagues, the company operates 333 distribution facilities worldwide and serves approximately 700,000 customer locations. For fiscal year 2022 that ended July 2, 2022, the company generated sales of more than $68 billion. Information about our Sustainability program, including Sysco’s 2022 Sustainability Report and 2022 Diversity, Equity & Inclusion Repo

Voir tous les emplois chez Sysco →

Emplois similaires

RELX
Site Reliability Engineering Lead
RELX
⚡ Postuler tôt Florida Sur site $118,300–$219,800
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
RELX
Site Reliability Engineer II
RELX
⚡ Postuler tôt Home based-Georgia · lieu restreint $71,600–$119,400
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
RELX
FinOps Senior Site Reliability Engineer II
RELX
⚡ Postuler tôt Boca Raton, FL (Yamato) Sur site $104,900–$174,700
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
AES
Engineer, Reliability
AES
⚡ Postuler tôt US, Louisville, CO Sur site $94,000–$112,625
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
AES
Senior Reliability Engineer
AES
⚡ Postuler tôt US, Louisville, CO Sur site $113,000–$141,525
● Nouveau 👁 Vu ✓ Postulé il y a 1 h
Anduril Industries
Senior Infrastructure Reliability Engineer
Anduril Industries
⚡ Postuler tôt Costa Mesa, California, United... Sur site $166,000–$220,000
● Nouveau 👁 Vu ✓ Postulé il y a 7 h
Anduril Industries
Site Reliability Engineer
Anduril Industries
⚡ Postuler tôt Waltham, Massachusetts, United... Sur site $166,000–$220,000
● Nouveau 👁 Vu ✓ Postulé il y a 7 h
Anduril Industries
Site Reliability Engineer, Intelligence Systems
Anduril Industries
⚡ Postuler tôt Reston, Virginia, United State... Sur site $146,000–$194,000
● Nouveau 👁 Vu ✓ Postulé il y a 7 h
Roku
Senior Software Engineer, MLOps/SRE
Roku
⚡ Postuler tôt Bengaluru, India Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 8 h

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez Sysco

Voir tous les emplois chez Sysco →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit