Jobs Companies Sysco Senior Site Reliability Engineer

Über diese Senior Site Reliability Engineer Stelle bei Sysco

Sysco · Hybrid · Sysco LABS - Sri Lanka

JOB DESCRIPTION

Senior Site Reliability Engineer

About the Role

Join the Sysco Commercial Technology (CT) Site Reliability Engineering team as a Senior Software Site Reliability Engineer, where you will improve the reliability, scalability, performance, security, and operational excellence of Sysco's Commercial Technology ecosystem. Our team is responsible for delivering end-to-end reliability across the technology stack—from cloud infrastructure and platform services to APIs, applications, and customer-facing digital experiences—supporting enterprise API platforms, digital commerce, B2B integrations, and other business-critical systems across the CT landscape.

This is a Software SRE role focused on applying software engineering principles to solve reliability challenges at scale. You will design and build automation, develop internal tools and platforms, enhance observability, reduce operational toil, and deliver engineering solutions that strengthen the reliability and resilience of distributed systems.

As a Senior Software Site Reliability Engineer, you will combine software engineering, systems thinking, and production reliability expertise to identify and address reliability gaps, improve incident response, and drive long-term reliability initiatives. Working closely with product, platform, and infrastructure engineering teams, you will help build highly available, scalable, and resilient services while embedding reliability throughout the software development lifecycle.

What You'll Do

  • Own the reliability, scalability, performance, and operational excellence of critical platforms and services across the Commercial Technology ecosystem.
  • Apply software engineering principles to design and build systems that improve reliability, resilience, and engineering productivity.
  • Design, develop, and maintain automation, internal platforms, self-service capabilities, and reliability engineering tools that eliminate operational toil.
  • Define and evolve service reliability through SLIs, SLOs, error budgets, observability standards, and actionable operational metrics.
  • Partner with product, platform, and infrastructure engineering teams to design reliable, scalable, and operable systems from inception through production.
  • Build engineering solutions that improve deployment safety, release automation, progressive delivery, and production readiness.
  • Investigate complex production issues using application code, distributed systems knowledge, logs, metrics, traces, and infrastructure telemetry to identify systemic failures and drive permanent solutions.
  • Lead incident reviews and postmortems, ensuring root causes are understood and translated into engineering improvements rather than operational workarounds.
  • Drive continuous reliability improvements by identifying recurring operational patterns and solving them through software, automation, or platform capabilities.
  • Improve the resilience of distributed systems through capacity planning, performance engineering, disaster recovery, and resilience testing.
  • Participate in architecture and design reviews to ensure systems are reliable, scalable, observable, secure, and operationally efficient.
  • Contribute to shared engineering libraries, frameworks, and platform capabilities that enable development teams to build and operate reliable services.
  • Participate in an on-call rotation, using operational experience to continuously improve system reliability, automate manual work, and reduce future operational burden.
  • Mentor engineers and champion Software SRE principles, engineering excellence, and a culture of reliability across the organization.

What We're Looking For?

  • 4+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Production Engineering, or a related engineering role.
  • Strong software engineering skills with experience designing, building, testing, and operating reliable production systems.
  • Proficiency in one or more programming languages such as Java, Go, Python, or JavaScript/TypeScript, with a focus on writing maintainable, production-quality code.
  • Solid understanding of distributed systems, cloud-native architectures, microservices, APIs, databases, messaging systems, caching, networking, and system design.
  • Experience applying Site Reliability Engineering principles, including SLIs, SLOs, error budgets, incident management, postmortems, toil reduction, and automation.
  • Experience building automation, internal tools, developer platforms, or operational tooling to improve reliability and engineering productivity.
  • Hands-on experience with observability practices, including metrics, logs, traces, profiling, and modern observability platforms such as Datadog, Prometheus, Grafana, OpenTelemetry, Splunk, or ELK.
  • Experience debugging complex production issues across applications, distributed systems, APIs, infrastructure, and cloud platforms.
  • Experience with Kubernetes, containers, CI/CD, Infrastructure as Code, and public cloud platforms.
  • Ability to read, understand, and debug application code to identify systemic issues and implement long-term engineering solutions.
  • Strong analytical and systems thinking skills, with the ability to translate operational challenges into scalable engineering improvements.
  • Excellent collaboration and communication skills, with experience working across product, software, platform, and infrastructure engineering teams.
  • Demonstrated ownership, curiosity, and a continuous improvement mindset, with the ability to drive reliability initiatives independently.

Nice to Have

  • Experience with large-scale distributed systems, enterprise SaaS, digital commerce, or high-traffic platforms.
  • Experience building internal developer platforms, automation frameworks, or reliability engineering tools.
  • Experience with AWS, GCP, or Azure and cloud-native technologies such as Kubernetes.
  • Experience with Terraform, Helm, Argo CD, Jenkins, GitHub Actions, or similar platform engineering tools.
  • Experience with service mesh, API gateways, messaging systems, caching, database reliability, or resilience engineering.
  • Experience with AIOps, AI-assisted operations, or modern observability platforms.

Why Join Us?

  • Build software that improves the reliability of business-critical platforms across Sysco's Commercial Technology ecosystem.
  • Solve large-scale distributed systems challenges using software engineering, automation, and platform engineering.
  • Shape the future of Software SRE by driving automation, observability, and engineering best practices.
  • Work across the entire technology stack—from cloud infrastructure and Kubernetes to APIs and applications.
  • Collaborate with talented engineers while making a measurable impact on reliability, scalability, and operational excellence.

Benefits 

  

  • Performance-based annual bonus  

  • Performance rewards and recognition  

  • Agile Benefits - special allowances for Health, Wellness & Academic purposes  

  • Paid birthday leave Team engagement allowance  

  • Comprehensive Health & Life Insurance Cover - extendable to parents and in-laws  

  • Hybrid work arrangement  

  

  

Sysco LABS is an Equal Opportunity Employer.

Bereit, sich bei Sysco zu bewerben?
Bei Sysco bewerben

Über Sysco

Sysco is the global leader in selling, marketing and distributing food products to restaurants, healthcare and educational facilities, lodging establishments and other customers who prepare meals away from home. Its family of products also includes equipment and supplies for the foodservice and hospitality industries. With more than 71,000 colleagues, the company operates 333 distribution facilities worldwide and serves approximately 700,000 customer locations. For fiscal year 2022 that ended July 2, 2022, the company generated sales of more than $68 billion. Information about our Sustainability program, including Sysco’s 2022 Sustainability Report and 2022 Diversity, Equity & Inclusion Repo

Alle Jobs bei Sysco ansehen →

Ähnliche Jobs

AES
Engineer, Reliability
AES
⚡ Früh bewerben US, Louisville, CO Vor Ort $94,000–$112,625
● Neu 👁 Gesehen ✓ Beworben vor 26 Min.
AES
Senior Reliability Engineer
AES
⚡ Früh bewerben US, Louisville, CO Vor Ort $113,000–$141,525
● Neu 👁 Gesehen ✓ Beworben vor 26 Min.
Anduril Industries
Senior Infrastructure Reliability Engineer
Anduril Industries
⚡ Früh bewerben Costa Mesa, California, United... Vor Ort $166,000–$220,000
● Neu 👁 Gesehen ✓ Beworben vor 5 Std.
Anduril Industries
Site Reliability Engineer
Anduril Industries
⚡ Früh bewerben Waltham, Massachusetts, United... Vor Ort $166,000–$220,000
● Neu 👁 Gesehen ✓ Beworben vor 5 Std.
Anduril Industries
Site Reliability Engineer, Intelligence Systems
Anduril Industries
⚡ Früh bewerben Reston, Virginia, United State... Vor Ort $146,000–$194,000
● Neu 👁 Gesehen ✓ Beworben vor 6 Std.
Roku
Senior Software Engineer, MLOps/SRE
Roku
⚡ Früh bewerben Bengaluru, India Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 7 Std.
Roku
Senior Software Engineer,  SRE
Roku
⚡ Früh bewerben Bengaluru, India Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 7 Std.
Roblox
Senior Site Reliability Engineer, Compute
Roblox
⚡ Früh bewerben San Mateo, CA, United States Vor Ort $243,290–$295,250
● Neu 👁 Gesehen ✓ Beworben vor 8 Std.
Genesys
Senior Operations Reliability Engineer - IAM
Genesys
⚡ Früh bewerben Ontario, Canada Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 22 Std.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Sysco

Alle Jobs bei Sysco ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos