Jobs Companies Sysco Senior Site Reliability Engineer

About this Senior Site Reliability Engineer role at Sysco

Sysco · Hybrid · Sysco LABS - Sri Lanka

JOB DESCRIPTION

Senior Site Reliability Engineer

About the Role

Join the Sysco Commercial Technology (CT) Site Reliability Engineering team as a Senior Software Site Reliability Engineer, where you will improve the reliability, scalability, performance, security, and operational excellence of Sysco's Commercial Technology ecosystem. Our team is responsible for delivering end-to-end reliability across the technology stack—from cloud infrastructure and platform services to APIs, applications, and customer-facing digital experiences—supporting enterprise API platforms, digital commerce, B2B integrations, and other business-critical systems across the CT landscape.

This is a Software SRE role focused on applying software engineering principles to solve reliability challenges at scale. You will design and build automation, develop internal tools and platforms, enhance observability, reduce operational toil, and deliver engineering solutions that strengthen the reliability and resilience of distributed systems.

As a Senior Software Site Reliability Engineer, you will combine software engineering, systems thinking, and production reliability expertise to identify and address reliability gaps, improve incident response, and drive long-term reliability initiatives. Working closely with product, platform, and infrastructure engineering teams, you will help build highly available, scalable, and resilient services while embedding reliability throughout the software development lifecycle.

What You'll Do

  • Own the reliability, scalability, performance, and operational excellence of critical platforms and services across the Commercial Technology ecosystem.
  • Apply software engineering principles to design and build systems that improve reliability, resilience, and engineering productivity.
  • Design, develop, and maintain automation, internal platforms, self-service capabilities, and reliability engineering tools that eliminate operational toil.
  • Define and evolve service reliability through SLIs, SLOs, error budgets, observability standards, and actionable operational metrics.
  • Partner with product, platform, and infrastructure engineering teams to design reliable, scalable, and operable systems from inception through production.
  • Build engineering solutions that improve deployment safety, release automation, progressive delivery, and production readiness.
  • Investigate complex production issues using application code, distributed systems knowledge, logs, metrics, traces, and infrastructure telemetry to identify systemic failures and drive permanent solutions.
  • Lead incident reviews and postmortems, ensuring root causes are understood and translated into engineering improvements rather than operational workarounds.
  • Drive continuous reliability improvements by identifying recurring operational patterns and solving them through software, automation, or platform capabilities.
  • Improve the resilience of distributed systems through capacity planning, performance engineering, disaster recovery, and resilience testing.
  • Participate in architecture and design reviews to ensure systems are reliable, scalable, observable, secure, and operationally efficient.
  • Contribute to shared engineering libraries, frameworks, and platform capabilities that enable development teams to build and operate reliable services.
  • Participate in an on-call rotation, using operational experience to continuously improve system reliability, automate manual work, and reduce future operational burden.
  • Mentor engineers and champion Software SRE principles, engineering excellence, and a culture of reliability across the organization.

What We're Looking For?

  • 4+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Production Engineering, or a related engineering role.
  • Strong software engineering skills with experience designing, building, testing, and operating reliable production systems.
  • Proficiency in one or more programming languages such as Java, Go, Python, or JavaScript/TypeScript, with a focus on writing maintainable, production-quality code.
  • Solid understanding of distributed systems, cloud-native architectures, microservices, APIs, databases, messaging systems, caching, networking, and system design.
  • Experience applying Site Reliability Engineering principles, including SLIs, SLOs, error budgets, incident management, postmortems, toil reduction, and automation.
  • Experience building automation, internal tools, developer platforms, or operational tooling to improve reliability and engineering productivity.
  • Hands-on experience with observability practices, including metrics, logs, traces, profiling, and modern observability platforms such as Datadog, Prometheus, Grafana, OpenTelemetry, Splunk, or ELK.
  • Experience debugging complex production issues across applications, distributed systems, APIs, infrastructure, and cloud platforms.
  • Experience with Kubernetes, containers, CI/CD, Infrastructure as Code, and public cloud platforms.
  • Ability to read, understand, and debug application code to identify systemic issues and implement long-term engineering solutions.
  • Strong analytical and systems thinking skills, with the ability to translate operational challenges into scalable engineering improvements.
  • Excellent collaboration and communication skills, with experience working across product, software, platform, and infrastructure engineering teams.
  • Demonstrated ownership, curiosity, and a continuous improvement mindset, with the ability to drive reliability initiatives independently.

Nice to Have

  • Experience with large-scale distributed systems, enterprise SaaS, digital commerce, or high-traffic platforms.
  • Experience building internal developer platforms, automation frameworks, or reliability engineering tools.
  • Experience with AWS, GCP, or Azure and cloud-native technologies such as Kubernetes.
  • Experience with Terraform, Helm, Argo CD, Jenkins, GitHub Actions, or similar platform engineering tools.
  • Experience with service mesh, API gateways, messaging systems, caching, database reliability, or resilience engineering.
  • Experience with AIOps, AI-assisted operations, or modern observability platforms.

Why Join Us?

  • Build software that improves the reliability of business-critical platforms across Sysco's Commercial Technology ecosystem.
  • Solve large-scale distributed systems challenges using software engineering, automation, and platform engineering.
  • Shape the future of Software SRE by driving automation, observability, and engineering best practices.
  • Work across the entire technology stack—from cloud infrastructure and Kubernetes to APIs and applications.
  • Collaborate with talented engineers while making a measurable impact on reliability, scalability, and operational excellence.

Benefits 

  

  • Performance-based annual bonus  

  • Performance rewards and recognition  

  • Agile Benefits - special allowances for Health, Wellness & Academic purposes  

  • Paid birthday leave Team engagement allowance  

  • Comprehensive Health & Life Insurance Cover - extendable to parents and in-laws  

  • Hybrid work arrangement  

  

  

Sysco LABS is an Equal Opportunity Employer.

Ready to apply to Sysco?
Apply to Sysco

About Sysco

Sysco is the global leader in selling, marketing and distributing food products to restaurants, healthcare and educational facilities, lodging establishments and other customers who prepare meals away from home. Its family of products also includes equipment and supplies for the foodservice and hospitality industries. With more than 71,000 colleagues, the company operates 333 distribution facilities worldwide and serves approximately 700,000 customer locations. For fiscal year 2022 that ended July 2, 2022, the company generated sales of more than $68 billion. Information about our Sustainability program, including Sysco’s 2022 Sustainability Report and 2022 Diversity, Equity & Inclusion Repo

See all jobs at Sysco →

Similar jobs

Anduril Industries
Senior Infrastructure Reliability Engineer
Anduril Industries
⚡ Apply early Costa Mesa, California, United... Onsite $166,000–$220,000
● New 👁 Seen ✓ Applied 2h ago
Anduril Industries
Site Reliability Engineer
Anduril Industries
⚡ Apply early Waltham, Massachusetts, United... Onsite $166,000–$220,000
● New 👁 Seen ✓ Applied 2h ago
Anduril Industries
Site Reliability Engineer, Intelligence Systems
Anduril Industries
⚡ Apply early Reston, Virginia, United State... Onsite $146,000–$194,000
● New 👁 Seen ✓ Applied 3h ago
Roku
Senior Software Engineer, MLOps/SRE
Roku
⚡ Apply early Bengaluru, India Onsite
● New 👁 Seen ✓ Applied 4h ago
Roku
Senior Software Engineer,  SRE
Roku
⚡ Apply early Bengaluru, India Onsite
● New 👁 Seen ✓ Applied 4h ago
Roblox
Senior Site Reliability Engineer, Compute
Roblox
⚡ Apply early San Mateo, CA, United States Onsite $243,290–$295,250
● New 👁 Seen ✓ Applied 5h ago
Genesys
Senior Operations Reliability Engineer - IAM
Genesys
⚡ Apply early Ontario, Canada Onsite
● New 👁 Seen ✓ Applied 18h ago
LSEG
Senior Manager, Site Reliability Engineering
LSEG
⚡ Apply early IND-BLR-Divyasree Technopolis Onsite
● New 👁 Seen ✓ Applied 19h ago
NL
Senior Cloud Platform & Site Reliability Engineering Lead
National Life Insurance Company
⚡ Apply early Addison, TX; Montpelier, VT Onsite $136,875–$200,750
● New 👁 Seen ✓ Applied 19h ago

Sign up for suggestions tailored to the jobs you open and the searches you save.

More jobs at Sysco

See all jobs at Sysco →

Apply now
🤖

Whoa — hold up

JobsRadar was built for real people having a rough time in their job search — not for automated requests. You're clicking way too fast and you're now temporarily blocked.

Come back later. If you're genuinely job hunting, we've got your back — just act like a human.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Get an edge on your job hunt.

Join our Telegram channel for the stuff that helps you land the role — salary benchmarks, the weekly market pulse, and new-feature drops. No spam, just signal.

Join the channel — it's free