Jobs โ€บ Companies โ€บ Weekday AI โ€บ Site Reliability Engineer

รœber diese Site Reliability Engineer Stelle bei Weekday AI

Weekday AI ยท Vor Ort ยท Hyderabad, Telangana, India

๐—ง๐—ต๐—ถ๐˜€ ๐—ฟ๐—ผ๐—น๐—ฒ ๐—ถ๐˜€ ๐—ณ๐—ผ๐—ฟ ๐—ผ๐—ป๐—ฒ ๐—ผ๐—ณ ๐˜๐—ต๐—ฒ ๐—ช๐—ฒ๐—ฒ๐—ธ๐—ฑ๐—ฎ๐˜†'๐˜€ ๐—ฐ๐—น๐—ถ๐—ฒ๐—ป๐˜๐˜€

๐—ฆ๐—ฎ๐—น๐—ฎ๐—ฟ๐˜† ๐—ฟ๐—ฎ๐—ป๐—ด๐—ฒ: ๐—ฅ๐˜€ ๐Ÿญ๐Ÿฎ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ - ๐—ฅ๐˜€ ๐Ÿฎ๐Ÿด๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ (๐—ถ๐—ฒ ๐—œ๐—ก๐—ฅ ๐Ÿญ๐Ÿฎ-๐Ÿฎ๐Ÿด ๐—Ÿ๐—ฃ๐—”)

Experience: 6+ yrs

Location: Hyderabad, Bengaluru, Pune, Chennai, Tamil Nadu, India, Mumbai, Maharashtra, India

Job Type: Full-time

We are seeking an experiencedย Kafka Platform Engineerย with strong expertise in distributed systems, large-scale messaging platforms, and production operations. This role is ideal for professionals who are passionate about building, maintaining, and optimizing highly available streaming infrastructure while ensuring reliability, scalability, and operational excellence across enterprise environments.

As a Kafka Platform Engineer, you will be responsible for managing mission-critical messaging platforms, improving platform performance, automating operational processes, and supporting production environments. You will collaborate with infrastructure, application, DevOps, and engineering teams to deliver resilient streaming solutions, troubleshoot complex production issues, and continuously enhance platform reliability through automation, monitoring, and best practices.

Requirements

Key Responsibilities

  • Design, deploy, manage, and optimize Apache Kafka clusters and large-scale messaging or streaming platforms.
  • Monitor platform health, system performance, and resource utilization using modern monitoring, logging, and alerting tools.
  • Troubleshoot production incidents, identify root causes, and implement long-term solutions to improve system stability.
  • Automate operational tasks, deployments, and maintenance activities using scripting and infrastructure automation techniques.
  • Collaborate with application and infrastructure teams to support messaging architecture, integrations, and production workloads.
  • Implement performance tuning, capacity planning, and scalability improvements for distributed systems.
  • Maintain high availability, fault tolerance, and disaster recovery strategies for messaging infrastructure.
  • Ensure platform security, system compliance, and operational best practices across production environments.
  • Develop operational documentation, runbooks, and knowledge-sharing resources to improve support efficiency.
  • Participate in on-call support, incident response, and continuous improvement initiatives to maintain service reliability.

What Makes You a Great Fit

  • 6+ years of experience managing distributed systems, production infrastructure, or platform engineering environments.
  • Strong hands-on experience withย Apache Kafkaย or large-scale messaging and event streaming platforms.
  • Deep understanding of distributed systems architecture, scalability, fault tolerance, and production operations.
  • Experience with monitoring, logging, alerting, and observability tools for enterprise infrastructure.
  • Proficiency in at least one scripting or programming language such asย Python,ย Bash, orย Java.
  • Strong knowledge of Linux system administration, networking fundamentals, and troubleshooting methodologies.
  • Experience automating operational workflows and improving platform reliability through scripting and infrastructure automation.
  • Excellent analytical, debugging, and problem-solving skills with a proactive operational mindset.
  • Strong communication and collaboration skills with the ability to work effectively across cross-functional engineering teams.
  • Passion for building reliable, secure, and highly available platform infrastructure while continuously improving operational excellence.
Bereit, sich bei Weekday AI zu bewerben?
Bei Weekday AI bewerben

รœber Weekday AI

At Weekday (backed by YC; also Product Hunt #1 product of the day), we are building the next frontier in hiring. We have built the largest database of white collar talent in India and have built outreach tools on top of it to generate highest response rates.

Alle Jobs bei Weekday AI ansehen โ†’

ร„hnliche Jobs

Jalasoft
Site Reliability Engineer - AWS and Azure
Jalasoft
โšก Frรผh bewerben Colombia ยท standortgebunden
โ— Neu ๐Ÿ‘ Gesehen โœ“ Beworben vor 3 Std.
Phantom
Staff Software Engineer (SRE)
Phantom
โšก Frรผh bewerben Remote ยท standortgebunden $200,000โ€“$250,000
โ— Neu ๐Ÿ‘ Gesehen โœ“ Beworben vor 3 Std.
Cohere
Site Reliability Engineer, Inference Infrastructure
Cohere
โšก Frรผh bewerben Toronto ยท standortgebunden ยฃ156,000โ€“ยฃ156,000
โ— Neu ๐Ÿ‘ Gesehen โœ“ Beworben vor 3 Std.
Outreach
Staff Site Reliability Engineer (COR, AZURE) - Prague, Czechia
Outreach
โšก Frรผh bewerben Prague Hybrid
โ— Neu ๐Ÿ‘ Gesehen โœ“ Beworben vor 4 Std.
Mainspringenergy
Staff Reliability Engineer
Mainspringenergy
โšก Frรผh bewerben Menlo Park, CA Vor Ort
โ— Neu ๐Ÿ‘ Gesehen โœ“ Beworben vor 4 Std.
Roblox
Senior Machine Learning Engineer, Reliability
Roblox
โšก Frรผh bewerben San Mateo, CA, United States Vor Ort $196,750โ€“$243,290
โ— Neu ๐Ÿ‘ Gesehen โœ“ Beworben vor 4 Std.
Roblox
Senior Site Reliability Engineer, Compute
Roblox
โšก Frรผh bewerben San Mateo, CA, United States Vor Ort $243,290โ€“$295,250
โ— Neu ๐Ÿ‘ Gesehen โœ“ Beworben vor 4 Std.
Gorgias
Senior SRE Engineer - Security
Gorgias
โšก Frรผh bewerben Paris Vor Ort โ‚ฌ84,348โ€“โ‚ฌ93,227
โ— Neu ๐Ÿ‘ Gesehen โœ“ Beworben vor 6 Std.
Lightmatter
Reliability Engineer (Hardware)
Lightmatter
โšก Frรผh bewerben Mountain View, CA Vor Ort $142,000โ€“$200,000
โ— Neu ๐Ÿ‘ Gesehen โœ“ Beworben vor 6 Std.

Registrieren fรผr Vorschlรคge, die auf die von Ihnen geรถffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Weekday AI

Alle Jobs bei Weekday AI ansehen โ†’

Jetzt bewerben
๐Ÿค–

Moment โ€” langsam

JobsRadar wurde fรผr echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben โ€” nicht fรผr automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorรผbergehend blockiert.

Kommen Sie spรคter wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen โ€” verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei fรผr das, was dir hilft, die Stelle zu bekommen โ€” Gehaltsbenchmarks, den wรถchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten โ€” kostenlos