Jobs Companies Weekday AI Site Reliability Engineer

Sobre este puesto de Site Reliability Engineer en Weekday AI

Weekday AI · Presencial · Hyderabad, Telangana, India

𝗧𝗵𝗶𝘀 𝗿𝗼𝗹𝗲 𝗶𝘀 𝗳𝗼𝗿 𝗼𝗻𝗲 𝗼𝗳 𝘁𝗵𝗲 𝗪𝗲𝗲𝗸𝗱𝗮𝘆'𝘀 𝗰𝗹𝗶𝗲𝗻𝘁𝘀

𝗦𝗮𝗹𝗮𝗿𝘆 𝗿𝗮𝗻𝗴𝗲: 𝗥𝘀 𝟭𝟮𝟬𝟬𝟬𝟬𝟬 - 𝗥𝘀 𝟮𝟴𝟬𝟬𝟬𝟬𝟬 (𝗶𝗲 𝗜𝗡𝗥 𝟭𝟮-𝟮𝟴 𝗟𝗣𝗔)

Experience: 6+ yrs

Location: Hyderabad, Bengaluru, Pune, Chennai, Tamil Nadu, India, Mumbai, Maharashtra, India

Job Type: Full-time

We are seeking an experienced Kafka Platform Engineer with strong expertise in distributed systems, large-scale messaging platforms, and production operations. This role is ideal for professionals who are passionate about building, maintaining, and optimizing highly available streaming infrastructure while ensuring reliability, scalability, and operational excellence across enterprise environments.

As a Kafka Platform Engineer, you will be responsible for managing mission-critical messaging platforms, improving platform performance, automating operational processes, and supporting production environments. You will collaborate with infrastructure, application, DevOps, and engineering teams to deliver resilient streaming solutions, troubleshoot complex production issues, and continuously enhance platform reliability through automation, monitoring, and best practices.

Requirements

Key Responsibilities

  • Design, deploy, manage, and optimize Apache Kafka clusters and large-scale messaging or streaming platforms.
  • Monitor platform health, system performance, and resource utilization using modern monitoring, logging, and alerting tools.
  • Troubleshoot production incidents, identify root causes, and implement long-term solutions to improve system stability.
  • Automate operational tasks, deployments, and maintenance activities using scripting and infrastructure automation techniques.
  • Collaborate with application and infrastructure teams to support messaging architecture, integrations, and production workloads.
  • Implement performance tuning, capacity planning, and scalability improvements for distributed systems.
  • Maintain high availability, fault tolerance, and disaster recovery strategies for messaging infrastructure.
  • Ensure platform security, system compliance, and operational best practices across production environments.
  • Develop operational documentation, runbooks, and knowledge-sharing resources to improve support efficiency.
  • Participate in on-call support, incident response, and continuous improvement initiatives to maintain service reliability.

What Makes You a Great Fit

  • 6+ years of experience managing distributed systems, production infrastructure, or platform engineering environments.
  • Strong hands-on experience with Apache Kafka or large-scale messaging and event streaming platforms.
  • Deep understanding of distributed systems architecture, scalability, fault tolerance, and production operations.
  • Experience with monitoring, logging, alerting, and observability tools for enterprise infrastructure.
  • Proficiency in at least one scripting or programming language such as PythonBash, or Java.
  • Strong knowledge of Linux system administration, networking fundamentals, and troubleshooting methodologies.
  • Experience automating operational workflows and improving platform reliability through scripting and infrastructure automation.
  • Excellent analytical, debugging, and problem-solving skills with a proactive operational mindset.
  • Strong communication and collaboration skills with the ability to work effectively across cross-functional engineering teams.
  • Passion for building reliable, secure, and highly available platform infrastructure while continuously improving operational excellence.
¿Listo para postularte en Weekday AI?
Postúlate en Weekday AI

Sobre Weekday AI

At Weekday (backed by YC; also Product Hunt #1 product of the day), we are building the next frontier in hiring. We have built the largest database of white collar talent in India and have built outreach tools on top of it to generate highest response rates.

Ver todos los empleos en Weekday AI →

Empleos similares

Jalasoft
Site Reliability Engineer - AWS and Azure
Jalasoft
⚡ Postúlate pronto Colombia · restringido por ubicación
● Nuevo 👁 Visto ✓ Postulado hace 4h
Phantom
Staff Software Engineer (SRE)
Phantom
⚡ Postúlate pronto Remote · restringido por ubicación $200,000–$250,000
● Nuevo 👁 Visto ✓ Postulado hace 4h
Cohere
Site Reliability Engineer, Inference Infrastructure
Cohere
⚡ Postúlate pronto Toronto · restringido por ubicación £156,000–£156,000
● Nuevo 👁 Visto ✓ Postulado hace 4h
Outreach
Staff Site Reliability Engineer (COR, AZURE) - Prague, Czechia
Outreach
⚡ Postúlate pronto Prague Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 5h
Mainspringenergy
Staff Reliability Engineer
Mainspringenergy
⚡ Postúlate pronto Menlo Park, CA Presencial
● Nuevo 👁 Visto ✓ Postulado hace 5h
Roblox
Senior Machine Learning Engineer, Reliability
Roblox
⚡ Postúlate pronto San Mateo, CA, United States Presencial $196,750–$243,290
● Nuevo 👁 Visto ✓ Postulado hace 5h
Roblox
Senior Site Reliability Engineer, Compute
Roblox
⚡ Postúlate pronto San Mateo, CA, United States Presencial $243,290–$295,250
● Nuevo 👁 Visto ✓ Postulado hace 5h
Gorgias
Senior SRE Engineer - Security
Gorgias
⚡ Postúlate pronto Paris Presencial €84,348–€93,227
● Nuevo 👁 Visto ✓ Postulado hace 7h
Lightmatter
Reliability Engineer (Hardware)
Lightmatter
⚡ Postúlate pronto Mountain View, CA Presencial $142,000–$200,000
● Nuevo 👁 Visto ✓ Postulado hace 7h

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en Weekday AI

Ver todos los empleos en Weekday AI →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis