Jobs Companies Candescent Site Reliability Engineer III

À propos de ce poste Site Reliability Engineer III chez Candescent

Candescent · Sur site · IN - Bengaluru - Office

Candescent is a forward-thinking technology company transforming how financial institutions deliver Intelligent Banking experiences. We unite digital banking, account opening, and branch solutions that power and connect digital banking, account opening, and branch solutions—creating seamless engagement across digital, remote, and in-person channels.

Our Experience-Led, Intelligence-Driven approach combines human-centered design with data, automation, and cloud-based innovation. Built on an API-first architecture, our extensible ecosystem enables institutions to adapt quickly, integrate easily, and unlock new opportunities for growth—turning every customer interaction into a moment of clarity, confidence, and connection.

Position: Site Reliability Engineer III

Experience: 6-9 Years

Location: Bangalore

We are looking for a strong Application Site Reliability Engineer (SRE) to support and improve the reliability of Java-based production systems running on Kubernetes in cloud environments.

This role focuses on application-level reliability, JVM deep troubleshooting, production incident management, and close collaboration with development teams — not infrastructure provisioning or CloudOps.

The ideal candidate understands how Java applications behave in production and can proactively improve performance, scalability, and operational maturity.

Key Responsibilities:

  • Support and operate production Java applications running on Kubernetes (GKE).
  • Troubleshoot complex application issues using logs, metrics, traces, heap dumps, and thread dumps.
  • Participate in incident response, root cause analysis, and blameless postmortems.
  • Collaborate closely with development teams to understand application architecture, dependencies, and failure patterns.
  • Analyze JVM behavior (heap, GC, memory leaks, OOM, thread contention) and recommend performance improvements.
  • Define and improve SLIs, SLOs, alerts, and dashboards.
  • Support application deployments, rollbacks, and runtime configuration changes.
  • Identify reliability, performance, and scalability gaps in application behavior.
  • Automate repetitive operational tasks to reduce toil.
  • Drive improvements in runbooks, operational readiness, and on-call effectiveness.
  • Advocate and influence adoption of shift-left reliability practices.

Must-Have Skills & Experience:

  • Strong hands-on experience supporting Java applications in production.
  • Deep understanding of JVM internals:
    • Heap & memory management
    • Garbage collection
    • tuning OOM analysis
    • Thread dump and performance analysis
  • Proven experience in incident response and production troubleshooting.
  • Experience operating applications on Kubernetes from an application/runtime perspective.
  • Strong experience with application observability:
    • Logs
    • Metrics
    • Monitoring tools
    • Distributed tracing
  • Solid understanding of SLIs, SLOs, and reliability-driven operations.
  • Experience with deployment strategies (rolling, blue/green, canary).
  • Ability to write scripts/automation (Python, Shell, or similar) to reduce operational toil.
  • Strong understanding of application architecture and service dependencies (databases, messaging systems, external APIs).
  • Ability to analyze and troubleshoot issues holistically across the entire application stack, rather than focusing on isolated components.
  • Strong collaboration and communication skills.
  • Demonstrates accountability and sound judgment during high-pressure production incidents.

Cloud & Platform Exposure

  • Experience working with applications deployed on Kubernetes in GCP.
  • Familiarity with GKE environments from an application operations perspective (not CloudOps or infrastructure engineering).
  • Understanding of cloud constructs relevant to application behavior (networking, IAM, storage, compute).

Good-to-Have Skills

  • CI/CD pipeline exposure (GitHub Actions, Jenkins).
  • Familiarity with GitOps practices.
  • Experience supporting cloud migrations or modernization initiatives.
  • Exposure to platform or infrastructure concepts supporting application workloads.

What We Value

  • Ownership mindset and reliability-first thinking.
  • Curiosity to investigate deep production issues.
  • Strong collaboration with development teams.
  • Bias toward automation and continuous improvement.
  • Clear communication during incidents and stakeholder updates.

Statement to Third Party Agencies
To ALL recruitment agencies: Candescent only accepts resumes from agencies on the preferred supplier list. Please do not forward resumes to our applicant tracking system, Candescent employees, or any Candescent facility. Candescent is not responsible for any fees or charges associated with unsolicited resumes.

Prêt à postuler chez Candescent ?
Postuler chez Candescent

À propos de Candescent

Candescent is the largest non-core digital banking provider. We bring together the transformative technologies that power and connect account opening, digital banking and branch solutions for banks and credit unions of all sizes on any core.

Voir tous les emplois chez Candescent →

Emplois similaires

AES
Senior Reliability Engineer
AES
⚡ Postuler tôt US, Louisville, CO Sur site $113,000–$141,525
● Nouveau 👁 Vu ✓ Postulé il y a 6 h
Allison Transmission
Senior Quality & Reliability Engineer
Allison Transmission
⚡ Postuler tôt Indianapolis, IN Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 5 j
IEX Group
Senior Systems Reliability Engineer
IEX Group
⚡ Postuler tôt Remote - Must reside in Califo... · lieu restreint $180,000–$225,000
● Nouveau 👁 Vu ✓ Postulé il y a 6 j
ST
Senior Site Reliability Engineer
StarRez
⚡ Postuler tôt Hyderabad, Telengana, IN Sur site
● Nouveau 👁 Vu ✓ Postulé il y a 3 sem.
CI
Senior Site Reliability Engineer - CLO
Circle
⚡ Postuler tôt San Francisco - remote first i... · lieu restreint $152,500–$205,000
● Nouveau 👁 Vu ✓ Postulé il y a 3 sem.
SA
Senior Site Reliability Engineer
Sanity
⚡ Postuler tôt Remote in the United States or... · lieu restreint
● Nouveau 👁 Vu ✓ Postulé il y a 4 sem.
Confluent
Staff Software Engineer I - SRE
Confluent
⚡ Postuler tôt IN Remote India · lieu restreint
● Nouveau 👁 Vu ✓ Postulé il y a 4 sem.
CI
Senior Site Reliability Engineer - Infra Ops
Circle
⚡ Postuler tôt San Francisco - remote first i... · lieu restreint $152,500–$205,000
● Nouveau 👁 Vu ✓ Postulé il y a 1 mois
CI
Staff Site Reliability Engineer
Circle
⚡ Postuler tôt San Francisco - remote first i... · lieu restreint $195,000–$257,500
● Nouveau 👁 Vu ✓ Postulé il y a 1 mois

Inscrivez-vous pour des suggestions adaptées aux emplois que vous ouvrez et aux recherches que vous enregistrez.

Plus d’emplois chez Candescent

Voir tous les emplois chez Candescent →

Postuler maintenant
🤖

Doucement — un instant

JobsRadar a été conçu pour de vraies personnes qui traversent une période difficile dans leur recherche d’emploi — pas pour des requêtes automatisées. Vous cliquez beaucoup trop vite et vous êtes maintenant temporairement bloqué.

Revenez plus tard. Si vous cherchez réellement un emploi, nous sommes de votre côté — agissez simplement comme un être humain.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Prenez une longueur d’avance dans votre recherche d’emploi.

Rejoignez notre canal Telegram pour ce qui vous aide à décrocher le poste — références salariales, le pouls hebdomadaire du marché et les annonces de nouveautés. Pas de spam, que du signal.

Rejoindre le canal — c’est gratuit