Jobs Companies Candescent Site Reliability Engineer III

About this Site Reliability Engineer III role at Candescent

Candescent · Onsite · IN - Bengaluru - Office

Candescent is a forward-thinking technology company transforming how financial institutions deliver Intelligent Banking experiences. We unite digital banking, account opening, and branch solutions that power and connect digital banking, account opening, and branch solutions—creating seamless engagement across digital, remote, and in-person channels.

Our Experience-Led, Intelligence-Driven approach combines human-centered design with data, automation, and cloud-based innovation. Built on an API-first architecture, our extensible ecosystem enables institutions to adapt quickly, integrate easily, and unlock new opportunities for growth—turning every customer interaction into a moment of clarity, confidence, and connection.

Position: Site Reliability Engineer III

Experience: 6-9 Years

Location: Bangalore

We are looking for a strong Application Site Reliability Engineer (SRE) to support and improve the reliability of Java-based production systems running on Kubernetes in cloud environments.

This role focuses on application-level reliability, JVM deep troubleshooting, production incident management, and close collaboration with development teams — not infrastructure provisioning or CloudOps.

The ideal candidate understands how Java applications behave in production and can proactively improve performance, scalability, and operational maturity.

Key Responsibilities:

  • Support and operate production Java applications running on Kubernetes (GKE).
  • Troubleshoot complex application issues using logs, metrics, traces, heap dumps, and thread dumps.
  • Participate in incident response, root cause analysis, and blameless postmortems.
  • Collaborate closely with development teams to understand application architecture, dependencies, and failure patterns.
  • Analyze JVM behavior (heap, GC, memory leaks, OOM, thread contention) and recommend performance improvements.
  • Define and improve SLIs, SLOs, alerts, and dashboards.
  • Support application deployments, rollbacks, and runtime configuration changes.
  • Identify reliability, performance, and scalability gaps in application behavior.
  • Automate repetitive operational tasks to reduce toil.
  • Drive improvements in runbooks, operational readiness, and on-call effectiveness.
  • Advocate and influence adoption of shift-left reliability practices.

Must-Have Skills & Experience:

  • Strong hands-on experience supporting Java applications in production.
  • Deep understanding of JVM internals:
    • Heap & memory management
    • Garbage collection
    • tuning OOM analysis
    • Thread dump and performance analysis
  • Proven experience in incident response and production troubleshooting.
  • Experience operating applications on Kubernetes from an application/runtime perspective.
  • Strong experience with application observability:
    • Logs
    • Metrics
    • Monitoring tools
    • Distributed tracing
  • Solid understanding of SLIs, SLOs, and reliability-driven operations.
  • Experience with deployment strategies (rolling, blue/green, canary).
  • Ability to write scripts/automation (Python, Shell, or similar) to reduce operational toil.
  • Strong understanding of application architecture and service dependencies (databases, messaging systems, external APIs).
  • Ability to analyze and troubleshoot issues holistically across the entire application stack, rather than focusing on isolated components.
  • Strong collaboration and communication skills.
  • Demonstrates accountability and sound judgment during high-pressure production incidents.

Cloud & Platform Exposure

  • Experience working with applications deployed on Kubernetes in GCP.
  • Familiarity with GKE environments from an application operations perspective (not CloudOps or infrastructure engineering).
  • Understanding of cloud constructs relevant to application behavior (networking, IAM, storage, compute).

Good-to-Have Skills

  • CI/CD pipeline exposure (GitHub Actions, Jenkins).
  • Familiarity with GitOps practices.
  • Experience supporting cloud migrations or modernization initiatives.
  • Exposure to platform or infrastructure concepts supporting application workloads.

What We Value

  • Ownership mindset and reliability-first thinking.
  • Curiosity to investigate deep production issues.
  • Strong collaboration with development teams.
  • Bias toward automation and continuous improvement.
  • Clear communication during incidents and stakeholder updates.

Statement to Third Party Agencies
To ALL recruitment agencies: Candescent only accepts resumes from agencies on the preferred supplier list. Please do not forward resumes to our applicant tracking system, Candescent employees, or any Candescent facility. Candescent is not responsible for any fees or charges associated with unsolicited resumes.

Ready to apply to Candescent?
Apply to Candescent

About Candescent

Candescent is the largest non-core digital banking provider. We bring together the transformative technologies that power and connect account opening, digital banking and branch solutions for banks and credit unions of all sizes on any core.

See all jobs at Candescent →

Similar jobs

AES
Senior Reliability Engineer
AES
⚡ Apply early US, Louisville, CO Onsite $113,000–$141,525
● New 👁 Seen ✓ Applied 2h ago
Allison Transmission
Senior Quality & Reliability Engineer
Allison Transmission
⚡ Apply early Indianapolis, IN Onsite
● New 👁 Seen ✓ Applied 5d ago
IEX Group
Senior Systems Reliability Engineer
IEX Group
⚡ Apply early Remote - Must reside in Califo... · location restricted $180,000–$225,000
● New 👁 Seen ✓ Applied 5d ago
ST
Senior Site Reliability Engineer
StarRez
⚡ Apply early Hyderabad, Telengana, IN Onsite
● New 👁 Seen ✓ Applied 3w ago
CI
Senior Site Reliability Engineer - CLO
Circle
⚡ Apply early San Francisco - remote first i... · location restricted $152,500–$205,000
● New 👁 Seen ✓ Applied 3w ago
SA
Senior Site Reliability Engineer
Sanity
⚡ Apply early Remote in the United States or... · location restricted
● New 👁 Seen ✓ Applied 4w ago
Confluent
Staff Software Engineer I - SRE
Confluent
⚡ Apply early IN Remote India · location restricted
● New 👁 Seen ✓ Applied 4w ago
CI
Staff Site Reliability Engineer
Circle
⚡ Apply early San Francisco - remote first i... · location restricted $195,000–$257,500
● New 👁 Seen ✓ Applied 1mo ago
CI
Senior Site Reliability Engineer - Infra Ops
Circle
⚡ Apply early San Francisco - remote first i... · location restricted $152,500–$205,000
● New 👁 Seen ✓ Applied 1mo ago

Sign up for suggestions tailored to the jobs you open and the searches you save.

More jobs at Candescent

See all jobs at Candescent →

Apply now
🤖

Whoa — hold up

JobsRadar was built for real people having a rough time in their job search — not for automated requests. You're clicking way too fast and you're now temporarily blocked.

Come back later. If you're genuinely job hunting, we've got your back — just act like a human.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Get an edge on your job hunt.

Join our Telegram channel for the stuff that helps you land the role — salary benchmarks, the weekly market pulse, and new-feature drops. No spam, just signal.

Join the channel — it's free