Jobs โ€บ Companies โ€บ Weekday AI โ€บ Linux SME / SRE Engineer

About this Linux SME / SRE Engineer role at Weekday AI

Weekday AI ยท Onsite ยท Bengaluru, Karnataka, India

๐—ง๐—ต๐—ถ๐˜€ ๐—ฟ๐—ผ๐—น๐—ฒ ๐—ถ๐˜€ ๐—ณ๐—ผ๐—ฟ ๐—ผ๐—ป๐—ฒ ๐—ผ๐—ณ ๐˜๐—ต๐—ฒ ๐—ช๐—ฒ๐—ฒ๐—ธ๐—ฑ๐—ฎ๐˜†'๐˜€ ๐—ฐ๐—น๐—ถ๐—ฒ๐—ป๐˜๐˜€

๐—ฆ๐—ฎ๐—น๐—ฎ๐—ฟ๐˜† ๐—ฟ๐—ฎ๐—ป๐—ด๐—ฒ: ๐—ฅ๐˜€ ๐Ÿฎ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ - ๐—ฅ๐˜€ ๐Ÿฏ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ (๐—ถ๐—ฒ ๐—œ๐—ก๐—ฅ ๐Ÿฎ๐Ÿฌ-๐Ÿฏ๐Ÿฌ ๐—Ÿ๐—ฃ๐—”)

Experience: 5+ yrs

Location: Bengaluru, Karnataka, India, Hyderabad, Telangana, India

Job Type: Full-time

We are looking for an experiencedย Linux SME / SRE Engineerย with strong expertise inย Core Linux Administration, RHEL, and PCS/Pacemaker clusteringย to support business-critical production environments.

The role focuses on maintaining highly available Linux infrastructure, resolving complex production issues, ensuring system reliability, and supporting clustered environments. The ideal candidate will have strong hands-on troubleshooting capabilities, a solid understanding of high-availability architectures, and the ability to work effectively with clients and technical stakeholders.

Requirements

Key Responsibilities

  • Administer and supportย Linux-based production environmentsย across critical infrastructure.
  • Perform day-to-dayย Core Linux administration, configuration, monitoring, maintenance, and troubleshooting.
  • Manage, monitor, configure, and troubleshootย PCS/Pacemaker high-availability clusters.
  • Ensure availability, reliability, stability, and performance of Linux infrastructure and clustered services.
  • Troubleshoot complex and critical production incidents and drive issues through to resolution.
  • Perform root-cause analysis and implement sustainable solutions for recurring infrastructure problems.
  • Monitor system and cluster health and proactively identify potential availability or performance issues.
  • Support failover, recovery, maintenance, and operational activities across high-availability environments.
  • Collaborate with clients, infrastructure teams, application teams, and other technical stakeholders on incidents and enhancements.
  • Participate in incident management, problem management, change management, and production maintenance activities.
  • Follow SRE practices for monitoring, reliability improvement, incident response, and operational efficiency.
  • Maintain technical documentation, operational procedures, troubleshooting guides, and support records.
  • Participate in rotational shifts to provide continuous production support.
  • Identify opportunities to automate repetitive infrastructure tasks and improve operational efficiency.
  • Support infrastructure changes, upgrades, patching, and maintenance activities in accordance with established processes.
  • Contribute to service reliability, availability, and continuous improvement initiatives.

What Makes You a Great Fit

  • 5โ€“9 years of overall experienceย in Linux administration, infrastructure engineering, SRE, or production support, with a maximum of 10 years preferred.
  • Minimumย 4 years of hands-on experience with PCS/Pacemaker cluster administration.
  • Strong expertise inย Core Linux Administrationย and production infrastructure support.
  • Strong hands-on experience withย RHEL (Red Hat Enterprise Linux).
  • Solid understanding ofย High Availability, clustering, failover, resource management, and cluster troubleshooting.
  • Proven experience supportingย critical production environmentsย with strict availability and reliability requirements.
  • Strong troubleshooting, debugging, root-cause analysis, and incident-resolution capabilities.
  • Experience working with production monitoring, incident management, and infrastructure maintenance processes.
  • Strong understanding ofย SRE and ITIL practicesย is desirable.
  • Excellent communication and client-facing skills with the ability to explain technical issues clearly to stakeholders.
  • Strong stakeholder-management and collaboration skills.
  • Ability to work effectively under pressure during critical production incidents.
  • Willingness to work inย rotational shifts, including scheduled production-support coverage.
  • Experience withย VMware administrationย is an advantage.
  • Exposure toย AWS or other cloud platformsย is desirable.
  • Knowledge ofย Oracle Databaseย and its infrastructure dependencies is an advantage.
  • Strong ownership mindset with a focus on system reliability, operational excellence, and continuous improvement.
Ready to apply to Weekday AI?
Apply to Weekday AI

About Weekday AI

At Weekday (backed by YC; also Product Hunt #1 product of the day), we are building the next frontier in hiring. We have built the largest database of white collar talent in India and have built outreach tools on top of it to generate highest response rates.

See all jobs at Weekday AI โ†’

Similar jobs

Fivetran
Staff Site Reliability Engineer
Fivetran
โšก Apply early Bengaluru, Karnataka, India, A... Onsite
โ— New ๐Ÿ‘ Seen โœ“ Applied 3h ago
Levi Strauss & Co.
Engineer III, Systems Reliability Engineering
Levi Strauss & Co.
โšก Apply early GCC Office โ€“ ITC Green Center,... Onsite
โ— New ๐Ÿ‘ Seen โœ“ Applied 12h ago
Levi Strauss & Co.
Engineer III, Systems Reliability Engineering
Levi Strauss & Co.
โšก Apply early GCC Office โ€“ ITC Green Center,... Onsite
โ— New ๐Ÿ‘ Seen โœ“ Applied 12h ago
Levi Strauss & Co.
Engineer III, Systems Reliability Engineering
Levi Strauss & Co.
โšก Apply early GCC Office โ€“ ITC Green Center,... Onsite
โ— New ๐Ÿ‘ Seen โœ“ Applied 12h ago
Netskope
Sr. Site Reliability Engineer, Engineering Stack Support
Netskope
โšก Apply early Bengaluru, Karnataka, India Onsite
โ— New ๐Ÿ‘ Seen โœ“ Applied 5d ago
Sabre
Site Reliability Engineer IV
Sabre
โšก Apply early Bengaluru, Karnataka, India Onsite
โ— New ๐Ÿ‘ Seen โœ“ Applied 1w ago
Weekday AI
Lead Site Reliability Engineer or Platform Engineer
Weekday AI
โšก Apply early Bengaluru, Karnataka, India Onsite
โ— New ๐Ÿ‘ Seen โœ“ Applied 1w ago
GSSTech Group
Senior DevOps / SRE Engineer
GSSTech Group
โšก Apply early Bengaluru, Karnataka, India Onsite
โ— New ๐Ÿ‘ Seen โœ“ Applied 2w ago
ServiceTitan
Staff Site Reliability Engineer
ServiceTitan
โšก Apply early India Bengaluru, Karnataka Onsite
โ— New ๐Ÿ‘ Seen โœ“ Applied 1mo ago

Sign up for suggestions tailored to the jobs you open and the searches you save.

More jobs at Weekday AI

See all jobs at Weekday AI โ†’

Apply now
๐Ÿค–

Whoa โ€” hold up

JobsRadar was built for real people having a rough time in their job search โ€” not for automated requests. You're clicking way too fast and you're now temporarily blocked.

Come back later. If you're genuinely job hunting, we've got your back โ€” just act like a human.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Get an edge on your job hunt.

Join our Telegram channel for the stuff that helps you land the role โ€” salary benchmarks, the weekly market pulse, and new-feature drops. No spam, just signal.

Join the channel โ€” it's free