Jobs Companies NTT Senior Associate Site Reliability Engineer

About this Senior Associate Site Reliability Engineer role at NTT

NTT · Onsite · hyderabad

Continue to make an impact with a company that is pushing the boundaries of what is possible. At NTT DATA, we are renowned for our technical excellence, leading innovations, and making a difference for our clients and society. Our workplace embraces diversity and inclusion – it’s a place where you can continue to grow, belong, and thrive.

Your career here is about believing in yourself and seizing new opportunities and challenges. It’s about expanding your skills and expertise in your current role and preparing yourself for future advancements. That’s why we encourage you to take every opportunity to further your career within our great global team.

Your day at NTT DATA
The Senior Associate Site Reliability Engineer (SRE) is a developing subject matter expert responsible for playing a key role in ensuring the reliability, availability, and performance of company systems and infrastructure.

This role takes guidance from a supervisor and collaborates with cross functional teams to improve system resiliency and support the development and deployment of highly reliable software application.

The Senior Associate Site Reliability Engineer has opportunities to learn from experienced professionals, gain hands-on experience, and grow their skills in ensuring the reliability and performance of critical systems and infrastructure.

Key responsibilities:
  • Monitors system health, performance metrics, and alerts to identify and respond to incidents promptly.
  • Works with Senior Site Reliability Engineers and teams to diagnose issues, troubleshoot problems, and restore services in a timely manner.
  • Assists in the deployment and release of software applications and infrastructure changes.
  • Collaborates with development teams to ensure smooth deployments, implement best practices, and minimize downtime during releases.
  • Collaborates with senior SREs and operations teams to automate routine tasks and improve operational efficiency.
  • Assists in capacity planning efforts, monitor resource utilization, and make recommendations for scaling infrastructure and services based on projected needs.
  • Collaborates with senior SREs to ensure adequate capacity to meet growing demands.
  • Documents incidents, their impact, and resolution procedures to maintain an incident knowledge base.
  • Participates in post-incident reviews, contributes to root cause analysis, and helps implement preventive measures to minimize future incidents.
  • Collaborates with security teams to implement security best practices and ensure compliance with industry standards and regulations.
  • Assists in monitoring and responding to security incidents, applying appropriate mitigation measures.
  • Works closely with development teams, operations teams, and other stakeholders to ensure smooth
  • Stays updated with the latest industry trends, emerging technologies, and best practices in Site Reliability Engineering.
  • Seeks opportunities to expand technical skills and knowledge through training, certifications, and self-study.
  • Performs any other related task as required.

To thrive in this role, you need to have:
  • Familiarity with infrastructure concepts, including cloud platforms (for example, AWS, Azure, Google Cloud), networking, and system administration.
  • Developing knowledge of programming or scripting languages (such as Python, Bash, or PowerShell) and version control systems (such as Git).
  • Relevant understanding of Linux/Unix systems and experience working with command-line tools.
  • Strong problem-solving and analytical skills, with attention to detail.
  • Excellent communication and collaboration skills, with the ability to work effectively in a team environment.
  • Passion for automation, reliability, and continuous improvement.
  • Familiarity with incident management processes, monitoring tools, and configuration management systems is beneficial.
  • Developing expertise in performance monitoring, optimization, and troubleshooting using tools such as Prometheus, Grafana, or New Relic.
  • Developing ability to optimize system performance, scalability, and reliability. experience with performance monitoring and tuning tools (for example, Prometheus, Grafana, or New Relic) to identify bottlenecks, analyze performance data, and implement optimization strategies.
  • Developing understanding of security principles, best practices, and compliance requirements.

Academic qualifications and certifications:
  • Bachelor's degree or equivalent in Computer Science, Information Technology, or a related field.
  • Relevant certifications, such as AWS Certified DevOps Engineer - Professional, Google Cloud Professional DevOps Engineer, or Certified Kubernetes Administrator (CKA) preferred.

Required experience:
  • Moderate level hands-on experience in a Site Reliability Engineering role or related roles, including experience in designing and maintaining highly available and scalable systems.
  • Moderate level experience in incident response procedures and troubleshooting techniques to identify and resolve system issues.
  • Moderate level experience in automation principles and tools (for example, Terraform, Jenkins, Git).
  • Moderate level experience with scripting languages and version control systems, such as Git, demonstrates an ability to automate tasks and work collaboratively.

Workplace type:

On-site Working

Equal Opportunity Employer
NTT DATA is proud to be an Equal Opportunity Employer with a global culture that embraces diversity. We are committed to providing an environment free of unfair discrimination and harassment. We do not discriminate based on age, race, colour, gender, sexual orientation, religion, nationality, disability, pregnancy, marital status, veteran status, or any other protected category. Accelerate your career with us. Apply today

Ready to apply to NTT?
Apply to NTT

About NTT

Is innovation part of your DNA? Do you want to enable a connected future for people, organizations, and society? Join our growing global NTT family and you’ll be part of the world’s largest ICT company (by revenue). We’ve combined the capabilities of 28 remarkable companies to become one, leading technology services provider. Together, we help our people, clients, and communities do great things with technology to create a more secure and connected future. We employ 40,000 people across 57 countries. By bringing together the world’s best technology companies and emerging innovators, we work together to deliver sustainable outcomes to businesses and the world. Innovation is part of our DNA. W

See all jobs at NTT →

Similar jobs

Sign up for suggestions tailored to the jobs you open and the searches you save.

More jobs at NTT

See all jobs at NTT →

Apply now
🤖

Whoa — hold up

JobsRadar was built for real people having a rough time in their job search — not for automated requests. You're clicking way too fast and you're now temporarily blocked.

Come back later. If you're genuinely job hunting, we've got your back — just act like a human.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Get an edge on your job hunt.

Join our Telegram channel for the stuff that helps you land the role — salary benchmarks, the weekly market pulse, and new-feature drops. No spam, just signal.

Join the channel — it's free