Jobs Companies Razer Site Reliability Engineer

Sobre este puesto de Site Reliability Engineer en Razer

Razer · Presencial · Chengdu

Joining Razer will place you on a global mission to revolutionize the way the world games. Razer is a place to do great work, offering you the opportunity to make an impact globally while working across a global team located across 5 continents. Razer is also a great place to work, providing you the unique, gamer-centric #LifeAtRazer experience that will put you in an accelerated growth, both personally and professionally.

Job Responsibilities/ 工作职责 :

Job Description Summary
We are seeking a skilled and driven Site Reliability Engineer (SRE) to join our growing infrastructure and platform engineering team. The ideal candidate will have hands-on experience in Amazon Web Services (AWS), strong troubleshooting capabilities, and a passion for building scalable, observable, and resilient systems using modern Infrastructure as Code (IaC) and automation tools.


Job Description

REQUIREMENTS:

  • Bachelor’s degree in Computer Science, Software Engineering, Information Technology, or a related field.

  • Minimum 3 years of experience in SRE, DevOps, cloud infrastructure, or system administration roles.

  • Hands-on expertise with AWS Cloud Services, including:

  • Compute & Containerization: EC2, Lambda, ECS, EKS, Auto Scaling

  • Networking: Load Balancers, VPC, Route 53, Security Groups, Firewalls

  • Storage & Databases: RDS, ElastiCache, Athena, S3

  • Messaging: SQS, SES

  • Deep understanding of Infrastructure as Code (IaC) tools such as Terraform and CloudFormation.

  • Proficiency in at least one programming/scripting language: Python, Node.js, Bash, Ruby, or related.

  • Experience operating and troubleshooting across Linux, Windows, and container-based environments.

  • Strong understanding of distributed systems, cloud networking (routers, switches), firewalls, DNS, and HTTP/TLS.

  • Experience implementing monitoring and alerting systems and working with incident management processes.

  • Experience with Zero Downtime Deployments, blue/green or canary deployments.

  • Familiarity with cost optimization and right-sizing AWS resources.

  • Exposure to multi-region, multi-account AWS architecture.

  • Understanding of API gateway, or edge networking (e.g., Akamai, CloudFront).

JOB DESCRIPTION:

  • Design, develop, and maintain Infrastructure as Code (IaC) using tools like Terraform or AWS CloudFormation, leveraging AI coding assistants to accelerate development and enforce best practices.

  • Implement and operate reliable, scalable cloud infrastructure primarily on AWS (e.g., EC2, ECS, RDS, S3, Lambda, ElastiCache, SQS, SES, Auto Scaling, Load Balancers)

  • Lead and participate in architecture reviews focusing on reliability, scalability, security, performance, and the cost-efficiency of infrastructure.

  • Develop and manage robust monitoring, alerting, and logging solutions (e.g., CloudWatch, Prometheus, Grafana, ELK), incorporating AIOps tools for predictive alerting, anomaly detection, and reducing alert fatigue.

  • Perform incident management, postmortems, root cause analysis, and implement continuous improvement strategies, utilizing AI-driven analytics to rapidly summarize logs and traces during outages.

  • Collaborate with software engineering teams to improve CI/CD pipelines, deployment automation, release management, and the deployment lifecycles of machine learning models.

  • Automate infrastructure operations, reduce manual toil, and improve reliability using scripting (Python, Bash, Node.js, or Ruby) and AI-powered workflow automation.

  • Maintain and troubleshoot environments involving web servers, databases, firewalls, DNS, load balancers, networking.

  • Ensure systems are compliant with security standards, including patching, hardening, secure access policies, and data privacy constraints specific to AI training data.

  • Provide on-call support, participate in incident rotations.

  • Monitor and maintain service-level objectives (SLOs), SLAs, and error budgets to ensure reliability targets are met.

  • Provide support and solution handling to incidents and tickets assigned.

Pre-Requisites/ 任职要求 :

Razer is proud to be an Equal Opportunity Employer. We believe that diverse teams drive better ideas, better products, and a stronger culture. We are committed to providing an inclusive, respectful, and fair workplace for every employee across all the countries we operate in. We do not discriminate on the basis of race, ethnicity, colour, nationality, ancestry, religion, age, sex, sexual orientation, gender identity or expression, disability, marital status, or any other characteristic protected under local laws. Where needed, we provide reasonable accommodations - including for disability or religious practices - to ensure every team member can perform and contribute at their best.

Are you game?

¿Listo para postularte en Razer?
Postúlate en Razer

Sobre Razer

At Razer, you’ll be at the forefront of the most exciting industry in the world: gaming. As gaming evolves, so does the ecosystem that powers it: hardware, software and services. Guided by our mission “ For Gamers. By Gamers .”, we create cutting-edge products and experiences that define the ultimate gameplay. Staying true to our mission, we’re relentlessly pushing boundaries and leading the charge in AI for gaming, shaping the future of the industry. At Razer, you won’t just witness this evolution; you’ll help drive it. Joining Razer means being part of a global mission to bring gamers closer to the games they love. Whether you’re crafting the next generation of gaming gear or powering our

Ver todos los empleos en Razer →

Empleos similares

Intel
Substrates Quality and Reliability Engineer
Intel
⚡ Postúlate pronto Malaysia, Penang Híbrido
● Nuevo 👁 Visto ✓ Postulado hace 1 mes
NTT
Site Reliability Engineer
NTT
⚡ Postúlate pronto hyderabad Presencial
● Nuevo 👁 Visto ✓ Postulado hace 9h
NTT
SRE Architect
NTT
⚡ Postúlate pronto hyderabad Presencial
● Nuevo 👁 Visto ✓ Postulado hace 9h
Xcel Energy
Reliability Engineer Intern- TX
Xcel Energy
⚡ Postúlate pronto Amarillo, TX, 79108 Híbrido $44,096–$48,256
● Nuevo 👁 Visto ✓ Postulado hace 9h
Xcel Energy
Performance Optimization Reliability Engineer Intern- TX
Xcel Energy
⚡ Postúlate pronto Earth, TX, 79031 Híbrido $44,096–$48,048
● Nuevo 👁 Visto ✓ Postulado hace 9h
Omnissa
Senior Site Reliability Engineer
Omnissa
⚡ Postúlate pronto Bengaluru, India Presencial
● Nuevo 👁 Visto ✓ Postulado hace 9h
Alcoa
Senior Electrical Reliability Engineer
Alcoa
⚡ Postúlate pronto AU PTL Portland Presencial
● Nuevo 👁 Visto ✓ Postulado hace 9h
NTT
Senior Associate Site Reliability Engineer
NTT
⚡ Postúlate pronto hyderabad Presencial
● Nuevo 👁 Visto ✓ Postulado hace 9h
Broadridge
Platform Site Reliability Engineer (SRE)
Broadridge
⚡ Postúlate pronto Manila - 6805 Ayala Ave Presencial
● Nuevo 👁 Visto ✓ Postulado hace 9h

Regístrate para recibir sugerencias adaptadas a los empleos que abres y las búsquedas que guardas.

Más empleos en Razer

Ver todos los empleos en Razer →

Postúlate ahora
🤖

Un momento — para

JobsRadar se creó para personas reales que están pasando un mal momento en su búsqueda de empleo — no para solicitudes automatizadas. Estás haciendo clic demasiado rápido y ahora estás bloqueado temporalmente.

Vuelve más tarde. Si de verdad estás buscando empleo, cuentas con nosotros — solo compórtate como una persona.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Toma ventaja en tu búsqueda de empleo.

Únete a nuestro canal de Telegram para lo que te ayuda a conseguir el puesto — referencias salariales, el pulso semanal del mercado y avisos de nuevas funciones. Sin spam, solo señal.

Únete al canal — es gratis