Sobre esta vaga de Senior Site Reliability Engineer na Omnissa
Job Description:
Job Description
Who are we?
Omnissa Workspace ONE UEM platform allows companies to make it easy for employees to work anywhere, any time, on any device, without compromising security. But making things easy for our customers is extremely challenging for us, so we are looking for unique thinkers of varying backgrounds that want to take on complex, highly technical, customer-impacting challenges. We, the UEM team, have a large customer base that spans across industry verticals including a majority of the Fortune 500 companies. That means that the work you do here has a broad, measurable impact on the businesses and communities that these customers represent.
We also want to help you have a positive impact on the community and the environment, whether through the Power of Differences (PODs) groups, various environmental improvement initiatives, matching donations to qualified non-profits, or time off for volunteering. And we want to have fun along the way, both while you are making a difference for our customers and your community.
Team Responsibilities
The SaaS Operations Team (SaaS Ops) drives the Workspace ONE UEM SaaS platform deployments, configuration, and support across global datacenters. We are often involved in the leading-edge use cases of SaaS transformation for our customers and work towards automated solutions for managing the infrastructure layer and the application layer that comprise our SaaS platform.
Role Responsibilities:
We need someone with operations team background and a DevOps mindset in solving complex problems, who maintains a high degree of ownership. The core team is currently made up of engineers across the US, UK, and India. As a SaaS Ops Engineer, you will work with other global departments to provide an overall 99.99% availability of enterprise and production services through day-to-day operations, support, maintenance and upkeep of the core application services and associated services like SQL Services. We use a wide range of cloud platforms and technologies such as Ansible, Terraform, Packer, AWS, VMware on AWS, Docker, Elasticsearch, StackStorm, etc.
The responsibilities will include but not be limited to:
- Deploy and maintain production environments that requires high availability
- Perform proactive troubleshooting & performance analysis of internal services and cloud environments
- Drive the product towards higher availability and reliability & assist with on-call support on a rotating schedule for incident escalations (24x7)
- Ensure our services meet stability, performance and availability requirements.
- Monitor usage, capacity, and performance of servers; liaise with users and/or vendors to address problems and changes in requirements.
- Build robust, self-healing features and automation that reduce operational effort and improve service up-time.
- Ability to work with global R&D, SRE and Cloud Services Teams to help design and develop deployment automation procedures for new cloud service offerings
- Working under pressure in production environments running production customer workloads and services
- Self-starter mindset with a strong drive to learn and own engineering initiatives to promote a culture of continuous improvement, and engineering excellence.
Required Skills
- 7+ years of experience in a Development, Operations, Site Reliability, Systems or comparable Cloud Engineering position.
- 7+ years of experience with Unix/Linux/ Windows Operating system
- Experience in one of the following Scripting languages: Python, PowerShell, Bash, Shell Script.
- Proficiency with Ansible or a knowledge of similar configuration management tools like Chef, Puppet etc.
- Experience with Jenkins or similar build automation tool.
- Experience with monitoring and logging services (e.g. Elasticsearch, Wavefront, Uptime, or similar)
- Experience with Configuring and troubleshooting Web applications to ensure SLA compliance
- Excellent written and verbal communication skills
Preferred Skills
- Experience with Database Administration Activities and troubleshooting especially with Microsoft SQL Server
- Experience with cloud-based infrastructure and services such as vSphere products, AWS and network software like BIG-IP, DYN, Route53 etc.