Jobs › Companies › Altera › AIOPs Observability/SRE Lead

Über diese AIOPs Observability/SRE Lead Stelle bei Altera

Altera · Vor Ort · San Jose, California, United States

Job Details:

Job Description:

Onsite Requirement: This position requires regular in-office work and is an onsite role based in San Jose, CA. Candidates must be able to work onsite in San Jose, CA.


About Altera

At Altera™, our independence as the world’s largest pure-play FPGA solutions provider gives us the focus, speed, and agility to innovate without compromise. With more than four decades of industry-leading FPGA expertise, our singular mission is to deliver high-performance, flexible FPGA solutions that enable customers to solve their most complex computing challenges.


As Altera continues to evolve and scale as an independent company, our IT organization is transforming its global infrastructure to support the needs of the business, engineering organizations, laboratories, and employees around the world.


About the Role

We are seeking a Site Reliability Engineering Manager to lead the SRE function. This role builds and manages a team of reliability engineers who drive platform stability, observability, automation, and operational excellence across the semiconductor company's critical IT and engineering systems.​


In this role, you will define and implement scalable, secure, and high-performance monitoring architectures that support both current and future business requirements. You will work closely with IT, Engineering, Integration teams, security teams, and external service providers to modernize network infrastructure, enable hybrid cloud connectivity, and ensure a smooth transition from existing environments to the future-state network.


The ideal candidate brings deep expertise in enterprise SRE architecture and transformation, strong hands-on knowledge of routing and switching technologies, and experience integrating cloud environments such as AWS and Azure with large-scale on-premises infrastructure.


Key Responsibilities

  • Lead and grow the SRE team including hiring, mentoring, and developing reliability engineering capabilities.​
  • Define SRE practice including SLIs, SLOs, error budgets, and reliability targets across critical platforms.​
  • Drive automation initiatives to eliminate toil and improve platform reliability and scalability.​
  • Oversee incident management, blameless post-mortems, and systematic reliability improvement programs.​
  • Collaborate with cloud, infrastructure, and application teams to embed reliability into platform design.​
  • Establish observability standards including unified monitoring, alerting, logging, and tracing strategies.​
  • Manage on-call processes, escalation procedures, and team wellbeing for 24x7 operations.​
  • Report on platform reliability, SLO compliance, and operational maturity to IT leadership.​
  • Collaborate with Integration teams, IT infrastructure teams, cybersecurity teams, and external service providers to design and implement reliable connectivity solutions.
  • Evaluate network technologies and solutions and provide technical recommendations based on business requirements, scalability, performance, security, and cost.
  • Ensure network transformation initiatives align with applicable security, regulatory, and compliance requirements.
  • Partner with cybersecurity teams to incorporate appropriate security controls, segmentation, firewall policies, VPN connectivity, and access controls into network architecture.
  • Develop and maintain comprehensive documentation for AIOps architectures, topologies, configurations, standards, policies, migration plans, and operational procedures.
  • Establish architecture standards, design principles, and best practices that promote consistency, scalability, reliability, and operational efficiency.
  • Provide technical leadership and guidance to Observability engineering and operations teams throughout architecture, implementation, migration, and optimization activities.
  • Identify opportunities for continuous improvement, automation, standardization, and modernization of network operations.
  • Troubleshoot and provide architectural guidance for complex Observability services
  • Stay current with emerging enterprise SRE, cloud networking, automation, and security technologies and assess their applicability to Altera’s environment.

Salary Range

The pay range below is for Bay Area California only. Actual salary may vary based on a number of factors including job location, job-related knowledge, skills, experiences, trainings, etc. We also offer incentive opportunities that reward employees based on individual and company performance.


$187,000 - $270,700 USD


We use artificial intelligence to screen, assess, or select applicants for the position. Applicants must be eligible for any required U.S. export authorizations.


#LI-MD1

Qualifications:

Minimum Qualifications

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related technical field, with 10+ years of professional experience in enterprise networking.
  • 10+ years of experience in site reliability engineering or DevOps with a focus on platform reliability.​
  • Proven experience leading technical engineering teams in a senior or management capacity.​
  • Deep understanding of SRE principles including SLIs, SLOs, error budgets, and reliability frameworks.​
  • Strong background in observability, incident management, and blameless post-mortem culture.​
  • Experience driving automation and toil reduction programs across infrastructure and platform teams.​
  • Knowledge of cloud platforms (Azure, AWS) and containerized environments.​
  • Excellent communication skills for stakeholder reporting and cross-functional collaboration.​

Preferred Qualifications

  • Experience establishing SRE practices in semiconductor, HPC, or EDA-dependent environments.​
  • Familiarity with AIOps platforms and AI-assisted incident management tooling.​
  • Knowledge of chaos engineering practices and resiliency testing frameworks.​

Job Type:

Regular

Shift:

Shift 1 (United States of America)

Primary Location:

San Jose, California, United States

Additional Locations:

Posting Statement:

All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance.
Bereit, sich bei Altera zu bewerben?
Bei Altera bewerben

Über Altera

About Altera Altera: Accelerating Innovators Altera provides leadership programmable solutions that are easy-to-use and deploy in applications from cloud to edge, offering limitless AI possibilities. Our end-to-end broad portfolio of products including FPGAs, CPLDs, Intellectual Property, development tools, System on Modules, SmartNICs and IPUs provide the flexibility to accelerate innovation. Altera is helping to shape the future through pioneering innovation that unlocks extraordinary possibilities for everyone on the planet.

Alle Jobs bei Altera ansehen →

Ähnliche Jobs

Altera
Digital Design Verification Engineer
Altera
⚡ Früh bewerben Iași, Romania (Remote) · standortgebunden
● Neu 👁 Gesehen ✓ Beworben vor 1 Tg.
Altera
AI Engineer – Cloud Software & AI Platforms - Contract
Altera
⚡ Früh bewerben San Jose, California, United S... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 2 Tg.
Altera
Senior M365 Engineer - SharePoint, Teams, OneDrive​
Altera
⚡ Früh bewerben San Jose, California, United S... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 2 Tg.
Altera
Senior Debug Verification Engineer
Altera
⚡ Früh bewerben San Jose, California, United S... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 2 Tg.
Altera
Software Engineer - Intern
Altera
⚡ Früh bewerben Toronto, Ontario, Canada Vor Ort CA$95,000–CA$100,000
● Neu 👁 Gesehen ✓ Beworben vor 2 Tg.
Altera
Senior SOC Timing Engineer
Altera
⚡ Früh bewerben Penang 15, Penang, Malaysia Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 2 Tg.
Altera
Senior Director, Silicon Architecture and Design Engineering (PCIe)
Altera
⚡ Früh bewerben San Jose, California, United S... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 2 Tg.
Altera
Workday Functional Solutions Architect
Altera
⚡ Früh bewerben San Jose, California, United S... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 2 Tg.
Altera
Senior Endpoint Engineer - Azure Virtual Desktop
Altera
⚡ Früh bewerben San Jose, California, United S... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 2 Tg.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Altera

Alle Jobs bei Altera ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos