×
Register Here to Apply for Jobs or Post Jobs. X

AIOPs Observability​/SRE Lead

Job in San Jose, Santa Clara County, California, 95199, USA
Listing for: Altera Corporation
Full Time position
Listed on 2026-10-02
Job specializations:
  • IT/Tech
    Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability, Cybersecurity
Salary/Wage Range or Industry Benchmark: 187000 - 270700 USD Yearly USD 187000.00 270700.00 YEAR
Job Description & How to Apply Below
## AIOPs Observability/SRE Lead Apply:
San Jose, California, United States:
Full time:
Posted Yesterday:
R03209#
** Job Details:**### ##
*
* Job Description:

**** Onsite Requirement:
** This position requires regular in-office work and is an onsite role based in San Jose, CA. Candidates must be able to work onsite in San Jose, CA.  
** About Altera
** At AlteraTM, our independence as the world’s largest pure-play FPGA solutions provider gives us the focus, speed, and agility to innovate without compromise. With more than four decades of industry-leading FPGA expertise, our singular mission is to deliver high-performance, flexible FPGA solutions that enable customers to solve their most complex computing challenges.  As Altera continues to evolve and scale as an independent company, our IT organization is transforming its global infrastructure to support the needs of the business, engineering organizations, laboratories, and employees around the world.

** About the Role
** We are seeking a Site Reliability Engineering Manager to lead the SRE function. This role builds and manages a team of reliability engineers who drive platform stability, observability, automation, and operational excellence across the semiconductor company's critical IT and engineering systems.  In this role, you will define and implement scalable, secure, and high-performance monitoring architectures that support both current and future business requirements.

You will work closely with IT, Engineering, Integration teams, security teams, and external service providers to modernize network infrastructure, enable hybrid cloud connectivity, and ensure a smooth transition from existing environments to the future-state network.  The ideal candidate brings deep expertise in enterprise SRE architecture and transformation, strong hands-on knowledge of routing and switching technologies, and experience integrating cloud environments such as AWS and Azure with large-scale on-premises infrastructure.

** Key Responsibilities
*** Lead and grow the SRE team including hiring, mentoring, and developing reliability engineering capabilities.
* Define SRE practice including SLIs, SLOs, error budgets, and reliability targets across critical platforms.
* Drive automation initiatives to eliminate toil and improve platform reliability and scalability.
* Oversee incident management, blameless post-mortems, and systematic reliability improvement programs.
* Collaborate with cloud, infrastructure, and application teams to embed reliability into platform design.
* Establish observability standards including unified monitoring, alerting, logging, and tracing strategies.
* Manage on-call processes, escalation procedures, and team wellbeing for 24x7 operations.
* Report on platform reliability, SLO compliance, and operational maturity to IT leadership.
* Collaborate with Integration teams, IT infrastructure teams, cybersecurity teams, and external service providers to design and implement reliable connectivity solutions.
* Evaluate network technologies and solutions and provide technical recommendations based on business requirements, scalability, performance, security, and cost.
* Ensure network transformation initiatives align with applicable security, regulatory, and compliance requirements.
* Partner with cybersecurity teams to incorporate appropriate security controls, segmentation, firewall policies, VPN connectivity, and access controls into network architecture.
* Develop and maintain comprehensive documentation for AIOps architectures, topologies, configurations, standards, policies, migration plans, and operational procedures.
* Establish architecture standards, design principles, and best practices that promote consistency, scalability, reliability, and…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary