Cloud Production Systems Engineer
Job in
Rayville, Richland Parish, Louisiana, 71269, USA
Listing for:
Meta
Full Time
position
Listed on 2026-10-05
Job specializations:
-
IT/Tech
Cloud Computing: Infrastructure & Operations, Systems Engineer
Salary/Wage Range or Industry Benchmark: 173000 - 245000 USD Yearly
USD
173000.00
245000.00
YEAR
Job Description & How to Apply Below
Location: RayvilleSummary:Meta is seeking an experienced Cloud Production Systems Engineer to join the Data Center Operations team. Our data centers and the cloud infrastructure powering tens of thousands of servers forms the foundation upon which Meta's rapidly scaling services operate. Meta is at the leading edge of the global data center industry in both design and cloud operations. This role requires a forward-thinking systems professional with deep experience leveraging diverse software tools to identify cloud automation solutions for complex operational challenges with our Cloud partnerships (AWS, GCP, Core Weave, and other Neo Cloud).
The ideal candidate performs deep data analysis to prioritize server repair automation in a hyperscale cloud environment, drives solutions through code, and collaborates effectively with globally distributed teams through clear written communication.
Required Skills:Cloud Production Systems Engineer Responsibilities:
Identify and root cause systemic issues across the cloud server fleet and drive resolutions to maximize uptime and utilization by leveraging hardware failure data and diagnostic telemetryWrite, review, and maintain code for diagnostic and cloud automation tooling that supports quality and efficient delivery of production servers at hyperscaleOwn and develop diagnostic tooling requirements that enable frontline operations teams to efficiently manage and repair the cloud server fleetDrive the escalation process for Data Center Operations to identify, root cause, and resolve complex cloud tooling and hardware issues affecting fleet healthExecute operational validation and verification activities for new cloud product integration into the production environmentCollaborate with cross-functional cloud tooling teams to provide an operations-centric perspective on open issues and contribute to their development roadmapsPerform deep data analysis to prioritize cloud automation opportunities for server repair workflows in a large-scale, heterogeneous hardware environmentBuild cross-functional relationships and influence policies and procedures to improve global cloud data center operations consistency and efficiencyMentor other engineers on evaluating and resolving cloud fleet issues and defining improvements to tools and operational processesTravel up to 25% to support global cloud data center operations and new site deploymentsMinimum Qualifications:Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience7+ years of experience in production systems engineering, infrastructure engineering, or system software development for large-scale cloud environments7+ years of experience with hardware lifecycle management, fleet automation, or cloud data center operations systems spanning compute, storage, or networking infrastructureExperience developing systems software or automation tooling in Python, Bash, PHP, C, or C++ for Linux-based production environments at scaleExperience with the configuration and maintenance of production systems, including web servers, load balancers, relational databases, storage systems, and messaging systemsExperience communicating technical designs and cloud infrastructure decisions through written documentation and cross-functional stakeholder alignment across engineering and operations teamsPreferred Qualifications:Experience designing or operating cloud configuration management and infrastructure-as-code systems for large heterogeneous hardware fleetsExperience with data analysis and visualization tools used to prioritize fleet health initiatives and drive operational decision‑makingExperience supporting global, multi-site data center infrastructure deployments, including hardware qualification and regional rollout coordinationFamiliarity with distributed systems monitoring, alerting, and automated remediation pipelines at hyperscalePublic Compensation:
$173,000/year to $245,000/year + bonus + equity + benefits
Industry:
Internet
Equal Opportunity:
Meta is proud to be an Equal Employment Opportunity and Affiant employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender, gender identity, gender expression, transgender status, sexual stereotypes, age, status as a protected veteran, status as an individual with a disability, or other applicable legally protected characteristics.
We also…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here: