Data Center Facility Operations Reliability Engineer — Opportunihub
Meta logo
Job Remote

Data Center Facility Operations Reliability Engineer

Meta · USA

At a glance

Type
Job
Organisation
Meta
Location
USA
Work mode
Remote
Deadline
Rolling / not stated
Posted
29 Jul 2026

About this job

<p>Meta is seeking an experienced and self-motivated Reliability Engineer to join our Asset Management &amp; Reliability team within Facility Operations. This person will work within Facility Operations to identify and manage asset reliability risks and various stages of end-to-end asset lifecycle for the Data Center Operations. Managing stakeholders spread across time zones and key to the success of our individual projects and overall asset management, and reliability program. In this role, you will be expected to own and execute within your focus areas, contribute to scoping and roadmapping for reliability initiatives, and solve complex technical problems to ensure seamless operations across new and existing sites.ResponsibilitiesLead operational Failure Mode and Effects Analysis (FMEA) and establish maintenance strategies.* Manage and govern content for the Global Maintenance Library, ensuring rigorous Maintenance Change Management processes are followed.* Drive Corrective Maintenance initiatives through deep Failure Mode Analysis to identify root causes and implement preventive measures.* Evaluate regulatory and compliance requirements, providing governance and ensuring adherence across all reliability activities.* Support and execute Condition-Based Monitoring and Diagnostics programs, including testing, vibration analysis, infrared (IR), partial discharge (PD), and robotics applications.* Assess and optimize Mean Time To Repair (MTTR) and Mean Time Between Failures (MTBF) metrics to drive continuous improvement.* Contribute to continuous production risk models to prioritize resources and mitigate operational vulnerabilities.* Manage non-critical scope and support supplier scope management to ensure external partners meet reliability and quality standards.* Implement and optimize sensor-based condition monitoring and maintenance (e.g., infrared, temperature, pressure, flow, leak detection) to remotely monitor and assign maintenance tasks.* Support technical insourcing versus outsourcing decision-making for maintenance and reliability activities.* Establish Service Bill of Materials (BOM) setups and drive initial and continuous spares selection processes.* Develop and execute a comprehensive whole-equipment sparing strategy to minimize downtime.* Identify First-of-Kind (FOK) parts and establish Parts Factory Base Part Number (FBPN) setups.* Conduct Form-Fit-Function technical analysis to evaluate and approve replacement components.* Act as the primary liaison to IBOS and ISCE teams for managing replacement parts backlog, timeliness, tooling requirements, and auditing.* Identify reliability opportunities and contribute to engineering excellence within the team.QualificationsBachelor's degree in Mechanical, Electrical, Reliability Engineering, or a similar technical discipline or 6+ years relevant industry experience will be considered in lieu of 8+ years relevant industry experience* 8+ years of experience in reliability engineering (related to electrical or mechanical cooling equipment) or related Engineering roles* Experienced in Reliability Centered Maintenance (RCM) and Failure Mode and Effects Analysis (FMEA) activities for maintenance, process, and equipment design optimization to meet reliability requirements* Proven ability to execute independently within defined scope and manage complex problems with guidance on ambiguous areas* Proficient in the usage of Enterprise Asset Management (EAM) solutions to extract data and develop meaningful insights* Knowledgeable of relevant ISO standards (ISO 14224, ISO 17359, ISO 55000)* Experience with project management and cross-functional collaboration* Ability to travel 25% domestically and internationally Experience with data center equipment such as critical cooling systems, generators, main switchboards, and network gear* Proficient in data analysis techniques that can include Process Control, Reliability modeling and prediction, Fault Tree Analysis, Weibull Tree Analysis, and Six Sigma (6σ) Methodology* Proficient in developing and executing comprehensive test plans for assets* Experience supporting technical initiatives and advocating for high-quality engineering standards* Certifications in Maintenance &amp; Reliability such as CMRP, CRL, or CRE</p>

How to apply

  1. 1 Read the full details above and confirm you meet the eligibility criteria.
  2. 2 Prepare your documents — an updated CV, and any cover letter, proposal or certificates required.
  3. 3 Click Apply on official site to complete your application on Meta’s official page.
  4. 4 Submit as early as possible — many close once filled.
Apply on official site

Sourced from jobicy. Always verify details on the official website. Opportunihub never charges you to apply.

Frequently asked questions

How do I apply for Data Center Facility Operations Reliability Engineer?

Review the full details and eligibility on this page, prepare your documents, then use the “Apply on official site” button to complete your application on Meta’s official page.

Is this opportunity remote or location-based?

This opportunity is remote-friendly and open to applicants who can work from anywhere.

Is Data Center Facility Operations Reliability Engineer free to apply for?

Opportunihub lists this Job for free. Legitimate Jobs do not ask for payment to apply — never pay a fee to submit an application.