[All remote jobs](https://jobicy.com/jobs.md)Open role[![Meta logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/750f86a9-221.jpeg)](https://jobicy.com/company/meta.md)Remote opportunity at[Meta](https://jobicy.com/company/meta.md)

# Data Center Facility Operations Reliability Engineer

Review the role, location requirements, compensation details, and application process before deciding whether this opportunity fits your next career move.

[Apply for this job](#job-application)[View company](https://jobicy.com/company/meta.md)Share29 Jul 2026Published44Listing views3Application actionsManually reviewedTrust & Safety status  Opportunity details

## About this role.

AI SummaryThis is a senior Reliability Engineer role at Meta, focused on asset management and reliability within data center facility operations. The position involves leading FMEA, developing maintenance strategies, driving corrective maintenance through root cause analysis, and implementing condition-based monitoring programs. The ideal candidate has 8+ years of experience in reliability engineering, expertise in RCM and FMEA, and proficiency with EAM solutions. This role offers an opportunity to work on critical infrastructure and collaborate across global teams.

## Role DNA

A quick view of the complexity, pace, ownership and collaboration implied by the job description.

### Job Complexity

4/5EasyHard

### Pace & Pressure

4/5RelaxedFast-paced

### Autonomy Level

4/5GuidedFull ownership

### Communication Load

5/5IndependentCollaborative

AI insightThe role requires deep technical expertise in reliability engineering, extensive experience, and the ability to solve complex problems independently; however, it is not at the executive level, hence a difficulty of 4.

## Salary analysis

Estimated compensation compared with the broader US market for similar roles.

Estimated job medianHighly competitive$150,000US market range$100k–$200k0$220k

AI insightThe salary is estimated based on market data for senior reliability engineers in data center operations. The median of $150,000 is competitive for the role, with potential for higher earnings based on experience and location.

## Core skills

Skills and capabilities most closely associated with this opportunity.

[Reliability Engineering](https://jobicy.com/jobs?search_keywords=Reliability%20Engineering.md)[FMEA](https://jobicy.com/jobs?search_keywords=FMEA.md)[RCM](https://jobicy.com/jobs?search_keywords=RCM.md)[Data Center Operations](https://jobicy.com/jobs?search_keywords=Data%20Center%20Operations.md)[Condition-Based Monitoring](https://jobicy.com/jobs?search_keywords=Condition-Based%20Monitoring.md)[Asset Management](https://jobicy.com/jobs?search_keywords=Asset%20Management.md)[Maintenance Strategies](https://jobicy.com/jobs?search_keywords=Maintenance%20Strategies.md)[EAM](https://jobicy.com/jobs?search_keywords=EAM.md)[Project Management](https://jobicy.com/jobs?search_keywords=Project%20Management.md)[Cross-functional Collaboration](https://jobicy.com/jobs?search_keywords=Cross-functional%20Collaboration.md)

Cover letter sampleDear Hiring Manager,

I am excited to apply for the Data Center Facility Operations Reliability Engineer position at Meta. With over 8 years of experience in reliability engineering, including expertise in FMEA and RCM, I have a proven track record of optimizing asset performance and reducing downtime in critical environments. My background in data center operations and proficiency with EAM solutions align perfectly with the responsibilities outlined.

I have successfully led cross-functional teams to implement condition-based monitoring programs and drive continuous improvement in MTTR and MTBF metrics. I am confident that my technical skills and collaborative approach will contribute to Meta's mission of delivering reliable infrastructure.

Thank you for considering my application. I look forward to the opportunity to discuss how my experience can support your team.

Sincerely,
[Your Name]

Copy   Sample interview questionsCan you describe your experience with Failure Mode and Effects Analysis (FMEA) and how you have used it to improve maintenance strategies?In my previous role, I led FMEA sessions for critical cooling systems, identifying failure modes like pump seal failures. We implemented predictive maintenance and redesigned seals, reducing downtime by 30%. I also used the results to update our maintenance library and procedures.

How do you prioritize corrective maintenance initiatives when multiple assets have issues?

I use risk-based prioritization, considering criticality, failure impact, and likelihood. I also factor in production risk models and mean time between failures (MTBF) data. For example, I once prioritized a generator issue over a cooling fan because it had higher business impact.

Explain your experience with condition-based monitoring technologies such as vibration analysis or infrared thermography.

I have implemented vibration analysis on pumps and fans to detect bearing degradation early. I also used infrared thermography to identify hot spots in electrical panels. These sensors fed into our EAM system, triggering maintenance tasks automatically.

How do you handle managing maintenance content and ensuring compliance with change management processes?

I established a governance framework where all maintenance procedure changes required approval from reliability and safety teams. I used a version control system in our EAM to track updates. For example, when we optimized a filter replacement schedule, I ensured proper documentation and training.

Describe a time you used data analysis to improve reliability metrics like MTTR or MTBF.

At my last job, I analyzed historical failure data for chillers and found a common root cause: improper refrigerant charge. I worked with the maintenance team to standardize charging procedures, which increased MTBF by 20% and reduced MTTR by 15% through better troubleshooting guides.

Meta is seeking an experienced and self-motivated Reliability Engineer to join our Asset Management & Reliability team within Facility Operations. This person will work within Facility Operations to identify and manage asset reliability risks and various stages of end-to-end asset lifecycle for the Data Center Operations. Managing stakeholders spread across time zones and key to the success of our individual projects and overall asset management, and reliability program. In this role, you will be expected to own and execute within your focus areas, contribute to scoping and roadmapping for reliability initiatives, and solve complex technical problems to ensure seamless operations across new and existing sites.ResponsibilitiesLead operational Failure Mode and Effects Analysis (FMEA) and establish maintenance strategies.* Manage and govern content for the Global Maintenance Library, ensuring rigorous Maintenance Change Management processes are followed.* Drive Corrective Maintenance initiatives through deep Failure Mode Analysis to identify root causes and implement preventive measures.* Evaluate regulatory and compliance requirements, providing governance and ensuring adherence across all reliability activities.* Support and execute Condition-Based Monitoring and Diagnostics programs, including testing, vibration analysis, infrared (IR), partial discharge (PD), and robotics applications.* Assess and optimize Mean Time To Repair (MTTR) and Mean Time Between Failures (MTBF) metrics to drive continuous improvement.* Contribute to continuous production risk models to prioritize resources and mitigate operational vulnerabilities.* Manage non-critical scope and support supplier scope management to ensure external partners meet reliability and quality standards.* Implement and optimize sensor-based condition monitoring and maintenance (e.g., infrared, temperature, pressure, flow, leak detection) to remotely monitor and assign maintenance tasks.* Support technical insourcing versus outsourcing decision-making for maintenance and reliability activities.* Establish Service Bill of Materials (BOM) setups and drive initial and continuous spares selection processes.* Develop and execute a comprehensive whole-equipment sparing strategy to minimize downtime.* Identify First-of-Kind (FOK) parts and establish Parts Factory Base Part Number (FBPN) setups.* Conduct Form-Fit-Function technical analysis to evaluate and approve replacement components.* Act as the primary liaison to IBOS and ISCE teams for managing replacement parts backlog, timeliness, tooling requirements, and auditing.* Identify reliability opportunities and contribute to engineering excellence within the team.QualificationsBachelor’s degree in Mechanical, Electrical, Reliability Engineering, or a similar technical discipline or 6+ years relevant industry experience will be considered in lieu of 8+ years relevant industry experience* 8+ years of experience in reliability engineering (related to electrical or mechanical cooling equipment) or related Engineering roles* Experienced in Reliability Centered Maintenance (RCM) and Failure Mode and Effects Analysis (FMEA) activities for maintenance, process, and equipment design optimization to meet reliability requirements* Proven ability to execute independently within defined scope and manage complex problems with guidance on ambiguous areas* Proficient in the usage of Enterprise Asset Management (EAM) solutions to extract data and develop meaningful insights* Knowledgeable of relevant ISO standards (ISO 14224, ISO 17359, ISO 55000)* Experience with project management and cross-functional collaboration* Ability to travel 25% domestically and internationally Experience with data center equipment such as critical cooling systems, generators, main switchboards, and network gear* Proficient in data analysis techniques that can include Process Control, Reliability modeling and prediction, Fault Tree Analysis, Weibull Tree Analysis, and Six Sigma (6σ) Methodology* Proficient in developing and executing comprehensive test plans for assets* Experience supporting technical initiatives and advocating for high-quality engineering standards* Certifications in Maintenance & Reliability such as CMRP, CRL, or CRE

Show more

[Apply now >](https://jobicy.com/jobs/149829-data-center-facility-operations-reliability-engineer.md)

>  Annual salary information is not provided for this position. Explore salary ranges for similar roles in our [Salary Directory ›](https://jobicy.com/salaries.md)

*

![Upload CV](data:image/svg+xml;base64,PHN2ZyB4bWxucz0iaHR0cDovL3d3dy53My5vcmcvMjAwMC9zdmciIHdpZHRoPSI2NSIgaGVpZ2h0PSI2NSIgZmlsbD0ibm9uZSIgeG1sbnM6dj0iaHR0cHM6Ly92ZWN0YS5pby9uYW5vIj48ZyBjbGlwLXBhdGg9InVybCgjQSkiPjxwYXRoIGQ9Ik0wIDBINjVWNjVIMFYwWiIgZmlsbD0iIzAyOWFlYiIvPjxnIGZpbGw9IiNmZmYiIHN0cm9rZT0iI2ZmZiIgc3Ryb2tlLXdpZHRoPSIyIj48cGF0aCBkPSJNMzMuMDQ5IDE1LjQ1NGExLjQzIDEuNDMgMCAwIDAtMi4wOTcgMGwtNy41NzkgOC4xNDdhMS4zOCAxLjM4IDAgMCAwIC4wOSAxLjk3MyAxLjQ0IDEuNDQgMCAwIDAgMi4wMDgtLjA4OGw1LjEwOS01LjQ5MnYyMC42MWExLjQxIDEuNDEgMCAwIDAgMS40MjEgMS4zOTdjLjc4NSAwIDEuNDIxLS42MjUgMS40MjEtMS4zOTd2LTIwLjYxbDUuMTA5IDUuNDkyYTEuNDQgMS40NCAwIDAgMCAyLjAwOC4wODggMS4zOCAxLjM4IDAgMCAwIC4wOS0xLjk3M2wtNy41NzktOC4xNDZ6TTE2Ljc2OSAzOC40YzAtLjc3My0uNjItMS40LTEuMzg1LTEuNFMxNCAzNy42MjcgMTQgMzguNHYuMTAybC4yMTUgNi4yMjljLjIyMyAxLjY4LjcwMSAzLjA5NSAxLjgxMyA0LjIxOHMyLjUxIDEuNjA3IDQuMTcyIDEuODMzYzEuNi4yMTggMy42MzYuMjE4IDYuMTYuMjE4aDExLjI4bDYuMTYtLjIxOGMxLjY2Mi0uMjI2IDMuMDYxLS43MDkgNC4xNzItMS44MzNzMS41ODktMi41MzggMS44MTMtNC4yMThDNTAgNDMuMTEzIDUwIDQxLjA1NSA1MCAzOC41MDNWMzguNGMwLS43NzMtLjYyLTEuNC0xLjM4NS0xLjRzLTEuMzg1LjYyNy0xLjM4NSAxLjRsLS4xOSA1Ljk1OGMtLjE4MiAxLjM3LS41MTUgMi4wOTUtMS4wMjYgMi42MTJzLTEuMjI4Ljg1My0yLjU4MyAxLjAzOGMtMS4zOTUuMTktMy4yNDMuMTkzLTUuODkzLjE5M0gyNi40NjJjLTIuNjUgMC00LjQ5OC0uMDAzLTUuODkzLS4xOTMtMS4zNTUtLjE4NC0yLjA3Mi0uNTIxLTIuNTgzLTEuMDM4cy0uODQ0LTEuMjQyLTEuMDI2LTIuNjEyYy0uMTg3LTEuNDEtLjE5MS0zLjI3OS0uMTkxLTUuOTU4eiIvPjwvZz48L2c+PGRlZnM+PGNsaXBQYXRoIGlkPSJBIj48cGF0aCBmaWxsPSIjZmZmIiBkPSJNMCAwaDY1djY1SDB6Ii8+PC9jbGlwUGF0aD48L2RlZnM+PC9zdmc+)

### Upload your resume now

To unlock remote work opportunities and be discovered by global employers.

This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.

Next step

## Apply now.

Follow the employer’s application method and review Jobicy’s safety guidance before sharing personal information.

Keep exploring

## Related remote jobs.

Matched by job category10 related opportunities[DevOps & Infrastructure](https://jobicy.com/categories/admin.md) [Browse all jobs](https://jobicy.com/jobs.md)
*
![Akamai Technologies logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/85a7cc0e-221-1.jpeg)
Akamai Technologies  Jul 28

### [Senior Site Reliability Engineer](https://jobicy.com/jobs/143565-senior-site-reliability-engineer-3.md)

Are you passionate about cutting edge technology?Do solving some of the Internet’s most difficult content delivery challenges interest you?Join our highly skilled Site Reliability teamOur team designs, develops, and manages…

*
![Defense Unicorns logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/44a44784-221-1.png)
Defense Unicorns  Jul 26

### [Platform Engineer](https://jobicy.com/jobs/143041-platform-engineer.md)

EMPLOYER IS A CONTRACTOR FOR THE U.S. GOVERNMENT. THIS POSITION WILL REQUIRE AN ACTIVE SECRET CLEARANCE TO APPLY. Role DescriptionWe are seeking a talented and experienced Platform Engineer to join our…

*
![GLS Germany GmbH & Co. OHG logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/e577d043-221.jpeg)
GLS Germany GmbH & Co. OHG  Jul 26

### [(Senior) Information Security Architect (f/m/d)](https://jobicy.com/jobs/142385-senior-information-security-architect-f-m-d.md)

(Senior) Information Security Architect (f/m/d) Germany Full time Unlimited Immediately The GLS Group CISO Team is responsible for information security worldwide. In your role you will report to the Manager Security…

*
![Experian logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2021/09/dcc5b29a570bb19b9f5c3e150db2fdfe.jpg)
Experian  Jul 24

### [Senior DevOps Engineer](https://jobicy.com/jobs/144464-senior-devops-engineer-5.md)

Company DescriptionExperian is a global data and technology company, powering opportunities for people and businesses around the world. We operate across a range of markets, from financial services to healthcare,…

*
![Zapier logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2021/03/Jobicy-210317121153-967239.png)
Zapier  Jul 23

### [Sr. IT Systems Engineer](https://jobicy.com/jobs/147417-sr-it-systems-engineer.md)

AI at ZapierAt Zapier, we build and use automation every day to make work more efficient, creative, and human. So if you’re using AI tools while applying here – that’s…

*
![Canonical Ltd. logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/b4d068a9-221-1.png)
Canonical Ltd.  Jul 21

### [OpenStack Engineering Manager](https://jobicy.com/jobs/149573-openstack-engineering-manager.md)

Canonical is a leading provider of open source software and operating systems to the global enterprise and technology markets. Our platform, Ubuntu, is very widely used in breakthrough enterprise initiatives…

*
![Canonical Ltd. logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/b4d068a9-221-1.png)
Canonical Ltd.  Jul 21

### [Cloud Engineering Manager](https://jobicy.com/jobs/149569-cloud-engineering-manager.md)

Canonical is a leading provider of open source software and operating systems to the global enterprise and technology markets. Our platform, Ubuntu, is very widely used in breakthrough enterprise initiatives…

*
![Canonical Ltd. logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/b4d068a9-221-1.png)
Canonical Ltd.  Jul 21

### [Senior Site Reliability Engineer](https://jobicy.com/jobs/149557-senior-site-reliability-engineer-5.md)

Canonical is a leading provider of open source software and operating systems to the global enterprise and technology markets. Our platform, Ubuntu, is very widely used in breakthrough enterprise initiatives…

*
![Canonical Ltd. logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/b4d068a9-221-1.png)
Canonical Ltd.  Jul 21

### [Site Reliability / Gitops Engineer](https://jobicy.com/jobs/149553-site-reliability-gitops-engineer.md)

Canonical is a leading provider of open source software and operating systems to the global enterprise and technology markets. Our platform, Ubuntu, is very widely used in breakthrough enterprise initiatives…

*
![Canonical Ltd. logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/b4d068a9-221-1.png)
Canonical Ltd.  Jul 21

### [Site Reliability Engineer](https://jobicy.com/jobs/149547-site-reliability-engineer.md)

Canonical is a leading provider of open source software and operating systems to the global enterprise and technology markets. Our platform, Ubuntu, is widely used in breakthrough enterprise initiatives such…