Point your AI agent at freehire and let it find you a job.

Get the CLI →

Oracle

Principal Data Center Facilities Engineer

Posted Updated
Discussion

Provides global facility engineering expertise to improve the reliability and resilience of Oracle’s colocation data center portfolio. The role identifies systemic risk through incident reviews, colocation facility audits, design-deviation reviews, and engineering analysis; guides mitigations to closure; and translates operational learning into colocation standards and policies.


Working across Data Center Operations, Facility Engineering, Design, Construction, Commissioning, Energy, and colocation providers, this engineer provides electrical and mechanical facility engineering expertise for mission-critical colocation infrastructure. The role supports incident escalation, validates corrective and preventive actions, evaluates recurring equipment and provider risks, and develops technical guidance that can be applied across Oracle’s global colocation portfolio.

Global Reliability and Risk Management
  • Review colocation data center design deviations identified through formal design-review processes and track approved mitigations through implementation and validation.
  • Analyze recurring critical-infrastructure failures, provider-performance trends, and site-level events to identify risks that may affect other facilities in the colocation portfolio.

  • Assess reliability risks associated with aging infrastructure, including exposure to loss of utility, cooling, generation, electrical distribution, and other critical systems.

  • Develop technical findings and recommendations that prioritize risk reduction and support informed engineering decisions.

  • Partner with colocation providers and internal engineering teams to address equipment, design, and maintenance risks that recur across colocation sites.

Incident Escalation and Corrective Action
  • Provide engineering support during significant data center infrastructure incidents and abnormal operating events.

  • Participate in root-cause analysis and post-incident reviews; identify lessons learned and validate corrective and preventive action plans.

  • Track CAPA and RCA action items to completion, escalating barriers and overdue actions when needed.

  • Confirm that approved remediation resolves the identified risk and determine whether similar conditions exist elsewhere in the colocation portfolio.

Facility Audits and Operational Assurance
  • Lead or support facility engineering audits, including engagement with mechanical, electrical, operations, and colocation-provider stakeholders.

  • Assess operations and maintenance maturity against established standards and identify gaps that could affect reliability.

  • Evaluate critical maintenance practices, including equipment lifecycle management, battery replacement practices, control-system settings, and maintenance standards.

  • Support technical assessments of high-risk facilities and recommend remediation plans with clear ownership and completion criteria.

Engineering Standards and Technical Governance
  • Develop, review, and maintain facility engineering standards, operating policies, and technical guidance based on fleet experience and industry practices.

  • Provide technical guidance for items such as CDU firmware releases, valve settings, maintenance standards, battery replacement policies, and liquid-cooling requirements.

  • Partner with the Energy team on coordination between onsite generation and data center short-circuit coordination studies.

  • Monitor applicable industry and technology developments, including liquid-cooling guidance, and recommend appropriate actions for Oracle facilities.

  • Represent Reliability Engineering in design reviews, technical working groups, and change reviews involving material impact to critical infrastructure.

Design and Platform Reliability
  • Provide facility engineering input for new rack platforms, liquid-cooling deployments, and infrastructure changes that may affect colocation reliability.

  • Participate in design and commissioning reviews to identify operational and reliability risks before turnover.

  • Review technical reports, simulations, and engineering documentation as needed to support reliability recommendations.

Knowledge Sharing
  • Share incident learning, audit findings, and technical guidance with regional Facility Engineering and Data Center Operations teams.

  • Act as an electrical or mechanical subject matter expert and support the development of technical capability across the organization.

  • Mentor junior engineers and contribute to training materials, procedures, and engineering playbooks.

Additional Responsibilities as Needed
Colocation Provider Management
  • Provide technical input for colocation-provider contract administration, including statements of work, change orders, cost forecasts, and other engineering documentation.

  • Coordinate with colocation providers and internal stakeholders on contractor scheduling, site access, and work execution in accordance with Oracle safety and operating expectations.

  • Support assessment of resource readiness for critical infrastructure work, including spare parts, generator fuel, water, consumables, and other provider-managed resources.

  • Review provider performance and maintenance-execution trends; escalate material gaps that could affect reliability across the colocation portfolio.

Minimum Qualifications
8 years of experience in operating data center critical infrastructure, critical facility operations, or a related field.
OR
Bachelor’s degree in Electrical Engineering, Mechanical Engineering, Industrial Engineering, Process Engineering, or a related field and 6 years of experience in operating data center critical infrastructure, critical facility operations, or a related field.
Required Skills
  • Demonstrated ability to troubleshoot and resolve complex technical issues across multiple technology domains.

  • Demonstrated ability to respond to and resolve incidents quickly to maintain business continuity.

Preferred Qualifications
12 years of experience in operating data center critical infrastructure, critical facility operations, or a related field.
OR
Bachelor’s degree in Electrical Engineering, Mechanical Engineering, Industrial Engineering, Process Engineering, or a related field and 8 years of experience in operating data center critical infrastructure, critical facility operations, or a related field.
  • Experience evaluating reliability risk across multiple colocation data centers, regions, or colocation providers.

  • Experience leading or supporting root-cause analysis, CAPA, facility audits, or operations-and-maintenance maturity assessments.

  • Deep electrical or mechanical expertise in critical infrastructure, including power, cooling, controls, generators, UPS, batteries, and fire and life-safety systems.

  • Ability to translate technical findings into standards, policies, and actionable remediation plans.

  • Experience working with operation and maintenance of building systems, uninterruptible power supplies, generators, cooling, building automation, and fire and life-safety systems.

Disclaimer:

Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.

Range and benefit information provided in this posting are specific to the stated locations only

US: Hiring Range in USD from: $102,300 to $209,500 per annum. May be eligible for bonus and equity.


Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.

Oracle US offers a comprehensive benefits package which includes the following:
1. Medical, dental, and vision insurance, including expert medical opinion
2. Short term disability and long term disability
3. Life insurance and AD&D
4. Supplemental life insurance (Employee/Spouse/Child)
5. Health care and dependent care Flexible Spending Accounts
6. Pre-tax commuter and parking benefits
7. 401(k) Savings and Investment Plan with company match
8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
9. 11 paid holidays
10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
11. Paid parental leave
12. Adoption assistance
13. Employee Stock Purchase Plan
14. Financial planning and group legal
15. Voluntary benefits including auto, homeowner and pet insurance

The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
As part of Oracle's onboarding process and consistent with applicable law, US-based employees are required to complete identity verification, which involves the collection and processing of their biometric information. Accommodations to this requirement may be granted following an individualized assessment.

Skills

What Principal Industrial Engineering jobs ask for — and how much of it you have →

See also

Industrial Engineering jobs by country — openings, pay and top skills →

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available