Job Title: Data Center Engineer - Data Center Engineer
Job ID: 956-1
Contract Duration: 12 Months+
Work Location: US, CA, Santa Clara
JOB DESCRIPTION:
Datacenter Site Operations Lead - Software Infrastructure SW Infra | Colossus Cloud DC Planning and Engineering The Role The SW Infra | Colossus Cloud DC Planning and Engineering team designs, scales, and operates the global datacenter infrastructure supporting NVIDIA's AI platforms, cloud environments, and rapidly expanding compute footprint.
We are seeking a Datacenter Site Operations Lead (Contract) to own operational governance, install-base visibility, compliance readiness, and cross-functional stakeholder engagement across the IPP datacenter portfolio.
This role serves as a critical liaison across IPP Platform, Colossus Infrastructure, SRE, DCEO-Ops, Engineering, and other infrastructure partners, ensuring datacenters transition successfully from build completion into production operations and remain reliable, measurable, compliant, and operationally ready at scale.
The ideal candidate combines strong datacenter operational expertise with disciplined governance, systems thinking, technical depth, and the ability to translate complex infrastructure conditions into clear, actionable outcomes.
What You'll Be Doing
Datacenter Operational Governance & Visibility
- Manage datacenter install-base and operational visibility across space, power, cooling, infrastructure monitoring, and capacity utilization.
- Establish and maintain clear operational KPIs and metrics to provide visibility into datacenter health, readiness, and performance.
- Ensure infrastructure capacity allocations are accurately tracked, monitored, and operated throughout the datacenter lifecycle.
- Maintain strong operational discipline across systems of record and infrastructure reporting.
- Cross-Functional Operations Coordination
- Serve as a primary operational liaison across Engineering, IPP Platform, Colossus Infrastructure, SRE, DCEO-Ops, and other key stakeholders.
- Coordinate resolution of operational issues, dependencies, and escalations, driving actions through closure with clear ownership and accountability.
- Create transparency around operational health, constraints, risks, and readiness across the global datacenter portfolio.
- Translate technical and operational complexity into concise, actionable information for engineering and leadership audiences.
- Operational Readiness & Production Transition
- Lead the operational transition of newly built datacenters into Production Operations.
- Validate detailed datacenter acceptance and operational-readiness criteria in partnership with Build Leads and Operations teams.
- Ensure required infrastructure, monitoring, documentation, processes, ownership, and capacity controls are established prior to operational acceptance.
- Identify readiness gaps early and drive closure of blockers before production handoff.
- Governance, Compliance & Systems of Record
- Audit operational governance processes and systems of record to ensure DCIM and asset-data accuracy.
- Maintain and improve operational procedures, including SOPs, acceptance criteria, reporting standards, and governance controls.
- Support compliance and audit readiness, including applicable infrastructure standards such as ISO 50001.
- Establish accountability for the quality, completeness, and consistency of operational infrastructure data.
- Continuous Improvement & Operational Excellence
- Drive continuous-improvement initiatives across infrastructure monitoring, operational processes, EHS programs, governance, and overall datacenter operational excellence.
- Identify recurring operational gaps and lead corrective actions that improve reliability, scalability, and execution.
- Capture lessons learned from datacenter transitions and operational events and incorporate them into future readiness and acceptance processes.
- Develop repeatable operating models and standards capable of scaling across a rapidly growing global datacenter footprint.
- What We Need to See
- Bachelor's degree in Electrical Engineering, Mechanical Engineering, Computer Engineering, Datacenter Engineering, or a related technical discipline, or equivalent practical experience.
- 7+ years of experience supporting complex datacenter, hyperscale infrastructure, cloud infrastructure, or high-density compute environments.
- Strong understanding of datacenter infrastructure, including power, cooling, space, racks, monitoring, capacity management, and operational readiness.
- Demonstrated experience coordinating across engineering, infrastructure, platform, SRE, facilities, and operations organizations.
- Experience establishing and managing operational KPIs, readiness metrics, governance processes, and systems of record.
- Experience with datacenter asset management and/or DCIM environments and maintaining infrastructure-data accuracy.
- Strong understanding of operational acceptance, production handoff, SOPs, compliance, and audit-readiness practices.
- Excellent written and verbal communication skills with the ability to distill complex technical and operational issues into clear actions, decisions, and executive-level reporting.
- What Sets You Apart
- You think in systems, understanding that successful datacenter operations require the orchestration of infrastructure, compute, people, processes, monitoring, governance, and data.
- You identify blockers and operational constraints early, establish clear ownership, and drive them to closure before they affect readiness or production.
- You bring a strong bias toward operational rigor, accountability, data accuracy, and measurable outcomes.
- You are comfortable operating across technical and leadership organizations and can influence outcomes without relying solely on organizational authority.
- You anticipate risk rather than react to it and continuously look for opportunities to improve reliability, efficiency, scalability, and operational maturity.
- You can operate effectively in a fast-moving global environment where datacenter scale, AI infrastructure requirements, and business priorities are continuously evolving.
Trinus Corporation, a leading provider of technology solutions and services with over 25 years of experience, is a certified WBE/MBE/SBE/SDB firm accredited by WBENC, NMSDC, and SBA.
Our mission is to shape the future of work by aligning the right mix of people, process, technology, and innovation to efficiently meet our clients' business objectives.
At Trinus, we understand that finding the right opportunity is pivotal in your career journey. Our staffing services go beyond mere placements; they are about matching your skills and aspirations with the perfect fit.
To learn more about us, please visit our website www.trinus.com