Graphcore Logo

Graphcore

Technical Services Director, Global Data Center & Lab Infrastructure

Posted 44 Minutes Ago
Be an Early Applicant
Hybrid
Austin, TX
Expert/Leader
Hybrid
Austin, TX
Expert/Leader
Leads global technical services teams supporting engineering labs, HPC platforms, and data center environments. Owns infrastructure reliability, security, capacity, procurement, budgets, supplier performance, operational standards, and lifecycle programs. Partners across engineering, IT, security, facilities, finance, supply chain, customers, and vendors to deliver resilient infrastructure for AI, silicon development, simulation, and validation workloads. The role is onsite in Austin and requires international travel.
The summary above was generated by AI
About Graphcore

Graphcore is a global leader in artificial intelligence computing systems. We design advanced semiconductors and data center hardware that deliver the specialized processing power needed to advance AI while improving the efficiency required for broad adoption.

As part of SoftBank Group, Graphcore belongs to a family of companies developing some of the world's most transformative technologies. Our AI Engineering Campus in Austin plays an important role in building the future of AI computing.

The Opportunity

As Technical Services Director, you will lead the teams that operate and evolve Graphcore's engineering labs, high-performance computing (HPC) platforms, and data center environments globally. You will be accountable for reliable, secure, cost-effective infrastructure that supports demanding engineering, AI, silicon-development, and validation workloads.

This role combines people leadership, infrastructure strategy, operational excellence, capacity and financial planning, procurement, and program delivery. You will partner with Engineering, Information Technology, Security, Finance, Facilities, Supply Chain, customers, and external suppliers. The position is based onsite in Austin and requires travel to company facilities, data centers, and supplier locations, including international travel.

What You'll Do
  • Lead, recruit, mentor, and develop the systems administration, lab operations, and technical services teams responsible for the facility supporting global Engineering and Research and Development.  
  • Own the reliability, efficiency, protection, safety, supportability, and continuous improvement of engineering labs, HPC systems, and infrastructure facilities.    
  • Establish service levels, operating standards, escalation paths, performance measures, monitoring, observability, automation, ticketing, and configuration-management practices.
  • Translate engineering and customer requirements into infrastructure roadmaps, capacity plans, procurement strategies, operating models, and executable investment programs.
  • Forecast compute, accelerator, storage, network, rack-space, power, cooling, and technical-support requirements across the infrastructure portfolio.
  • Lead global infrastructure programs from requirements and business-case development through procurement, deployment, operational readiness, service handoff, expansion, refresh, and decommissioning.
  • Own infrastructure procurement and supplier performance, including specifications, bills of materials, quotations, commercial negotiations, purchase orders, logistics, delivery schedules, and deployment coordination.
  • Develop and lead operating budgets, capital plans, expense frameworks, projections, lifecycle plans, and investment proposals for infrastructure operations and growth.  
  • Build strategic relationships with hardware vendors, colocation providers, integrators, maintenance partners, and other technical service suppliers.
  • Ensure HPC environments are efficiently utilized, maintained, and capable of supporting large-scale engineering, AI, simulation, and hardware-validation workloads.
  • Maintain appropriate security, safety, operational, and compliance controls; support reviews involving frameworks such as ISO 27001, SOC/SSAE, and PCI DSS where applicable.
  • Engage internal and external customers on infrastructure capabilities, requirements, service delivery, resilience, security, and compliance.
  • Provide concise executive reporting on infrastructure health, capacity, risks, budgets, supplier performance, service quality, and coordination of programs.
What You'll Bring
  • Significant leadership experience in global technical infrastructure, engineering labs, systems engineering, HPC, Information Technology operations, or data center services.
  • Proven success leading and developing geographically distributed technical teams, including managers and senior individual contributors.
  • Deep knowledge of HPC and data center environments, including server architecture, Linux-based systems, accelerators, high-speed interconnects, parallel storage, cluster management, workload scheduling, and monitoring.
  • Experience supporting semiconductor development, silicon bring-up, hardware validation, systems engineering, or similarly complex engineering environments.
  • A record of developing and executing multi-site infrastructure strategy, capacity plans, lifecycle programs, and operating models aligned with engineering and business priorities.
  • Experience operating business-critical compute and lab infrastructure with clear expectations for availability, performance, security, safety, and support.
  • Experience leading complex cross-functional programs involving Engineering, Facilities, Information Technology, Finance, Security, Supply Chain, and third-party vendors.  
  • Strong procurement and commercial experience, including requirements, supplier evaluation, contract negotiation, purchase-order processes, logistics, and delivery management.
  • Experience developing global budgets, cost models, forecasts, investment proposals, and lifecycle plans for technical infrastructure.
  • Working knowledge of rack integration, structured cabling, power, cooling, physical security, capacity management, operational controls, and relevant compliance frameworks.
  • Strong strategic, analytical, problem-solving, and communication skills, including the ability to influence senior leaders and explain complex issues to technical and nontechnical stakeholders.
  • Ability and willingness to work onsite in Austin and travel internationally as required.
  • A bachelor's degree in engineering, computer science, information technology, or another relevant technical discipline, or equivalent practical experience.
Preferred Qualifications
  • Experience designing, deploying, or operating large-scale HPC clusters or supercomputing platforms.
  • Knowledge of GPU-accelerated computing, AI infrastructure, and NVIDIA reference architectures.
  • Experience with cluster schedulers, high-performance storage, high-speed networking, and infrastructure automation technologies.
  • Experience developing a global lab strategy or managing lab and data center operations across multiple countries, including the United States, United Kingdom, or India.
  • Experience operating colocation facilities and managing third-party data center providers.
  • Knowledge of data center power distribution, cooling systems, environmental monitoring, and capacity engineering.
  • ITIL, program-management, or project-management certification, or equivalent practical experience implementing service-management and delivery frameworks.
  • Experience with Visio, Bluebeam, AutoCAD, or other infrastructure design and drafting tools.

These qualifications are helpful, not mandatory. We encourage you to apply even if you do not meet every preferred qualification.

U.S. Benefits Overview

Graphcore offers competitive compensation and a benefits package designed to support employees' health, financial well-being, families, flexibility, and development. Benefits for eligible U.S. employees may include:

  • Medical, dental, and vision coverage, together with mental health and well-being support.
  • A 401(k) retirement plan with company matching, plus company-provided life and disability coverage.
  • Flexible paid time off, paid company holidays, flexible working hours, and parental leave.
  • Learning and career development support, including development plans, training allowances, and online learning resources.
  • Office amenities and team-led social, community, and well-being activities.

Benefits vary by work location, employment status, and plan eligibility and are subject to the terms of the applicable plans and company policies. This overview is not a contract or guarantee of benefits.

Equal Opportunity and Accommodations

Graphcore is an equal opportunity employer. We consider qualified applicants without regard to race, color, religion, creed, sex, pregnancy, sexual orientation, gender identity or expression, national origin, ancestry, age, disability, genetic information, veteran status, or any other status protected by applicable law.

Graphcore is committed to an inclusive and accessible hiring process. If you need a reasonable accommodation to participate in the application or interview process, please let the recruiting team know.

Candidate Privacy

Personal information submitted during the recruiting process will be handled in accordance with Graphcore's applicable candidate privacy notices.

Similar Jobs at Graphcore

5 Days Ago
Hybrid
Senior level
Senior level
Artificial Intelligence • Semiconductor
Designs and validates thermal solutions for AI data center hardware, specializing in direct liquid cooling, cold plates, heatsinks, CDUs/CDMs, simulations, testing, and control algorithms. Leads thermal specifications, design reviews, qualification, root-cause analysis, and vendor collaboration. Works cross-functionally with system architects, chip and board designers, and software or firmware engineers to deliver reliable, high-performance server cooling systems.
Top Skills: AnsysCdu/CdmCold Plate CoolingComsolDirect Liquid CoolingFan And Thermal Control AlgorithmsFlowthermHeatsinks
Senior level
Artificial Intelligence • Semiconductor
Performs mechanical and thermal testing and validation of AI server and data center hardware. Operates laboratory, machine shop, CNC, and material-handling equipment; maintains test instruments; reads engineering documentation; develops test procedures; analyzes test data; supports troubleshooting, root-cause analysis, prototype fabrication, hardware installation, and equipment movement while maintaining laboratory safety.
Top Skills: CadCnc Milling MachinesData Center EquipmentDirect Liquid CoolingDrill PressesForkliftsPallet JacksServer HardwareThermal Testing Instrumentation
8 Days Ago
Hybrid
Mid level
Mid level
Artificial Intelligence • Semiconductor
Seeking a Staff Hardware Engineer to support Graphcore's AI hardware platforms through troubleshooting, validation, and engineering support. Responsibilities include diagnosing failures, collaborating with teams, and providing mentorship to junior engineers.
Top Skills: Ai Compute PlatformsBashHpc SystemsPythonServer Hardware Architectures

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account