Accenture Logo

Accenture

AI & HPC Infrastructure Engineer

Posted 3 Hours Ago
Be an Early Applicant
In-Office
Manhattan, NY, USA
87K-266K Annually
Senior level
In-Office
Manhattan, NY, USA
87K-266K Annually
Senior level
Design, deploy, and operate AI and accelerated computing infrastructure across cloud, on-prem, and hybrid environments. Build and manage XPU clusters, orchestrate workloads with Kubernetes/Slurm/Run:ai, integrate model-serving and governance, tune performance using NVIDIA tools, create documentation and runbooks, and provide troubleshooting and optimization for large-scale training, inference, and agentic AI workloads. Travel to client sites as required.
The summary above was generated by AI

We Are:

The Global AI Infrastructure team is at the center of enabling infrastructure reinvention for the next era of digital solutions powered by AI, accelerated computing, and high-performance workloads. We bring together deep technical expertise across cloud, on-premises, and hybrid environments to design, build, and operate advanced infrastructure that powers AI platforms, GPU-accelerated workloads, large-scale models, simulations, and emerging agentic AI solutions at scale. Our solutions enable some of our most strategic and mission-critical clients to unlock new levels of performance, efficiency, governance, and innovation. Our remit spans the full lifecycle-from strategy and architecture through implementation and operations-driving modernization across the entire infrastructure stack. We collaborate across the ecosystem to harness emerging technologies, fuel growth, and transform industries. In this rapidly growing market, our team is leading the way in shaping how enterprises leverage AI infrastructure to drive breakthrough innovation and reimagine what is possible.

Key Responsibilities:

  • Design and implement AI infrastructure and accelerated computing solutions, aligning system architecture and deployment roadmaps to industry-specific performance, scalability, resiliency, and governance needs

  • Deploy, configure, and manage XPU-based clusters (GPU, DPU, LPU, CPU) across bare-metal and containerized environments using workload schedulers (Slurm, Run:ai), Kubernetes orchestration, and container platforms to deliver scalable AI infrastructure services including Bare-Metal-aaS, GPUaaS, AIaaS, Token-aaS, model serving, and agentic AI frameworks

  • Integrate AI infrastructure platforms with existing IT systems, data pipelines, security frameworks, model-serving endpoints, and enterprise governance controls

  • Design and implement agentic AI infrastructure by integrating platform services, model endpoints, tool and function calling, retrieval patterns, and workflow orchestration with observability, identity, and policy controls through secure, deterministic APIs to support governed enterprise use cases

  • Build and integrate MCP servers, tools, connectors, and adapters that allows agents to monitor, troubleshoot, and tune infrastructure to ensure high availability, low-latency networking, and workload resiliency

  • Architect and deploy with NVIDIA platform tools including Base Command Manager (BCM), NGC, NCCL, NVLink, and CUDA along with LLM inference engines (TensorRT-LLM), production serving frameworks (vLLM, SGLang), inference orchestration (Triton Inference Server, NVIDIA Dynamo, llm-d), and GPU benchmarking and validation tools (MLPerf, NCCL tests, fio, iperf) to deploy, tune, profile, and validate AI cluster performance across compute and networking layers including multi-node training and inference workloads

  • Develop and maintain documentation including architecture diagrams, configuration baselines, and operational runbooks

  • Provide technical guidance, troubleshooting, and optimization across AI workloads including large-scale training, inference, multi-node simulations, and agentic pipelines while leveraging digital twins to validate infrastructure and drive performance, scalability, energy efficiency, and token cost optimization

Travel may be required for this role.  The amount of travel will vary from 25% to 100% depending on business need and client requirements.

Required Skills and Qualifications:

  • Minimum of 5+ years of experience designing, deploying, and managing AI infrastructure and accelerated computing environments across on-premises, cloud, and hybrid environments for hyperscaler, neocloud, large enterprise, Telco/Mobile, Financial Services, Life Sciences, Manufacturing, and/or Retail clients.

  • Minimum of 5+ years of hands-on experience with accelerated computing platforms, including GPUs, DPUs, LPUs, CPUs, high-speed interconnects such as InfiniBand or Ethernet, data center networking such as SONiC, and AI storage architectures including NVMe, NVMe-oF, parallel file systems, VAST, Weka, or DDN.

  • Minimum of 5+ years of experience with cluster management, workload scheduling, orchestration, observability, and infrastructure automation using platforms and tools such as Kubernetes, Slurm, Run:ai, AWS, Azure, GCP, VMware, Nutanix, Python, Terraform, and Ansible.

  • Bachelor's degree or equivalent (minimum 12 years) work experience. If Associate's Degree, must have minimum 6 years work experience.

Preferred Skills and Qualifications:

  • 2+ years of experience implementing MLOps, LLMOps, agentic AI, and DevSecOps frameworks to enable secure, automated, governed, and reproducible AI workflows.

  • 2+ years of experience developing APIs, integration services, automation workflows, or platform services using Python and modern API patterns such as REST, OpenAPI, JSON/YAML schemas, webhooks, and event-driven integrations.

  • Experience designing and implementing agentic AI infrastructure, including LLM inference, tool/function calling, retrieval-augmented generation (RAG), agent orchestration, secure API integration, policy-based governance, and deterministic platform APIs.

  • Experience building and integrating MCP servers, tools, connectors, and adapters that allow agents to monitor, troubleshoot, and tune infrastructure for high availability, low-latency networking, workload resiliency, and intelligent observability.

  • Experience using NVIDIA platform tools including Base Command Manager (BCM), NGC, NCCL, NVLink, CUDA, TensorRT-LLM, Triton Inference Server, NVIDIA Dynamo, llm-d, vLLM, SGLang, MLPerf, NCCL tests, fio, and iperf to deploy, tune, profile, and validate AI cluster performance.

  • Experience managing the deployment of 1,000+ GPU clusters for AI, HPC, and agentic AI workloads with infrastructure services enabled.

  • Design and build experience in AI Cloud platforms from CoreWeave, Nebius, and other specialty cloud providers.

  • Knowledge of machine learning and AI frameworks such as TensorFlow, PyTorch, JAX, Jupyter notebooks, and Google Colab environments.

  • Industry certifications in NVIDIA infrastructure, public cloud providers, data science, infrastructure automation, networking, or security are a plus.

Compensation at Accenture varies depending on a wide array of factors, which may include but are not limited to the specific office location, role, skill set, and level of experience. As required by local law, Accenture provides a reasonable range of compensation for roles that may be hired as set forth below.
We anticipate this job posting will be posted until 10/15/2026.
Accenture offers a market competitive suite of benefits including medical, dental, vision, life, and long-term disability coverage, a 401(k) plan, bonus opportunities, paid holidays, and paid time off. See more information on our benefits here:

U.S. Employee Benefits | Accenture

Role Location Annual Salary Range
California $94,400 to $266,300
Cleveland $87,400 to $213,000
Colorado $94,400 to $230,000
District of Columbia $100,500 to $245,000
Illinois $87,400 to $230,000
Maine $80,400 to $196,000
Maryland $94,400 to $230,000
Massachusetts $94,400 to $245,000
Minnesota $94,400 to $230,000
New York $87,400 to $266,300
New Jersey $100,500 to $266,300
Virginia $87,400 to $245,000
Washington $100,500 to $245,000

About Accenture

Accenture is a leading global professional services company that helps the world’s leading businesses, governments and other organizations build their digital core, optimize their operations, accelerate revenue growth and enhance citizen services—creating tangible value at speed and scale. We are a talent- and innovation-led company with approximately 791,000 people serving clients in more than 120 countries. Technology is at the core of change today, and we are one of the world’s leaders in helping drive that change, with strong ecosystem relationships. We combine our strength in technology and leadership in cloud, data and AI with unmatched industry experience, functional expertise and global delivery capability. Our broad range of services, solutions and assets across Strategy & Consulting, Technology, Operations, Industry X and Song, together with our culture of shared success and commitment to creating 360° value, enable us to help our clients reinvent and build trusted, lasting relationships. We measure our success by the 360° value we create for our clients, each other, our shareholders, partners and communities.

Visit us at www.accenture.com 

What We Believe 

We have an unwavering commitment to diversity with the aim that every one of our people has a full sense of belonging within our organization. As a business imperative, every person at Accenture has the responsibility to create and sustain an inclusive environment. 

Inclusion and diversity are fundamental to our culture and core values. Our rich diversity makes us more innovative and more creative, which helps us better serve our clients and our communities. Read more here 

Requesting An Accommodation 

Accenture is committed to providing equal employment opportunities for persons with disabilities or religious observances, including reasonable accommodation when needed. If you are hired by Accenture and require accommodation to perform the essential functions of your role, you will be asked to participate in our reasonable accommodation process. Accommodations made to facilitate the recruiting process are not a guarantee of future or continued accommodations once hired.

If you would like to be considered for employment opportunities with Accenture and have accommodation needs such as for a disability or religious observance, please call us toll free at 1 (877) 889-9009 or send us an email or speak with your recruiter.

Equal Employment Opportunity Statement 

We believe that no one should be discriminated against because of their differences. All employment decisions shall be made without regard to age, race, creed, color, religion, sex, national origin, ancestry, disability status, military veteran status, sexual orientation, gender identity or expression, genetic information, marital status, citizenship status or any other basis as protected by applicable law. Our rich diversity makes us more innovative, more competitive, and more creative, which helps us better serve our clients and our communities.

For details, view a copy of the  Accenture Equal Opportunity Statement

Accenture is an EEO and Affirmative Action Employer of Veterans/Individuals with Disabilities.

Accenture is committed to providing veteran employment opportunities to our service men and women.

Other Employment Statements 

Applicants for employment in the US must have work authorization that does not now or in the future require sponsorship of a visa for employment authorization in the United States.

Candidates who are currently employed by a client of Accenture or an affiliated Accenture business may not be eligible for consideration.

Job candidates will not be obligated to disclose sealed or expunged records of conviction or arrest as part of the hiring process. Further, at Accenture a criminal conviction history is not an absolute bar to employment. 

The Company will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay or the pay of another employee or applicant. Additionally, employees who have access to the compensation information of other employees or applicants as a part of their essential job functions cannot disclose the pay of other employees or applicants to individuals who do not otherwise have access to compensation information, unless the disclosure is (a) in response to a formal complaint or charge, (b) in furtherance of an investigation, proceeding, hearing, or action, including an investigation conducted by the employer, or (c) consistent with the Company's legal duty to furnish information.

California requires additional notifications for applicants and employees. If you are a California resident, live in or plan to work from Los Angeles County upon being hired for this position, please click here for additional important information.

Please read Accenture’s Recruiting and Hiring Statement for more information on how we process your data during the Recruiting and Hiring process.

Accenture New York, New York, USA Office

1345 6th Ave, New York, NY, United States, 10105

Similar Jobs

3 Hours Ago
Hybrid
Senior level
Senior level
eCommerce • Healthtech • Pet • Retail • Pharmaceutical
Lead end-to-end fulfillment center operations across two shifts, manage Area Managers and hourly teams, drive KPIs (on-time shipments, defects, productivity), implement continuous improvement, optimize WMS and inventory/order management, ensure safety and staffing, and support forecasting, training, and cross-functional collaboration.
Top Skills: Green BeltInventory Management SystemLean Six SigmaOrder Management SystemWarehouse Management System (Wms)
3 Hours Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
89K-135K Annually
Senior level
89K-135K Annually
Senior level
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Lead enterprise customer implementations of Samsara's IoT platform, managing deployments, launch plans, training, and adoption tracking. Serve as primary customer contact, coordinate cross-functional teams, mentor peers, and ensure customers achieve time-to-value and ongoing product usage across fleets and industrial assets.
Top Skills: AppsDriver WorkflowsEquipment MonitoringIotSamsara PlatformVehicle TelematicsVideo-Based Safety
3 Hours Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
98K-132K Annually
Mid level
98K-132K Annually
Mid level
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Manage strategic public-sector customers using the Samsara IoT platform to drive adoption, safety, efficiency, and sustainability. Create joint success plans, run executive business reviews and workshops, recommend workflow changes, mentor CSM and support teams, and coordinate cross-functional stakeholders to remove barriers and deliver business value.
Top Skills: AppsIotSamsara PlatformVehicle TelematicsVideo-Based Safety

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account