Senior Principal AI/HPC Architect
Make an impact with NTT DATA
Join a company that is pushing the boundaries of what is possible. We are renowned for our technical excellence and leading innovations, and for making a difference to our clients and society. Our workplace embraces diversity and inclusion β itβs a place where you can grow, belong and thrive.
Your day at NTT DATA
The Senior Principal AI Infrastructure Architect is a highly skilled and advanced subject matter expert, responsible for leading the design of complex AI platform and managed-service solutions and driving the strategic vision and direction for the company's largest enterprise clients. The role sits at the centre of NTT DATA 's AI Factories practice and is focused on the hardware foundations β GPU and accelerator compute, host CPU platforms, high-performance storage and AI fabric β that underpin enterprise-scale training, fine-tuning and inference workloads.
Key Responsibilities:
- Lead the end-to-end design of large, complex AI infrastructure solutions β covering accelerated compute (NVIDIA H100/H200/B200 and GB200 NVL72, AMD Instinct MI300X/MI325X, Intel Gaudi 3), CPU host platforms (Intel Xeon, AMD EPYC, NVIDIA Grace), high-throughput storage tiers and lossless AI fabric β for enterprise, sovereign AI and AI Factory clients.
- Architect reference designs built on NVIDIA DGX/HGX SuperPOD, Dell AI Factory with NVIDIA, Cisco Nexus HyperFabric AI, HPE / Lenovo / Supermicro accelerated compute and equivalent platforms, balancing single-node performance with cluster-scale efficiency.
- Size and validate GPU clusters against real workloads β foundation-model pre-training, distributed fine-tuning, RAG, real-time and batch inference β using the right combination of NVLink/NVSwitch domains, InfiniBand NDR/XDR or Ultra Ethernet / NVIDIA Spectrum-X fabrics and tiered NVMe and parallel storage (VAST, WEKA, DDN, Pure FlashBlade, NetApp ONTAP AI, Dell PowerScale).
- Define the supporting datacenter design: high-density power (50β140 kW/rack), direct-to-chip and rear-door liquid cooling, structured cabling for AI fabrics and modular deployment models across on-prem, colo and sovereign-cloud footprints.
- Work closely with the sales team to drive the presales process for AI infrastructure pursuits β client discovery, technical workshops, proposal writing, executive presentations and bid defence.
- Translate clients' AI ambitions and business outcomes into a hardware and platform roadmap, positioning NTT DATA 's end-to-end portfolio β silicon, systems, storage, fabric, MLOps stack and managed services β to land service-led AI solutions.
- βLead integration of compute, storage, networking, the AI software stack (CUDA, ROCm, Triton, NIM, NVIDIA AI Enterprise, Run:ai, Slurm, Kubernetes / Kubeflow) and managed-service operating models across multiple domains, delivery units and geographies.
- Build business cases,TCOand unit-economics models (cost per token, cost per training run, GPU-hour economics) and end-to-end transition roadmaps for cloud-to-private AI migrations and sovereign AI deployments.
- Define architectural principles for AI infrastructure β acceleratorutilisation, data gravity, multi-tenancy, model lifecycle, energy efficiency β and apply them to influence architectural outcomes and governance.
- Develop As-Is, Vision, FMO and To-Be AI platform architectures,identifygaps and develop transition roadmaps.
- Synthesisecurrent and future trends in AI silicon, memory hierarchies (HBM3e, CXL),interconnectsand AI software stacks with client strategic imperatives to create compelling, evidence-based solutions.
- Contribute to NTT DATA 's AI Factories knowledge base by sharing reference architectures, sizing tools and lessons learned with internal teams and clients.
Knowledge and Attributes
- Deep, hands-on knowledge of AI hardware : GPU and accelerator portfolios (NVIDIA Hopper / Blackwell, AMD MI300/MI325, Intel Gaudi 3, emerging custom silicon), host CPU platforms (Intel Xeon, AM