Sr Solution Architect (76719)
Job Fit Check
Base Career helps you apply smarter for this job.
Key skills for this role
Role Overview
AI Architect The Role The AI Architect is responsible for defining, designing, and governing end‑to‑end artificial intelligence system architectures that align with business objectives, data strategies, and enterprise technology standards. This role provides technical leadership across AI solution lifecycles, from ideation to production, ensuring scalability, security, interoperability, and regulatory compliance. Competency Focus: AI Infrastructure and architecture design, cloud-native architecture, model governance, large‑scale distributed systems Keywords: HPC Architect, HPC Architecture and System Design Responsibilities: Architect, deploy, and operate large-scale accelerator clusters, including NVIDIA DGX platforms, discrete NVIDIA and AMD GPUs, and TPU-based systems, ensuring high availability, scalability, and performance. Design and architect high‑bandwidth, low‑latency interconnect architectures, leveraging technologies such as InfiniBand, NVLink, and RoCE to support distributed AI training and inference workloads. Architect and design end‑to‑end AI training and inference platforms across on‑premises and public cloud environments (Azure, AWS, GCP), incorporating elastic GPU resource orchestration and automated scaling mechanisms. Architect and engineer high‑performance, large‑scale data delivery and storage solutions, including petabyte‑scale object storage and distributed file systems (e.g., VAST Data, WekaIO, DDN) optimized for AI and high‑throughput workloads. Design and architect streaming and batch data ingestion pipelines optimized for AI/ML workflows, enabling efficient data preprocessing, feature ingestion, and model training at scale. Architect and enforce secure GPU and compute isolation mechanisms, utilizing Kubernetes primitives such as RBAC, namespace isolation, and network policies to ensure multi‑tenant security, governance, and compliance. Evaluate, benchmark, and qualify emerging AI hardware platforms and software frameworks, conducting performance, scalability, and cost‑efficiency assessments to inform technology adoption decisions. Mentor engineers in AI infra best practices, observability, and capacity management Define the reference architecture for enterprise-wide AI adoption. Understanding on Sovereign AI Qualifications & Experience B. Tech/B.E. in Computer Science, Artificial Intelligence, Data Science, or related discipline; M. Tech/MS preferred 12+ years in infrastructure/cloud engineering, with 4+ years focused purely on AI/ML systems. Deep expertise in GPU cluster management, distributed compute, and container orchestration. Hands-on experience with Kubernetes for AI workloads, GPU scheduling, and Ray/Kubeflow pipelines. Basic Understanding of LLM training, fine-tuning, quantization, and model optimization. Certifications Required: NVIDIA Certified Associate – AI Infrastructure NVIDIA Professional Certification for AI Networking and AI Infrastructure Certified Kubernetes Administrator Cloud Certification (AWS, Azure, GCP) How You’ll Grow At HCLTech, we provide continuous opportunities for you to discover your spark and grow through meaningful, hands-on experiences. We encourage you to collaborate across diverse teams, engage in vendor interactions, and build strong professional networks that br
Key Skills for This Role
Full Job Posting
Job Summary
AI Architect The Role The AI Architect is responsible for defining, designing, and governing end‑to‑end artificial intelligence system architectures that align with business objectives, data strategies, and enterprise technology standards. This role provides technical leadership across AI solution lifecycles, from ideation to production, ensuring scalability, security, interoperability, and regulatory compliance. Competency Focus: AI Infrastructure and architecture design, cloud-native architecture, model governance, large‑scale distributed systems Keywords: HPC Architect, HPC Architecture and System Design Responsibilities: Architect, deploy, and operate large-scale accelerator clusters, including NVIDIA DGX platforms, discrete NVIDIA and AMD GPUs, and TPU-based systems, ensuring high availability, scalability, and performance. Design and architect high‑bandwidth, low‑latency interconnect architectures, leveraging technologies such as InfiniBand, NVLink, and RoCE to support distributed AI training and inference workloads. Architect and design end‑to‑end AI training and inference platforms across on‑premises and public cloud environments (Azure, AWS, GCP), incorporating elastic GPU resource orchestration and automated scaling mechanisms. Architect and engineer high‑performance, large‑scale data delivery and storage solutions, including petabyte‑scale object storage and distributed file systems (e.g., VAST Data, WekaIO, DDN) optimized for AI and high‑throughput workloads. Design and architect streaming and batch data ingestion pipelines optimized for AI/ML workflows, enabling efficient data preprocessing, feature ingestion, and model training at scale. Architect and enforce secure GPU and compute isolation mechanisms, utilizing Kubernetes primitives such as RBAC, namespace isolation, and network policies to ensure multi‑tenant security, governance, and compliance. Evaluate, benchmark, and qualify emerging AI hardware platforms and software frameworks, conducting performance, scalability, and cost‑efficiency assessments to inform technology adoption decisions. Mentor engineers in AI infra best practices, observability, and capacity management Define the reference architecture for enterprise-wide AI adoption. Understanding on Sovereign AI Qualifications & Experience B. Tech/B.E. in Computer Science, Artificial Intelligence, Data Science, or related discipline; M. Tech/MS preferred 12+ years in infrastructure/cloud engineering, with 4+ years focused purely on AI/ML systems. Deep expertise in GPU cluster management, distributed compute, and container orchestration. Hands-on experience with Kubernetes for AI workloads, GPU scheduling, and Ray/Kubeflow pipelines. Basic Understanding of LLM training, fine-tuning, quantization, and model optimization. Certifications Required: NVIDIA Certified Associate – AI Infrastructure NVIDIA Professional Certification for AI Networking and AI Infrastructure Certified Kubernetes Administrator Cloud Certification (AWS, Azure, GCP) How You’ll Grow At HCLTech, we provide continuous opportunities for you to discover your spark and grow through meaningful, hands-on experiences. We encourage you to collaborate across diverse teams, engage in vendor interactions, and build strong professional networks that br
About HCLTech
Supercharging progress through technology, engineering, and digital transformation.
Visit company websiteApply for this job in 1 click
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
More jobs at HCLTech
Functional Lead - Oracle Financial Cloud
, IND
Senior Administrator - English, Arabic, Microsoft Windows
, IND
The employer is seeking a senior service desk language specialist to provide Level 2 remote desktop and end-user computing support. The role requires advanced troubleshooting, escalation handling, knowledge management, a
Senior Administrator - English, Arabic, Microsoft Windows
, IND
The employer is seeking a senior service desk language specialist to provide Level 2 remote desktop and end-user computing support. The role requires advanced troubleshooting, escalation handling, knowledge management, a
Functional Lead - Oracle Financial Cloud
, IND
Senior Administrator - English, Arabic, Microsoft Windows
, IND
Senior Administrator - English, Arabic, Microsoft Windows
, IND
SME - IBM AIX, Power HA (166020)
, IND
SME - Program & Project Management (165346)
, IND
Technical Architect (158792)
, IND
Automation Test Lead Embedded (152157)
, IND
Senior Technical Lead (164957)
, USA