Job Description
HPC Solutions Architect
\n
Location: Dallas, TX (Hybrid)
\n
Type: Direct Hire
\n
• Competitive base salary + performance bonus
\n
• 100% company-paid benefits
\n
Overview
\n
We are seeking an HPC Solutions Architect to drive the technical design, integration, and delivery of high-performance computing (HPC) solutions supporting advanced compute workloads.
\n
This is a highly technical, customer-facing role focused on designing scalable, high-performance architectures across compute, storage, networking, Kubernetes, and security domains. The position spans the full solution lifecycle—from requirements discovery and workload analysis through proof-of-concept, deployment, and ongoing optimization.
\n
The ideal candidate brings deep expertise in HPC architectures, strong hands-on experience with performance tuning and system design, and the ability to translate complex customer requirements into scalable, production-ready solutions.
\n
Key Responsibilities
\n
Customer Engagement & Technical Discovery
\n
• Work directly with customers to understand HPC workload requirements, performance targets, and technical objectives
\n
• Lead technical discovery sessions to assess application characteristics, bottlenecks, and scalability challenges
\n
• Serve as a trusted technical advisor throughout the solution lifecycle
\n
Solution Architecture & Design
\n
• Design and document end-to-end HPC architectures across compute (CPU/GPU), storage, networking, orchestration, and security
\n
• Recommend hardware and software solutions aligned with performance, scalability, and efficiency goals
\n
• Develop architecture blueprints, integration guides, and implementation plans
\n
Performance Optimization & Workload Engineering
\n
• Support proof-of-concept and benchmarking initiatives to validate solution performance
\n
• Perform workload profiling, system tuning, and performance optimization
\n
• Conduct workload reviews to improve scalability, resilience, and efficiency
\n
Implementation & Delivery
\n
• Provide technical guidance during deployment to ensure seamless integration into customer environments
\n
• Act as a liaison between customers and internal engineering, product, and operations teams
\n
• Support solution delivery from design through implementation and optimization
\n
Cross-Functional Collaboration
\n
• Partner with engineering and product teams to incorporate customer feedback into platform improvements
\n
• Build relationships with vendors across GPU, networking, and storage ecosystems
\n
• Collaborate with internal teams to refine HPC solution offerings
\n
Innovation & Technical Leadership
\n
• Stay current on HPC technologies including GPUs, accelerators, interconnects, and orchestration frameworks
\n
• Represent the organization in customer workshops, solution reviews, and technical presentations
\n
• Contribute to best practices, reference architectures, and reusable design patterns
\n
Required Experience
\n
• Proven experience in HPC solution architecture, systems integration, or large-scale distributed systems design
\n
• Strong expertise across:
\n
- \n
- GPU and CPU architectures (CUDA, NVIDIA ecosystem)
- Workload schedulers such as Slurm and Kubernetes
- High-performance networking (InfiniBand, RDMA, RoCE)
- Distributed storage systems (Lustre, GPFS, Ceph, VAST)
- Kubernetes and container orchestration for HPC workloads
- Security integration including identity, encryption, and compliance
\n
\n
\n
\n
\n
\n
\n
• Strong Linux systems knowledge including tuning, performance profiling, and system-level optimization
\n
• Ability to translate customer requirements into detailed architecture and integration plans
\n
• Strong communication skills with experience leading workshops, technical reviews, and customer engagements
\n
• Experience working cross-functionally with engineering, product, and operations teams
\n
• Ability to present complex technical solutions to both technical and executive audiences
\n
Preferred Experience
\n
• Experience delivering HPC or AI/ML workloads from design through deployment and optimization
\n
• Familiarity with containerized HPC environments (e.g., Kubernetes, Singularity)
\n
• Experience with automation and infrastructure deployment practices
\n
• Background in proof-of-concept delivery, benchmarking, and workload migration
\n
• Awareness of emerging HPC technologies including next-generation GPUs and interconnects
\n
• Bachelor’s or Master’s degree in Computer Science, Engineering, Physics, or related field
\n
• Relevant certifications such as AWS Solutions Architect, Azure Solutions Architect Expert, GCP Professional Cloud Architect, CCNP, RHCE, CKA, or CKS
