Remote Jobs RockRemote Jobs Rock

Senior Solution Architect, AI Infrastructure

๐Ÿ•’ 2 days ago
AI InfrastructureAccelerated ComputingGPU TechnologyNetworking

๐Ÿ“œ Description

  • Work with NVIDIA Cloud Partners and OEMs on large data center GPU server and networking system deployments.
  • Guide customer discussions on network design, compute/storage, and support bring up of server/network/cluster deployments.
  • Become the primary technical driver for customers during the build, deployment, and production of GPU infrastructure.
  • Conduct regular technical meetings for product roadmap, cluster issue debugging, and feature discussions.
  • Analyze and debug compute/network configuration and performance issues to deliver performant clusters.

๐Ÿ› ๏ธ Requirements

  • BS/MS/PhD in Electrical Engineering, Computer Science or equivalent experience.
  • 6+ years supporting Solution Engineering or similar roles with direct customer interaction.
  • Experience with high performance Networking or CPU/GPU application acceleration.
  • Familiarity with schedulers such as SLURM, LSF, UGE, etc.
  • Experience with benchmarking tools such as HPL, NCCL tests, MLPerf, and Kubernetes.
  • Ability to travel to customer sites up to 20% of the time.
  • Experience with bring up and deployment of large GPU clusters and optimizing high-speed networks.
  • Active Security Clearance is highly desirable.
Full job description

NVIDIA is looking for a Senior AI Infrastructure Solutions Architect to join its Public Sector Continuous Bring Up team This person will be passionate about crafting an entire market and serving Public Sector customers during the AI transformation. The ideal candidate will have a strong technical background in accelerated computing technology and artificial intelligence. They will apply these skills to support government programs within the Public Sector market. They will collaborate closely with product, engineering, and customers to accelerate NVIDIA technology in the design to deployment of large-scale GPU infrastructure. A successful candidate will demonstrate skill in transcending boundaries, working effectively with product teams, account managers, field organization, and customers to drive success. You must be a U.S. Citizen to apply for this position. What youโ€™ll be doing: Working with NVIDIA Cloud Partners and OEMs in Public Sector on large data center GPU server and networking system deployments. Guide customer discussions on network design, compute/storage, and support bring up of server/network/cluster deployments. You will need to visit customer data center during bring up phase. Become the primary technical driver for customers during the build, deployment, construction, integration, and production of GPU infrastructure and applications throughout the entire customer lifecycle. Work as the customer's trusted advisor conducting regular technical customer meetings for product roadmap, cluster issue debugging, feature discussions and introduction to new technology solutions. Partner with other SAs, Account Managers, Engineering, Product, and business leaders to align on strategies, assess technical needs, and secure business opportunities for NVIDIA. Analyze and debug compute/network configuration and performance issues to deliver performant clusters. Prepare and deliver technical content to customers including presentations, workshops, reference architectures, tutorials, publications. Lead communication with customers and NVIDIA Management. Provide constructive feedback to engineering and product regarding product requirements, customer experience, documentation, and tools. What we need to see: The position requires solving complex multidisciplinary problems. Responsible for leading the resolution of technical issues across multiple engineering teams and coordinating the solutions with the customer. BS/MS/PhD in Electrical Engineering, Computer Science or equivalent experience. 6+ years supporting Solution Engineering (or similar Sales Engineering, Solution Architecture) including experience working directly with partners and customers. Experience with high performance Networking or CPU/GPU application acceleration. Background with schedulers such as SLURM, LSF, UGE, etc. Experience with benchmarking tools such as HPL, NCCL tests, MLPerf as well as Kubernetes experience. An ability to travel to customer sites up to 20% of the time. Ways to stand out from the crowd: Familiarity with NVIDIA GPUs, NVIDIA Networking technologies (e.g. NICs, RoCE, InfiniBand), and systems technology such as NCCL, DCGM, UFM, Mission Control, and Base Command Manager. Experience building and/or integration artificial intelligence solutions. Experience with bring up and deployment of large GPU clusters, including deploying and optimizing high-speed networks (InfiniBand/Ethernet), with a clear understanding of how network architecture impacts GPU cluster performance. Experience with MPI (Message Passing Interface). Experience working with enterprise developers and strong customer-facing skills. Active Security Clearance is highly desirable. NVIDIA is widely considered to be one of the technology worldโ€™s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative, independent, and focused on serving the mission of the U.S. Federal Government, we want to hear from you! Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until September 15, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law. NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA.

LangChain

Deployed Architect, Professional Services (Austin)

LangChain๐Ÿ‘ฅ 1001 - 5000 employees๐Ÿข Technology, Information And Internet๐Ÿค B2B
๐Ÿ•’ 9 days ago

As a Deployed Architect, you'll design and optimize production-grade AI infrastructure and agent systems, directly impacting customer success and shaping best practices.

KubernetesTerraformHelmAI/ML Applications
LangChain

Deployed Architect, Professional Services (Dallas)

LangChain๐Ÿ‘ฅ 1001 - 5000 employees๐Ÿข Technology, Information And Internet๐Ÿค B2B
๐Ÿ•’ 9 days ago

As a Deployed Architect, you'll design, deploy, and optimize AI infrastructure and agent systems for enterprise customers, combining software development and customer-facing skills.

KubernetesTerraformHelmAI/ML Applications
LangChain

Deployed Architect, Professional Services (Remote)

LangChain๐Ÿ‘ฅ 1001 - 5000 employees๐Ÿข Technology, Information And Internet๐Ÿค B2B
๐Ÿ•’ 9 days ago

As a Deployed Architect, you will design, deploy, and optimize AI infrastructure and agent systems for enterprise customers, ensuring scalable and secure solutions.

KubernetesTerraformHelmGCP
Nvidia

Senior Solutions Architect, Networking Solutions

Nvidia๐Ÿ‘ฅ 10,000+ employees๐Ÿข Computer Hardware
๐Ÿ•’ 2 days ago

Join NVIDIA as a Senior Solutions Architect to drive innovative networking solutions for cloud, HPC, and enterprise customers while collaborating with sales and engineering teams.

Networking SolutionsSales EngineeringCloud ApplicationsHPC

Trusted by Remote Workers