hirq
← All jobs

jobgether

GPU Cluster Architect

US · Remote · Full-time · IT

Apply well, not just fast

Create a free account and upload your resume to get a match score, keyword gaps, a tailored resume, a cover letter and interview prep for this job.

About the role

Distributed SystemsSRELLMs
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a GPU Cluster Architect based in the United States. This is a remote, high-impact architecture role focused on designing next-generation AI infrastructure at massive scale. You’ll make end-to-end architectural decisions spanning GPU compute, high-performance networking, storage, reliability, and control planes. Your work will shape how tens of thousands of GPUs are interconnected, powered, cooled, monitored, and optimized across multiple data center sites. You’ll model demanding AI and machine-learning workloads, including large language model training and inference, to guide critical performance and infrastructure tradeoffs. The role combines deep systems expertise with hands-on collaboration across networking, storage, site reliability, and data center engineering teams. You’ll operate in a fast-moving, engineering-led environment where scalability, performance, reliability, and innovation are central to the work. This is an opportunity to influence the architecture of large-scale AI infrastructure while solving complex distributed systems challenges.