Youl'll be designing, deploying, andoperatingthe switch fabrics and interconnects that power large-scale training and inference workloads.16th September, ****About the RoleOne of our clients are building out their network backbone that connects and orchestrates thousands of GPUs at scale.You'llbe joining a team working withcutting-edgeNvidia hardware and high-performance networking technologies that few localorganisationsoperateat this scale.The role:Design, build, andoperatehigh-performance network fabrics supporting large-scale GPU clustersDeploy and tune switch fabrics for low-latency, high-throughput AI/ML workloadsWork closely with infrastructure and platform teams tooptimisenetwork performance for distributed training and inferenceTroubleshoot and resolve complex network issues across HPC and GPU-dense environmentsPlan capacity and topology for ongoing GPU infrastructure expansionPartner with vendors (including Nvidia) on hardware qualification, firmware, and fabric designContribute to standards and best practices for network architecture as the environment scales16th September, ****About the RoleOne of our clients are building out their network backbone that connects and orchestrates thousands of GPUs at scale.You'llbe joining a team working withcutting-edgeNvidia hardware and high-performance networking technologies that few localorganisationsoperateat this scale.The role:Design, build,
andoperatehigh-performance network fabrics supporting large-scale GPU clustersDeploy and tune switch fabrics for low-latency, high-throughput AI/ML workloadsWork closely with infrastructure and platform teams tooptimisenetwork performance for distributed training and inferenceTroubleshoot and resolve complex network issues across HPC and GPU-dense environmentsPlan capacity and topology for ongoing GPU infrastructure expansionPartner with vendors (including Nvidia) on hardware qualification, firmware, and fabric designContribute to standards and best practices for network architecture as the setting scalesExperience needed:Hands-on experience withhigh performancecomputing (HPC) networking environmentsStrong understanding of GPU infrastructure and the networking demands of AI/ML workloadsExperience with switch fabric design, deployment, and operations at scaleFamiliarity with Nvidia networking technologies (e.g.InfiniBand, Spectrum-X,NVLink/NVSwitchecosystems)We'llalso consider candidates without direct HPC/AI experience if they bring:Large-scale network engineering experience from ahyperscaleror major cloud provider (e.g.AWS, Azure, Google Cloud) or equivalent big-tech environmentProven experience designing oroperatinglarge, complex switch fabrics in production — this background is scarce in the Australian market and highly valuedA track recordof operating at scale (thousands of nodes/ports) rather than traditional enterprise networking
#J-*****-Ljbffr