Network Infrastructure Architect
About the role
SambaNova is accelerating from rack-scale to global-scale AI compute clusters. We're looking for a Network Architect join us on this exciting new chapter and own this critical layer end to end.
Responsibilities
div]:bg-bg-000/50 [&_pre>div]:border-0.5 [&_pre>div]:border-border-400 [&_.ignore-pre-bg>div]:bg-transparent [&_.standard-markdown_:is(p,blockquote,h1,h2,h3,h4,h5,h6)]:pl-2 [&_.standard-markdown_:is(p,blockquote,ul,ol,h1,h2,h3,h4,h5,h6)]:pr-8 [&_.progressive-markdown_:is(p,blockquote,h1,h2,h3,h4,h5,h6)]:pl-2 [&_.progressive-markdown_:is(p,blockquote,ul,ol,h1,h2,h3,h4,h5,h6)]:pr-8">
_*]:min-w-0 gap-3 [&_>_*:last-child]:mb-0 print:block print:[&_>_*_+_*]:mt-3 standard-markdown">
In this role, you'll be architecting the network fabrics that power hyperscale AI clusters, from RDU-accelerated servers to global-scale deployments, owning the fabric, interconnect, and spine-leaf design decisions that shape how our compute platforms are built. You'll partner closely with hardware engineering, supply chain, and customer-facing teams to translate workload requirements into platform roadmaps, taking a vendor-agnostic approach that avoids lock-in. Your architecture decisions will directly shape the products we ship and how fast the world's largest AI clusters come online.
Required Qualifications
12+ years designing or architecting infrastructure for hyperscale, AI, HPC, or large-scale data center environments
Deep working knowledge of AI cluster networking, including InfiniBand, RoCEv2, Ethernet-based AI fabrics, 400G/800G interconnects, and GPU-to-GPU east-west traffic patterns
Experience with lossless Ethernet designs, including QoS, ECN, PFC, congestion management, buffer tuning, and failure-domain isolation
Strong understanding of spine-leaf architectures and associated technologies (VXLAN, EVPN, BGP, ECMP, underlay/overlay design, network automation)
Familiarity with high-speed optics and cabling for AI clusters, including 400G/800G transceivers, DAC/AOC, fiber topology, and link budgets
Experience translating customer and workload requirements into product requirements, platform roadmaps, and architecture tradeoffs
Vendor-agnostic approach to platform selection and interoperability across switch, NIC, optical, and fabric ecosystems
Ability to partner across hardware engineering, supply chain, operations, and customer-facing teams to bring platforms from concept through deployment
Preferred Qualifications
Experience with AI compute platforms: GPU/accelerator systems, rack-scale compute, NVLink/NVSwitch-class fabrics, PCIe/CXL, DPUs/IPUs
Understanding of rack-level power, thermal, mechanical, and serviceability requirements, and how network choices drive them (NIC selection, cabling density, PCIe lane allocation)
Background in large-scale fleet operations: lifecycle management, reliability, observability, telemetry, field issue resolution
Base Salary Range:
Base Pay Range
$250,000—$330,000 USD