FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in network solutioning for GPU clouds and AI factories, with a strong focus on designing and implementing complex data center networking architectures. Proficient in Kubernetes networking, InfiniBand, and cloud infrastructure, with a proven ability to collaborate with cross-functional teams and present technical concepts effectively.
Highest-signal resume keywords
Network SolutioningData Center Network DesignKubernetes NetworkingInfiniBand ExpertiseCustomer-Facing Experience
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Network ArchitectureInfiniBand Subnet ManagementLinux NetworkingKubernetes NetworkingProgramming in PythonProgramming in GoData Center NetworkingNCCL/RDMA TrafficOversubscription DesignMulti-Cloud Connectivity
Soft Skills
Excellent CommunicationPresentation SkillsCollaborationCustomer EngagementWorkshop Facilitation
Tools & Technologies
GitCIInfrastructure-as-CodeAnsibleTerraformKubernetesNVIDIA Network OperatorDRABashHelm
Industry Keywords
GPU CloudsAI FactoriesHPC NetworkingSpine-Leaf TopologyClos TopologyRail-Optimised FabricsVPNsHybrid CloudSDNOpen-Source Contributions
Tech Stack
Tools & technologiesAnsibleCloudDNSKubernetesLinuxPythonTerraformGo
About the role
Key responsibilities & impact- Own network solutioning for k0rdent AI, Mirantis' platform for building and operating GPU clouds and AI factories
- Design interconnects for large GPU clusters, tenant isolation and workload network access from containers, virtual machines and bare-metal nodes
- Design and publish network reference architectures and solution designs from single-rack to multi-thousand-GPU clusters
- Define InfiniBand and RoCEv2 Ethernet compute fabrics, including rail-optimised and fat-tree/Clos designs, oversubscription and failure domains
- Define front-end, storage and management networks and their connectivity to customer data centers, public clouds and hybrid environments
- Specify multi-tenant isolation using InfiniBand PKeys, VRFs, EVPN-VXLAN and Kubernetes network policy
- Document scale limits and trade-offs involving cost, performance, operability and vendor lock-in
- Define Linux host networking for NICs, DPUs and SuperNICs, including PCI passthrough, SR-IOV, IOMMU, NUMA and GPU-NIC affinity
- Design Kubernetes networking for AI workloads using CNIs, Multus, SR-IOV, RDMA plugins, NVIDIA Network Operator and DRA
- Design KubeVirt networking for GPU and RDMA traffic
- Research and evaluate emerging AI networking technologies and standards
- Prototype and benchmark alternative designs in the lab and compare them with current practice
- Publish research notes, design proposals, papers, blog posts and talks
- Build and run proofs of concept with customers, partners and on lab/customer hardware
- Write automation and tooling using Python, Go, Bash, Ansible, Helm, Kubernetes manifests and Terraform
- Validate and benchmark fabrics and host configurations using NCCL tests, perftest and ib_write_bw
- Act as network subject matter expert in customer discovery, design reviews and architecture workshops
- Collaborate with hardware/networking partners and cross-functional Solutions Architecture, Partner Management, Product and Engineering teams
- Present at industry events, webinars and partner summits
- Run hands-on technical workshops and write reference architectures, solution briefs, blog posts and enablement content
Requirements
What you’ll need- Bachelor's degree in Computer Science, Electrical Engineering, Telecommunications or a related field, or equivalent practical experience
- 8+ years in network engineering or network architecture
- At least 3 years in data center, HPC or cloud infrastructure networking
- Customer-facing experience as a solutions architect, pre-sales engineer, consultant or technical lead
- Expertise in data center network design, including spine-leaf and Clos topologies, rail-optimised GPU fabrics, oversubscription and ECMP
- Expertise in InfiniBand subnet management, partitioning, adaptive routing and NCCL/RDMA traffic
- Expertise in RoCEv2, including PFC, ECN, DCQCN, QoS, MTU and buffer tuning
- Knowledge of BGP, EVPN-VXLAN, VRFs, hybrid and multi-cloud connectivity, VPNs, IP address and DNS planning
- Knowledge of Linux networking, PCI passthrough, SR-IOV, IOMMU, NUMA affinity, iproute2, ethtool and devlink
- Knowledge of Kubernetes networking, CNI plugins, Multus, SR-IOV, RDMA device plugins, network policy and service exposure
- Knowledge of KubeVirt and VM networking, passthrough and SR-IOV
- Programming or scripting in at least one language; Python or Go preferred
- Comfortable using Git, CI and infrastructure-as-code
- Excellent written and spoken English
- Comfortable presenting to large audiences and running hands-on workshops
- Experience working in an international, distributed company across time zones and work cultures
- Willingness to work some meetings outside standard local hours
- Up to 25% travel
- Nice-to-have experience includes NVIDIA networking, GPU cloud/neocloud/HPC deployments, bare-metal provisioning, network automation, SDN, GSLB, NVMe-oF, Mirantis products, open-source contributions, standards participation and additional languages
Benefits
Comp & perks- Travel of up to 25% for customer engagements, partner meetings, lab work and industry events
- Compensation and benefits set according to the local market and employment arrangement
- Work with a distributed team committed to openness and technical excellence
- Opportunity to shape the product narrative and influence go-to-market success
