Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Supermicro

Network Engineer

Supermicro

. Manage the rollout and maintenance of business-critical applications and services .

Posted 10/6/2026full-timeSan Jose • California • United StatesMid-LevelSenior💰 $120,000 - $140,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in managing HPC/AI applications, including testing, optimization, and customer support. Proficient in utilizing AI/ML frameworks and tools, with a strong foundation in Linux and networking environments.

Highest-signal resume keywords
HPC/AI Application ManagementDeep Learning and Machine LearningLinux/Networking DebuggingAI/ML Frameworks (PyTorch, TensorFlow, ONNX)DevOps and Cloud Environments (Docker, Kubernetes)

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
System-Level Rack TestingBIOS Settings OptimizationShell Scripting (Windows and Linux)Workload/Scheduler Management (Slurm)Benchmarking (MLPerf, HPL-AI)Technical Documentation CreationProof-of-Concept Design and TestingCustomer Acceptance VerificationAutomation ScriptingPerformance Testing
Soft Skills
TeamworkCommunication
Tools & Technologies
NVIDIA Development Toolkits (CUDA, oneAPI, ROCm)DockerKubernetesOpenStackOpenShiftAzureAWS
Certifications & Qualifications
CCNA
Industry Keywords
HPCAIMachine LearningDeep LearningCloud ComputingNetworkingDebuggingPerformance OptimizationCustomer SupportTechnical Training

Tech Stack

Tools & technologies
AWSAzureCloudDockerKubernetesLinuxOpenShiftOpenStackPyTorchShell ScriptingTensorflow

About the role

Key responsibilities & impact
  • Manage the rollout and maintenance of business-critical applications and services
  • Resolve escalated service issues
  • Execute comprehensive system-level rack tests on NVIDIA and AMD GPUs, ARM-based, Intel Xeon and AMD EPYC processors
  • Test functionality, compatibility, performance, stress and reliability using proprietary in-house tools
  • Establish knowledge of HPC/AI applications and benchmarks
  • Deliver training sessions to customers and partners
  • Address complex customer support issues and build processes for HPC/AI solutions
  • Conduct proof-of-concept design and testing
  • Provide optimized benchmarks for HPC/AI applications
  • Fine-tune BIOS settings and optimize OS/network configurations
  • Develop simulation configurations for various workloads
  • Deliver on-site deployment services and customer acceptance verification
  • Provide post-level 1 and 2 support
  • Create and maintain technical notes, blogs, diagrams and other documentation
  • Identify and document hardware and software quality issues
  • Collaborate with Product Management and Engineering teams on product enhancements
  • Contribute to HPC roadmap development and plan software and hardware upgrades
  • Document and analyze test plans, reports and logs
  • Contribute to test utilities and automation scripts

Requirements

What you’ll need
  • BS/MS in Electrical Engineering, Computer Engineering or Computer Science
  • 5+ years of work-related experience in Deep Learning and Machine Learning
  • 5+ years of Linux/networking debugging/testing or relevant experience preferred
  • Experience with leading AI/ML frameworks such as PyTorch, TensorFlow, ONNX, etc.
  • Experience with DevOps or cloud environments, including Docker/Containers and Kubernetes
  • Hands-on experience with workload/scheduler managers such as Slurm for rack/cluster
  • Familiarity with MLPerf Training/Inference benchmark, LLM, HPL-AI or RCCL/NCCL
  • Programming experience with Windows and Linux shell scripting
  • Strong teamwork and communication skills
  • Familiarity with Intel/AMD/NVIDIA development toolkits such as CUDA, oneAPI, ROCm is a plus
  • Experience with server/network hardware debugging and troubleshooting is a plus
  • CCNA, OpenStack, OpenShift, Azure or AWS is a plus

Benefits

Comp & perks
  • Comprehensive benefits package
  • Potential bonus program participation
  • Potential equity award program participation