FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in GPU node and cluster management, including installation, configuration, and maintenance of GPU drivers and system software. Proficient in automation and scripting with Python, Bash, and Ansible, alongside strong troubleshooting and problem-solving capabilities.
Highest-signal resume keywords
GPU Cluster ManagementLinux Systems AdministrationPython ScriptingAnsible AutomationTroubleshooting Skills
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
GPU Functionality ValidationAutomated TestingSystem Software MaintenanceFirmware InstallationCluster Health Monitoring
Soft Skills
Problem-Solving SkillsClear Communication
Tools & Technologies
NVIDIA CUDAAMD ROCmBashAnsible
Industry Keywords
Systems EngineeringInfrastructureHigh-Performance ComputingAI Workloads
Tech Stack
Tools & technologiesAnsibleLinuxNode.jsPython
About the role
Key responsibilities & impact- Execute GPU node and cluster bring-up processes across new hardware platforms
- Validate GPU node functionality through automated and manual testing
- Troubleshoot hardware and software issues across GPU and OS layers
- Contribute to automation using Python, Bash, and Ansible
- Support onboarding and offboarding of systems within GPU clusters
- Document processes, findings, and improvements for internal teams
- Install, configure, and maintain GPU drivers, firmware, and system software
- Monitor GPU cluster health and respond to hardware alerts and failures
- Help build and scale high-performance GPU infrastructure powering next-generation AI workloads
Requirements
What you’ll need- 2–5 years of experience in systems engineering, infrastructure, or similar roles
- Hands-on experience with Linux systems and server hardware
- Familiarity with GPUs (NVIDIA and/or AMD) in production or lab environments
- Experience with scripting (Python, Bash)
- Exposure to automation tools such as Ansible
- Strong troubleshooting and problem-solving skills
- Ability to work across teams and communicate technical concepts clearly
- Familiarity with GPU drivers and system software (NVIDIA CUDA, AMD ROCm)
- Must be located in the United States
- Must be legally authorized to work in the United States
- Must not require employment visa sponsorship
Benefits
Comp & perks- 100% company-paid insurance premiums for employee medical, dental and vision plans
- 401(k) plan that matches 100% up to 4%, with immediate vesting
- Professional Development Reimbursement of $2,500 each year
- 11 Holidays + Paid Time Off Accrual + Rollover Plan
- Increased PTO at 3 year and 10 year anniversary
- 1 month paid sabbatical every 5 years
- Anniversary Bonus each year
- $500 stipend for remote office setup in first year + $400 each following year
- Internet reimbursement up to $75 per month
- Gym membership reimbursement up to $50 per month
- Company paid Wellable subscription
