FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Staff Data Center Infrastructure Software Engineer
Designworks Talent LLC. Lead the design, development, configuration, and automation of AI infrastructure clusters .
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and operating large-scale Linux-based infrastructure, with a strong focus on Kubernetes, containerization, and automation tools. Capable of optimizing performance for AI workloads and ensuring operational excellence in mission-critical environments.
Highest-signal resume keywords
Kubernetes DeploymentInfrastructure-As-CodeAutomation ToolsLarge-Scale Linux InfrastructureGPU Cluster Management
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Infrastructure-As-CodeAutomationKubernetesContainerizationDistributed SystemsLinuxCloud InfrastructureBare-Metal ProvisioningHardware Lifecycle ManagementPerformance Optimization
Tools & Technologies
TerraformAnsibleMAASIronicXCATIPMIRedfishPXE Boot
Industry Keywords
AI InfrastructureOperational ExcellenceMission-Critical InfrastructureHybrid Work Arrangement
Tech Stack
Tools & technologiesAnsibleCloudDistributed SystemsKubernetesLinuxTerraform
About the role
Key responsibilities & impact- Lead the design, development, configuration, and automation of AI infrastructure clusters
- Develop infrastructure-as-code, automation, and provisioning systems for compute, networking, and storage
- Deploy and optimize Kubernetes, container, and distributed computing platforms
- Optimize GPU, networking, storage, and system performance for large-scale AI workloads
- Troubleshoot complex issues across hardware, operating systems, networking, storage, and software stacks
- Build reliability, observability, and operational excellence practices for mission-critical infrastructure
- Support the platform from server and rack installation through software deployment, networking, configuration, cluster bring-up, and automation
- Help make the platform operational and ready for customer workloads
- Shape architecture, tooling, and culture as an early engineer on the team
Requirements
What you’ll need- 5+ years of experience designing, building, or operating large-scale Linux-based infrastructure
- Hands-on experience with Kubernetes, containerization, and distributed systems in production environments
- Experience with infrastructure-as-code and automation tools such as Terraform, Ansible, or similar frameworks
- Strong experience operating cloud or datacenter-scale infrastructure
- Preferred: experience with bare-metal provisioning and hardware lifecycle management platforms such as MAAS, Ironic, or xCAT
- Preferred: experience with IPMI, Redfish, PXE boot, and automated operating system deployment at scale
- Preferred: experience managing GPU clusters in datacenter or cloud environments
- U.S. work authorization is required
- Visa sponsorship is not currently available
- Willingness to work in a hybrid arrangement with a minimum of three days per week in the office
Benefits
Comp & perks- Certain roles are eligible for merit increases, annual bonus, and long term incentives based on individual performance
- Medical, dental, and vision insurance for U.S.-based employees
- 401(k) plan and company match
- Paid holidays per calendar year
- Hybrid work arrangement