Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Designworks Talent LLC

Staff Data Center Infrastructure Software Engineer

Designworks Talent LLC

. Lead the design, development, configuration, and automation of AI infrastructure clusters .

Posted 10/2/2026full-timeBellevue • Washington • United StatesLeadWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in designing and operating large-scale Linux-based infrastructure, with a strong focus on Kubernetes, containerization, and automation tools. Capable of optimizing performance for AI workloads and ensuring operational excellence in mission-critical environments.

Highest-signal resume keywords
Kubernetes DeploymentInfrastructure-As-CodeAutomation ToolsLarge-Scale Linux InfrastructureGPU Cluster Management

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Infrastructure-As-CodeAutomationKubernetesContainerizationDistributed SystemsLinuxCloud InfrastructureBare-Metal ProvisioningHardware Lifecycle ManagementPerformance Optimization
Tools & Technologies
TerraformAnsibleMAASIronicXCATIPMIRedfishPXE Boot
Industry Keywords
AI InfrastructureOperational ExcellenceMission-Critical InfrastructureHybrid Work Arrangement

Tech Stack

Tools & technologies
AnsibleCloudDistributed SystemsKubernetesLinuxTerraform

About the role

Key responsibilities & impact
  • Lead the design, development, configuration, and automation of AI infrastructure clusters
  • Develop infrastructure-as-code, automation, and provisioning systems for compute, networking, and storage
  • Deploy and optimize Kubernetes, container, and distributed computing platforms
  • Optimize GPU, networking, storage, and system performance for large-scale AI workloads
  • Troubleshoot complex issues across hardware, operating systems, networking, storage, and software stacks
  • Build reliability, observability, and operational excellence practices for mission-critical infrastructure
  • Support the platform from server and rack installation through software deployment, networking, configuration, cluster bring-up, and automation
  • Help make the platform operational and ready for customer workloads
  • Shape architecture, tooling, and culture as an early engineer on the team

Requirements

What you’ll need
  • 5+ years of experience designing, building, or operating large-scale Linux-based infrastructure
  • Hands-on experience with Kubernetes, containerization, and distributed systems in production environments
  • Experience with infrastructure-as-code and automation tools such as Terraform, Ansible, or similar frameworks
  • Strong experience operating cloud or datacenter-scale infrastructure
  • Preferred: experience with bare-metal provisioning and hardware lifecycle management platforms such as MAAS, Ironic, or xCAT
  • Preferred: experience with IPMI, Redfish, PXE boot, and automated operating system deployment at scale
  • Preferred: experience managing GPU clusters in datacenter or cloud environments
  • U.S. work authorization is required
  • Visa sponsorship is not currently available
  • Willingness to work in a hybrid arrangement with a minimum of three days per week in the office

Benefits

Comp & perks
  • Certain roles are eligible for merit increases, annual bonus, and long term incentives based on individual performance
  • Medical, dental, and vision insurance for U.S.-based employees
  • 401(k) plan and company match
  • Paid holidays per calendar year
  • Hybrid work arrangement