FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in incident management processes, cloud services, and infrastructure engineering, with a strong focus on automation, troubleshooting, and continuous improvement. Proficient in analyzing logs and traffic flows to enhance system performance and reliability.
Highest-signal resume keywords
Incident Management ProcessesCloud ServicesInfrastructure EngineeringScripting and Software DevelopmentXmatters / Whistler Workflow Integration
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Technology Infrastructure EngineeringTroubleshooting SkillsAutomationRoot Cause AnalysisLog AnalysisTraffic Flow UnderstandingMicro-ServicesCloud PlatformsConfiguration ManagementPerformance Optimization
Soft Skills
Clear Communication Skills
Tools & Technologies
XmattersWhistlerDevOps ToolsMonitoring ToolsSelf-Healing Capabilities
Industry Keywords
ComputeStorageNetworkMobilityVirtualization
Tech Stack
Tools & technologiesCloud
About the role
Key responsibilities & impact- Proactively maintain mission-critical infrastructure, cloud platforms, micro-services, tools, and processes
- Collaborate with CCC, TDO, SRE, DevOps, and Engineering practitioners
- Lead major incident response and orchestrate faster mitigation
- Monitor, detect, triage, troubleshoot, and resolve site issues before they impact customers and availability
- Understand the end-to-end technology stack and take corrective action to mitigate failures
- Assist Walmart store and distribution center associates with day-to-day issues related to store functions
- Support internal escalations through call functioning
- Follow SOPs to troubleshoot and resolve issues reported by Store Operations
- Analyze logs, traffic flows, dependencies, trends, and metrics
- Research and recommend incident-resolution actions; develop and maintain procedures and documentation
- Drive continuous improvement by eliminating, automating, or streamlining waste
- Build tools, automation, and self-healing capabilities with DevOps, Engineering, and SRE partners
- Perform root cause analysis and focus on immediate restoration
- Define and maintain CCC onboarding processes for new systems
- Share knowledge globally between CCC teams
- Improve alert detection and time to mitigate using recoverability tools
- Transition observability projects into the command center
- Act as a technical focal point and perform other assigned duties
Requirements
What you’ll need- Bachelor's degree in computer science, computer engineering, information systems, information technology, or related area and 3 years’ experience in technology infrastructure engineering across areas such as compute, storage, network, mobility or virtualization-related technologies; OR 5 years’ experience in technology infrastructure engineering across areas such as compute, storage, network, mobility or virtualization-related technologies
- Xmatters / Whistler workflow integration with scalability, resiliency and performance
- Flexible in shift and support hours
- Expert level understanding of incident management processes and procedures
- Deep technical understanding of core infrastructure, cloud services, platforms and micro-services
- Ability to understand and capture key data from logs at an expert level
- Ability to understand traffic flows and key dependencies between services
- Expert level troubleshooting skills using a diverse set of tools and methods
- Scripting and software development to automate and enhance existing solutions
- Ability to gather requirements and build solutions into a product
- Clear communication skills
- Preferred: Master’s degree and relevant experience in configuration management, automation for provisioning and orchestration, and developing/managing SLOs/SLIs
Benefits
Comp & perks- Incentive awards for performance
- Maternity and parental leave
- PTO
- Health benefits
- Flexibility for associates to manage their personal lives
