FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in building and optimizing data streaming infrastructure, managing cloud environments, and implementing infrastructure-as-code practices. Proficient in monitoring and incident management, with a strong foundation in distributed systems principles.
Highest-signal resume keywords
Amazon MSK (Kafka) ExperienceTerraform for Infrastructure-as-CodePython and Shell ScriptingPostgreSQL Database ManagementCI/CD Pipeline Implementation
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data Streaming InfrastructureCloud Infrastructure ManagementDistributed Systems PrinciplesMonitoring and Alerting ImplementationRelational and NoSQL DatabasesIncident Response and Root Cause AnalysisAutomation with AnsibleContainer Orchestration with KubernetesNetworking FundamentalsSecurity Fundamentals
Soft Skills
Problem-SolvingCommunicationDocumentationOwnership in Incident Management
Tools & Technologies
AWSGoogle Cloud PlatformJenkinsGitStarburst GalaxyAWS AthenaApache KafkaConfluent KafkaCI/CD ToolsCloud Automation Tools
Industry Keywords
SREDevOpsPlatform EngineeringHigh AvailabilityFault ToleranceDisaster RecoveryCapacity PlanningCost OptimizationCloud AutomationIncident Management
Tech Stack
Tools & technologiesAnsibleApacheAWSCloudDistributed SystemsGoogle Cloud PlatformJenkinsKafkaKubernetesNoSQLPostgresPythonShell ScriptingTerraform
About the role
Key responsibilities & impact- Build, operate, and optimize data streaming infrastructure using Amazon MSK (Kafka) or Google Pub/Sub
- Design, deploy, and manage highly available database and caching platforms across multi-cloud environments
- Develop and maintain infrastructure-as-code, CI/CD pipelines, and cloud automation using Terraform and Python
- Implement monitoring, alerting, and observability for data platform services
- Partner with application development teams to troubleshoot, tune, and optimize application performance and data access layers
- Administer and optimize Starburst Galaxy and AWS Athena
- Participate in incident response, root cause analysis, and post-incident reviews
- Share on-call rotation to support mission-critical data infrastructure
- Review application source code and collaborate on targeted fixes and optimization guidance
- Evaluate emerging tools and AI-assisted automation approaches
- Contribute to capacity planning, disaster recovery, security hardening, and cost optimization
Requirements
What you’ll need- 3–5 years of software engineering or infrastructure experience, with at least 2 years in SRE, DevOps, or platform engineering operating production systems at scale
- Hands-on experience designing, deploying, and managing cloud infrastructure on AWS and/or Google Cloud Platform, including networking, identity, and security fundamentals
- Production experience operating data streaming platforms; hands-on work with Apache Kafka, including Amazon MSK or Confluent Kafka, and understanding of partitions, consumer groups, delivery semantics, and backpressure
- Production experience with relational and NoSQL databases; PostgreSQL required
- Strong Python and Shell scripting for automation and operational solutions
- Strong experience with infrastructure-as-code such as Terraform, CI/CD such as Jenkins and Git, Ansible, and container orchestration such as Kubernetes in production environments
- Experience implementing and automating monitoring, logging, and alerting for distributed systems
- Proven ability to contribute to root cause analysis for production incidents
- Strong problem-solving, communication, and documentation skills with ownership in on-call and incident management environments
- Working understanding of distributed systems principles including high availability, fault tolerance, consistency models, and disaster recovery
- Background investigation required
- Dedicated, private workspace and reliable internet with minimum speeds of 50 MB download and 5 MB upload when working from home
Benefits
Comp & perks- Comprehensive technology setup including a laptop, monitors, headset, keyboard, and mouse
- Monthly connectivity reimbursement to help offset internet costs for employees eligible to work from home
