FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Customer Reliability Engineer, Infrastructure
Astronomer. Provide solutions to customers using Astronomer's products .
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in managing and troubleshooting cloud-native distributed systems, with a strong focus on customer experience and operational efficiency. Proficient in automation, monitoring, and providing feedback to product development teams to enhance solutions.
Highest-signal resume keywords
Kubernetes ManagementCloud Infrastructure ExperienceLinux ProficiencyPython ScriptingDevOps Practices
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
KubernetesCloud InfrastructureLinuxObservability ToolsPython ScriptingCI/CDDistributed Systems ManagementAutomationMonitoring SystemsSite Reliability Engineering
Soft Skills
Strong Communication Skills
Tools & Technologies
AWSGCPAzureAirflowBig Data OrchestrationInfrastructure as Code
Industry Keywords
Cloud-NativeCustomer ExperienceOperational TasksMonitoringTriage Issues
Tech Stack
Tools & technologiesAirflowAWSAzureCloudDistributed SystemsGoogle Cloud PlatformKubernetesLinuxPython
About the role
Key responsibilities & impact- Provide solutions to customers using Astronomer's products
- Troubleshoot customer environments and actively triage issues with customers
- Provide product-development teams with feedback on customer needs and pain points
- Build monitoring and alerting systems
- Build and maintain automation for efficient daily operational tasks
- Help direct product architecture and contribute to architectural improvements
- Own the customer experience by prioritizing and solving issues, meeting SLAs, and providing guidance through production
- Participate remotely within a fully distributed team
- Enhance and enrich customer documentation
- Work on a cloud-native product connecting customers to dozens of other systems
- Help maintain 24x7 coverage through a specified 6-hour pager period during the work day
- Participate in paid on-call rotation for weekend coverage
Requirements
What you’ll need- 3+ years of experience, preferably with large, complex cloud infrastructures operating at scale
- 2+ years of experience with Kubernetes
- Experience managing a production distributed system with at least one major cloud provider: AWS, GCP, or Azure
- Strong network experience with one of the major clouds
- Strong Linux experience
- Knowledge of operating and monitoring distributed systems
- Experience with observability tools
- Previous experience handling customer issues, internal and external
- Strong communication skills
- DevOps or CI/CD experience
- Python scripting
- Bonus: experience as a Site Reliability Engineer
- Bonus: experience with Kubernetes Custom Resources
- Bonus: Airflow/Big Data Orchestration experience
- Bonus: IaC experience
- Willingness to work at least 3 days per week at the Hyderabad office
- Must be located in Hyderabad or willing to relocate to Hyderabad
Benefits
Comp & perks- Paid on-call rotation for weekend coverage
- Flexible hybrid work model
- Participation in a fully distributed team
- 24x7 coverage through a specified 6-hour pager period during the work day