FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in managing and optimizing cloud infrastructure, particularly AWS and Alicloud, while leveraging AI frameworks for operational automation. Proficient in deploying and troubleshooting Kafka and Redis clusters, with a strong focus on CI/CD practices and collaboration with development teams.
Highest-signal resume keywords
Kafka OperationsRedis OperationsAWS Cloud PlatformCI/CD ToolsPython Programming
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
PythonGoJavaSQLDockerKubernetesCI/CDAIOpsAnomaly DetectionRoot-Cause Analysis
Soft Skills
Problem-SolvingTeam Collaboration
Tools & Technologies
GitHub ActionsAnsibleTerraformOpenAIDifyAgnoLangChain
Industry Keywords
Cloud InfrastructureOperational AutomationProduction EnvironmentsDevOpsChange Management
Tech Stack
Tools & technologiesAnsibleAWSAzureCloudDockerGoogle Cloud PlatformJavaKafkaKubernetesPythonRedisSQLTerraformGo
About the role
Key responsibilities & impact- Handle production incidents and conduct post-mortem analysis for system stability improvements
- Design, deploy, monitor, and troubleshoot Kafka and Redis clusters in production environments
- Work closely with development teams to ensure seamless application and system deployments
- Manage and optimize AWS and Alicloud infrastructure for performance, cost, and reliability
- Develop DevOps platforms such as online load testing and change management systems
- Leverage LLMs or AI frameworks including OpenAI, Dify, Agno, and LangChain to enhance infrastructure-operation automation
- Implement intelligent alert triage, root-cause analysis, and chat-based operations
- Integrate AI-driven insights into operational processes to improve reliability, reduce noise, and support engineering decision-making
Requirements
What you’ll need- 5+ years of hands-on experience in Kafka and Redis operations in large-scale production environments, with ability to cooperate with developers to optimize code
- Proficient in Python, Go, or Java (at least one language) and SQL
- Hands-on experience with Docker and Kubernetes
- Strong experience with CI/CD tools such as GitHub Actions, Ansible, and Terraform
- At least 3 years of experience with AWS cloud platform
- GCP, Azure, or Ali Cloud experience is a plus
- Excellent problem-solving and troubleshooting skills
- Strong team collaboration attitude and ability to develop partnerships with other teams and business
- Practical experience building or operating AIOps systems, including anomaly detection, alert correlation, automated healing, or RCA
- Familiarity with LLM-based DevOps automation, such as chat-based operations assistants or AI-driven observability workflows
- Experience using or integrating Dify, Agno, or LangChain into operational workflows
Benefits
Comp & perks- Competitive salary and company benefits
- Work-from-home arrangement (the arrangement may vary depending on the work nature of the business team)
- Opportunities for career growth and continuous learning
- Collaborate with world-class talent in a user-centric global organization with a flat structure
- Innovative, results-driven work environment
- Equal opportunity employer
