FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates extensive experience in building and operating large-scale distributed platforms using Java and Python, with a strong focus on system design, microservices architecture, and operational excellence. Proficient in establishing engineering standards and mentoring teams while driving innovation in AI and cloud-native technologies.
Highest-signal resume keywords
Java ProficiencyPython ProficiencyMicroservices ArchitectureKubernetes OrchestrationOperational Excellence
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Distributed SystemsAPI DesignDomain-Oriented DesignLow-Level DesignCI/CDSparkMLOps PracticesObservabilityIncident ManagementData Protection
Soft Skills
Ownership MindsetTechnical JudgmentProblem-Solving
Tools & Technologies
KubernetesHadoopHDFSHiveKafkaGenAILLM Tooling
Certifications & Qualifications
Bachelor's Degree in Computer ScienceMaster's Degree in Engineering
Industry Keywords
Cloud-Native InfrastructureEvent-Driven SystemsService ContractsFeature StoresModel Lifecycle
Tech Stack
Tools & technologiesCloudDistributed SystemsHadoopHDFSJavaKafkaKubernetesMicroservicesPythonSpark
About the role
Key responsibilities & impact- Design, build, and operate distributed platform services and APIs using Java and Python on Kubernetes or equivalent cloud-native infrastructure
- Apply domain-oriented design, system design, and low-level design rigor to ambiguous platform problems
- Contribute to Spark-based data pipelines and processing jobs when needed
- Build AI, GenAI, and agentic-powered platform capabilities, including developer tooling, workflow automation, agentic services, and AI-assisted operational tooling
- Own platform reliability, scalability, and observability, including SLOs, distributed tracing, alerting, and incident response
- Guide architecture and system design decisions across microservices, event-driven systems, and enterprise integrations
- Partner with data engineers, data scientists, and ML engineers on platform primitives for data pipelines, feature stores, and model-serving paths
- Establish engineering standards through architecture, design, code, and production readiness reviews
- Set technical direction and best practices across multiple teams and mentor engineers without a management title
- Evaluate and adopt emerging platform, cloud, and AI/agentic technologies
- Build and operate foundational data infrastructure powering AI, analytics, and decision-making across Walmart
Requirements
What you’ll need- 12+ years of professional software engineering experience building and operating large-scale distributed platforms
- Strong proficiency in Java and Python
- Operational Excellence and Engineering Excellence mindset, including SLO/SLA ownership, incident management, blameless post-incident reviews, reliability and quality metrics, and continuous improvement
- Bachelor's or Master's degree in Computer Science, Engineering, or related technical field, or equivalent practical experience
- Deep experience with microservices architectures, distributed systems, event-driven patterns, API design, service contracts, and enterprise system integration
- Strong system design, domain-oriented design (DDD), and low-level design (LLD) skills
- Experience with CI/CD, containerization, Kubernetes or equivalent orchestration, infrastructure automation, and production operations
- Spark required at a working level; familiarity with Hadoop/HDFS, Hive, or Kafka
- Hands-on experience integrating model APIs and GenAI/LLM tooling
- Experience building or operating agentic workflows, including tool-calling agents and multi-step autonomous systems
- Experience with embeddings/vector search and MLOps practices, including feature stores, model lifecycle, and retraining pipelines
- Knowledge of observability, distributed tracing, logging, metrics, alerting, incident response, capacity planning, and fault tolerance
- Knowledge of authentication, authorization, data protection, and secure API design
- Ability to translate ambiguous technical/business problems into clear execution plans
- Strong ownership mindset and technical judgment; comfortable operating with high autonomy in a 0-to-1 environment
- Must work from the Chennai office for daily work
Benefits
Comp & perks- Incentive awards for performance
- Maternity and parental leave
- PTO
- Health benefits
