FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and scaling Retrieval-Augmented Generation (RAG) pipelines, managing vector databases, and developing data pipelines using tools like Glue and Databricks. Proficient in data mining, ETL processes, and ensuring data quality while effectively communicating technical requirements to diverse stakeholders.
Highest-signal resume keywords
Retrieval-Augmented Generation (RAG) PipelinesData Pipeline DevelopmentVector Database ManagementData Mining and ETL ProcessesRelational and NoSQL Databases
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data MiningExtract-Transform-Load (ETL)Data Pipeline DevelopmentVector Database ManagementKnowledge GraphsSemantic LayersData Quality ResolutionCI/CD PipelinesVector EmbeddingsAutomated Data Guardrails
Soft Skills
Problem-SolvingAnalytical SkillsCritical ThinkingAttention to DetailCommunication Skills
Tools & Technologies
PineconeMilvusWeaviateGlueDatabricksSynapseDataprocPostgreSQLDB2MongoDB
Tech Stack
Tools & technologiesETLMongoDBNoSQLPostgres
About the role
Key responsibilities & impact- Design and scale Retrieval-Augmented Generation (RAG) pipelines
- Transform unstructured IT logs and documentation into optimized vector embeddings
- Manage the health and performance of vector databases such as Pinecone, Milvus, or Weaviate
- Build knowledge graphs and semantic layers
- Create automated data guardrails to detect noise, bias, and personally identifiable information
- Identify and resolve data quality issues at the source
- Build, deploy, and maintain CI/CD pipelines for data infrastructure
- Ensure data context remains fresh and reliable
Requirements
What you’ll need- Expertise in data mining, data storage, and Extract-Transform-Load (ETL) processes
- Experience in data pipeline development and tooling, such as Glue, Databricks, Synapse, or Dataproc
- Experience with relational and NoSQL databases, including PostgreSQL, DB2, and MongoDB
- Excellent problem-solving, analytical, and critical thinking skills
- Ability to manage multiple projects simultaneously while maintaining attention to detail
- Ability to communicate with technical and non-technical colleagues and translate business needs into technical requirements
Benefits
Comp & perks- Flexible, supportive environment
- Hands-on experience
- Learning opportunities
- Opportunities to certify in four major platforms
- Hybrid-friendly culture
- Be Well programs supporting financial, mental, physical, and social health
- Personalized development goals
- Continuous feedback
- Certifications with Microsoft, Google, and Amazon
- Coaching and hands-on experiences
- Career development and cutting-edge learning opportunities
