FREE ACCESS
5,000–10,000 jobs/day
See all jobs on Scoutfield
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and building LLM pipelines and AI-assisted data-to-report systems, with a strong focus on multilingual tools and evaluation methodologies. Proficient in managing inference costs and utilizing agentic coding tools for production software development.
Highest-signal resume keywords
LLM Pipeline DevelopmentPython ProgrammingCloud Platform DeploymentAgentic Coding Tools ExperienceLLM Evaluation Design
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Software DevelopmentNatural-Language AnalysisMultilingual Questionnaire ToolsTest Set ConstructionAutomated GradingRegression TestingCost ManagementBatch ProcessingData Services IntegrationProduction Monitoring
Tools & Technologies
AWSAzureGoogle CloudClaude CodeOpenAI CodexCursorAntigravityGitHub Copilot
Industry Keywords
LLM-Powered SystemsSurvey WorkflowsData-Residency RequirementsInference CostsReasoning Complexity
Tech Stack
Tools & technologiesAWSAzureCloudJavaScriptPythonTypeScriptGo
About the role
Key responsibilities & impact- Design and build LLM pipelines and agents for survey workflows
- Build AI-assisted data-to-report pipelines
- Develop natural-language analysis tools for survey data
- Develop multilingual questionnaire translation tools
- Build field support tools for interviewers and supervisors
- Build evaluation harnesses and define acceptance criteria with technical and survey experts
- Construct test sets from expert-produced ground truth and run non-inferiority comparisons
- Route work across models based on reasoning complexity, volume, cost, latency, and data-residency requirements
- Manage inference costs through token usage tracking, context limits, prompt caching, and batch processing
- Connect models to data services and full-stack applications
- Use agentic coding tools to produce, review, test, and own production software
- Monitor production systems and conduct disciplined error analysis
- Report positive and negative results
- Document methods and results for public release
- Help transfer tools to partner-country institutions
Requirements
What you’ll need- Bachelor's degree in Computer Science, Engineering, or a related field (or equivalent years experience)
- Minimum 4 years of professional software development experience, including building LLM-powered systems that real users relied on in production
- U.S. citizenship required by federal government contract
- Proficiency in Python or another modern programming language, such as JavaScript/TypeScript or Go
- Hands-on experience building with frontier-model APIs, including tool use and retrieval-augmented generation
- Experience designing LLM evaluations, including test-set construction, automated and human grading, regression testing, and ship/no-ship decisions
- Working knowledge of LLM cost, latency, and model-selection trade-offs
- Experience building and deploying on a major cloud platform, such as AWS, Azure, or Google Cloud
- Experience using agentic coding tools such as Claude Code, OpenAI Codex, Cursor, Antigravity, or GitHub Copilot agent mode
- Experience with agent frameworks and AI evaluation tooling
Benefits
Comp & perks- Remote work within the United States
- Occasional travel
- Reasonable accommodations for disability, veterans, and sincerely held religious beliefs
- Equal opportunity employment
- Confidential accommodation support during the application and employment process
