Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
GitGuardian

Junior Data Engineer

GitGuardian

. Build data foundations for AI agents through warehouse and data model design .

Posted 10/7/2026full-timeParis • FranceJuniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in data engineering, focusing on SQL and Python for data modeling and pipeline orchestration. Collaborates effectively with cross-functional teams to define business metrics and ensure data quality and accessibility.

Highest-signal resume keywords
SQLPythonData ModelingSnowflakePipeline Orchestration

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
SQLPythonData ModelingDimensional ModelingBusiness MetricsData Pipeline DevelopmentData Quality AssuranceData IngestionData DocumentationData Architecture
Soft Skills
CommunicationCollaborationProblem-SolvingAdaptabilityAttention to Detail
Tools & Technologies
SnowflakeClickHouseDagsterDockerTerraformAWSPySparkSnowpark
Industry Keywords
Data EngineeringAI AgentsFinance DataHR DataMarketing AnalyticsSales AnalyticsProduct AnalyticsHigh-Growth StartupsScale-UpsAgentic Systems

Tech Stack

Tools & technologies
AWSDockerPySparkPythonSQLTerraform

About the role

Key responsibilities & impact
  • Build data foundations for AI agents through warehouse and data model design
  • Expand warehouse coverage to Finance and HR data with rigor, access control, and confidentiality
  • Strengthen Marketing, Sales, and Product Analytics data domains on Snowflake and ClickHouse
  • Develop trustworthy data models and clearly defined business metrics
  • Write SQL and Python data models using Snowpark and PySpark
  • Own data topics end to end, from ingestion through business-user consumption
  • Collaborate with Sales, Marketing, Finance, HR, and Product stakeholders to define business metrics
  • Build and orchestrate pipelines with Dagster
  • Deploy pipelines with Docker and Terraform on AWS
  • Monitor pipelines, investigate data issues, and respond to company-wide data requests
  • Maintain code quality through reviews, tests, and documentation
  • Collaborate with AI Engineers to support internal AI tools
  • Participate in interviews involving Python coding, data architecture, communication, and reasoning

Requirements

What you’ll need
  • A first experience in data engineering, such as an internship, apprenticeship or up to about 2 years in a role
  • Strong SQL and good Python skills
  • Understanding of data modeling concepts, including dimensional modeling, fact and dimension tables, and business metrics
  • Ability to navigate an existing, complex code base and learn quickly
  • Comfort communicating with non-technical stakeholders, asking the right questions, and explaining choices
  • High standard for quality; data should be correct, tested, and documented
  • Fluent English in an international environment
  • Availability to work from the office 3 days a week
  • Experience with Snowflake and/or ClickHouse is beneficial but not required
  • Familiarity with agentic systems or LLM-based architectures applied to data pipelines is beneficial but not required
  • Experience in high-growth startups or scale-ups is beneficial but not required

Benefits

Comp & perks
  • Diverse, equitable and inclusive workforce commitment
  • Coffee with team members during the final onsite interview