Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
General Motors

Summer Intern, Data Scaling, Embodied AI

General Motors

. Develop data pipelines and tooling for large-scale, multimodal autonomous driving datasets .

Posted 10/9/2026internshipSunnyvale • California • United StatesEntry Level💰 $11,800 - $14,600 per monthWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in developing data pipelines and tooling for autonomous driving datasets, with strong programming skills in Python and experience in machine learning frameworks such as PyTorch and TensorFlow. Capable of analyzing data quality and performance while collaborating effectively in cross-functional teams.

Highest-signal resume keywords
Python ProgrammingMachine Learning FrameworksData EngineeringDistributed SystemsQuantitative Analysis

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Data ProcessingMachine Learning PipelinesSQLC++Data Quality SystemsDistributed Data-Processing FrameworksHigh-Performance ComputingCloud InfrastructureData MiningExperiment Reproducibility
Soft Skills
Analytical SkillsProblem-Solving SkillsCollaborationWritten CommunicationPresentation Skills
Tools & Technologies
PyTorchTensorFlowJAXGPU-Based WorkflowsWorkflow OrchestrationVisualization ToolsDashboardsMetricsDistributed TrainingAutonomous Vehicles
Industry Keywords
Autonomous DrivingMultimodal Sensor DataRoboticsComputer VisionDeep LearningSelf-Supervised LearningImitation LearningReinforcement LearningLarge-Scale DatasetsTime-Series Data

Tech Stack

Tools & technologies
CloudDistributed SystemsPythonPyTorchSQLTensorflowC++

About the role

Key responsibilities & impact
  • Develop data pipelines and tooling for large-scale, multimodal autonomous driving datasets
  • Analyze data quality, coverage, distribution, and performance impact using quantitative methods
  • Prototype and evaluate approaches for data mining, curation, labeling, sampling, or scenario discovery
  • Improve training throughput, data loading, experiment reproducibility, or resource utilization in distributed computing environments
  • Collaborate with machine learning researchers, data engineers, infrastructure engineers, and autonomy teams
  • Build visualizations, metrics, dashboards, and evaluation workflows to communicate data and model behavior
  • Contribute to production-quality software through design reviews, code reviews, automated testing, continuous integration, and documentation
  • Present technical findings and document experiments, results, and recommendations

Requirements

What you’ll need
  • Currently enrolled in or pursuing a Master’s or Ph.D. degree in Computer Science, Machine Learning, Data Science, Electrical Engineering, or a related technical field
  • Demonstrated experience through coursework, research, academic projects, or professional work in machine learning, data engineering, distributed systems, or a related area
  • Strong programming skills in Python
  • Experience working with data processing, databases, machine learning pipelines, or large-scale datasets
  • Strong analytical and problem-solving skills, with the ability to use quantitative analysis to guide decisions
  • Ability to work collaboratively in a cross-functional, team-oriented environment
  • Strong written, verbal, and presentation skills
  • Availability to work full-time, 40 hours per week, during the internship period
  • Experience with PyTorch, TensorFlow, JAX, or another machine learning framework
  • Experience with distributed training, parallel computing, high-performance computing, cloud infrastructure, or GPU-based workflows
  • Familiarity with multimodal sensor data, autonomous vehicles, robotics, computer vision, or large-scale time-series data
  • Experience with SQL, data warehouses, workflow orchestration, data quality systems, or distributed data-processing frameworks
  • Familiarity with foundation models, self-supervised learning, imitation learning, reinforcement learning, or deep learning
  • Experience with C++ or another systems programming language
  • Must be graduating between December 2027 and June 2028
  • Intent to return to a degree program after completion of the internship

Benefits

Comp & perks
  • One-time lump sum taxable stipend payment for eligible students selected for the 2027 Student Program
  • Paid U.S. GM holidays
  • GM Family First Vehicle Discount Program
  • Potential for growth within GM based on performance and business needs
  • Intern events and opportunities to network with company leaders and peers
  • Mentorship and hands-on experience with the data and ML foundations behind autonomy