Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
NVIDIA

Technical Product Manager – AI Infra Resilience

NVIDIA

. Define the resilience platform and own the product roadmap and delivery for specific platform features, including common telemetry interfaces, health-check contracts, attribution hooks, observability APIs, and joint deployment experience .

Posted 9/18/2026full-timeSanta Clara • California • United StatesSeniorLead💰 $208,000 - $379,500 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in product management and solutions architecture with a focus on resilience platforms, including experience in data center operations and observability. Proven ability to engage with technical customers and translate their needs into effective product strategies.

Highest-signal resume keywords
Product ManagementData Center OperationsKubernetesOpen-Source ContributionAPI Development

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Product RoadmapTelemetry InterfacesHealth-Check ContractsObservability APIsDeveloper PlatformsSDKsAgent FrameworksSRE-Focused ProductsFeature DesignArchitecture Tradeoffs
Soft Skills
Strong CommunicationCollaborationCustomer EngagementFeedback GatheringAdaptability
Tools & Technologies
GitHubContainer OrchestrationDeveloper-Facing APIsCLIs
Certifications & Qualifications
Bachelor's Degree in Computer Science
Industry Keywords
Resilience PlatformObservabilityInfrastructure OperationsTechnical Product

Tech Stack

Tools & technologies
Kubernetes

About the role

Key responsibilities & impact
  • Define the resilience platform and own the product roadmap and delivery for specific platform features, including common telemetry interfaces, health-check contracts, attribution hooks, observability APIs, and joint deployment experience
  • Translate developer pain points into features and integration partnerships
  • Collaborate with the open-source developer community by prioritizing GitHub issues, gathering feedback, supporting contributors, and channeling community signals into the roadmap
  • Collaborate with engineering on feature design, prioritization, execution, and architecture tradeoffs
  • Align with Product, Engineering, Product Marketing, and Field teams on requirements, roadmaps, messaging, and engagements

Requirements

What you’ll need
  • 12+ years in product management, solutions architecture, or software engineering on a technical product
  • Bachelor's degree in Computer Science or equivalent experience
  • Technical depth in 2 or more of: data center operations, GPU infrastructure, network and storage, container orchestration (Kubernetes), developer platforms and SDKs, agent frameworks
  • Proven capability to connect with senior technical customers and translate requirements into product strategy
  • Comfort operating in fast-paced environments
  • Strong written and verbal communication across developer to executive audiences
  • Strong experience building products for data center infrastructure operations and observability
  • Practical experience delivering or contributing to an open-source product, including interacting with contributors on GitHub
  • Experience crafting developer-facing APIs, SDKs, or CLIs at scale
  • Background as an SRE or building SRE-focused products

Benefits

Comp & perks
  • Equity
  • Benefits 📊 Check your resume score for this job Improve your chances of getting an interview by checking your resume score before you apply. Check Resume Score