Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Mirantis

Technical Product Manager, Observability

Mirantis

. Own the vision, roadmap, and priorities for k0rdent AI observability across GPU compute, east-west fabric, high-performance storage, DPU/SmartNIC telemetry, workload schedulers, inference serving, and data services .

Posted 9/15/2026full-timeRemote • United StatesMid-LevelSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in product management for observability solutions, with a strong focus on GPU compute, cloud-native monitoring, and integration of telemetry sources. Proficient in collaborating with engineering and field teams to drive product direction and meet customer needs.

Highest-signal resume keywords
Product ManagementKubernetes ObservabilityPrometheusOpenTelemetryGPU Observability

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Observability Product OwnershipDistributed TracingLog AggregationAI Workload ProfilingHigh-Performance Storage TelemetryTelemetry Integration StrategiesMetrics and Alerting Pipeline ArchitecturePerformance AnalysisWorkload SchedulingData Service Telemetry
Soft Skills
CollaborationTechnical Trade-Off EvaluationCustomer Representation
Tools & Technologies
JaegerTempoLokiElasticsearchInfiniBandRoCEv2NVIDIA BlueField DPUVAST DataWekaDDN
Industry Keywords
AI ObservabilityCloud-Native MonitoringEast-West Fabric TelemetryTelemetry APIsSLURM Job Scheduling

Tech Stack

Tools & technologies
CloudElasticSearchKubernetesPrometheus

About the role

Key responsibilities & impact
  • Own the vision, roadmap, and priorities for k0rdent AI observability across GPU compute, east-west fabric, high-performance storage, DPU/SmartNIC telemetry, workload schedulers, inference serving, and data services
  • Translate requirements from NeoClouds, GPU clouds, telcos, sovereign clouds, and enterprise platform teams into clear product direction
  • Partner with engineering to define requirements and evaluate trade-offs
  • Manage the observability backlog using feedback from production deployments and design partners
  • Track and shape responses to emerging observability standards and technologies, including OpenTelemetry, DCGM GPU metrics, InfiniBand/RoCE fabric counters, storage platform telemetry APIs, and AI workload profiling
  • Define integration strategies for vendor telemetry sources into a unified, operator-facing observability plane
  • Partner with product marketing and field teams on positioning, technical briefs, and reference architectures
  • Represent Mirantis with customers, analysts, and ecosystem partners

Requirements

What you’ll need
  • 5+ years in product management or a senior technical role owning an observability product or operating large-scale monitoring infrastructure
  • Working knowledge of Prometheus, OpenTelemetry, distributed tracing (Jaeger, Tempo), and log aggregation (Loki, Elasticsearch/OpenSearch)
  • Fluency in Kubernetes observability, cloud-native monitoring, or metrics and alerting pipeline architecture
  • Ability to work directly with engineering on technical trade-offs and with field teams in competitive GPU cloud and NeoCloud deals
  • Exposure to GPU observability, including DCGM metrics, AI workload profiling and performance analysis
  • Familiarity with east-west fabric telemetry, including InfiniBand counters, RoCEv2 congestion metrics (ECN, PFC, DCQCN), or switch-level fabric health
  • Experience with high-performance storage telemetry from platforms such as VAST Data, Weka, or DDN, including IOPS, latency, and throughput instrumentation at scale
  • Familiarity with NVIDIA BlueField DPU telemetry, SR-IOV, or offload pipeline observability
  • Exposure to workload-level visibility for SLURM job scheduling, inference serving stacks (vLLM, Triton, TensorRT-LLM), or data service telemetry from vector databases (Milvus, Qdrant) and relational databases in AI pipelines

Benefits

Comp & perks
  • Build the observability foundation for the AI cloud era, working directly with leading GPU cloud operators, NeoClouds, sovereign clouds, and AI-first enterprises
  • Collaborate with a world-class, distributed team committed to openness and technical excellence
  • Shape the product narrative and influence go-to-market success
  • Remote work arrangement