Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Cloudera

GraphRAG Engineer

Cloudera

. Own the semantic, vector, and graph storage layer powering the core context engine for enterprise AI utilities and the Internal Developer Portal .

Posted 9/22/2026full-timeBudapest • HungaryMid-LevelSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in managing and optimizing Neo4j and PostgreSQL databases within AWS and GCP environments, with a strong focus on building and maintaining enterprise knowledge graphs and automated ingestion pipelines. Proficient in implementing advanced AI-driven workflows and ensuring compliance with data governance standards.

Highest-signal resume keywords
Neo4j Database ManagementPostgreSQL ExpertiseGraph Database OptimizationInfrastructure-as-Code (Terraform)AI-Driven Workflows

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Neo4jCypherPgvectorApache AvroKubernetesDockerKafkaLangChainLlamaIndexTerraform
Tools & Technologies
AWSGCPDatadogGrafanaHashiCorp Vault
Industry Keywords
Knowledge GraphAI UtilitiesSDLC Context GraphGraphRAG EngineDLP PII Scrubbing

Tech Stack

Tools & technologies
ApacheAWSCloudDockerGoogle Cloud PlatformGrafanaKafkaKubernetesMicroservicesNeo4jPostgresSDLCSQLTerraformVault

About the role

Key responsibilities & impact
  • Own the semantic, vector, and graph storage layer powering the core context engine for enterprise AI utilities and the Internal Developer Portal
  • Provision, tune, and maintain production-grade Neo4j graph database and pgvector/PostgreSQL vector storage clusters across AWS and GCP
  • Engineer high-throughput index structures, cosine similarity vector indexes, and query optimizations for sub-second responses
  • Build automated ingestion pipelines parsing Git repositories, ASTs, Jira issue links, Apache Avro schemas, and CI/CD metadata into an enterprise knowledge graph
  • Connect distributed pipeline engines to hybrid retrievers combining SQL, Cypher graph traversals, and dense vector embeddings
  • Enable AI-driven developer workflows and autonomous coding agents
  • Configure circuit breakers, confidence scoring thresholds, and step-limit constraints for cyclic agent safeguards and governance
  • Integrate microservices and knowledge stores with the central Enterprise AI Gateway
  • Maintain version-controlled system prompt structures in localized .ai/ spoke directories while following DLP PII scrubbing rules and token rate limits
  • Implement automated failover, backup restoration, and multi-cloud storage tier cost controls across AWS and GCP
  • Lead deployment of the SDLC Context Graph and GraphRAG Engine for CAB compliance, semantic code/schema lineage tracking, and enterprise LLM proxy integrations

Requirements

What you’ll need
  • Deep operational and development experience with Neo4j, including Cypher, APOC, and causal clustering, or enterprise Knowledge Graphs
  • Proven expertise with pgvector (PostgreSQL), embeddings management, hybrid search techniques, and LangChain, LlamaIndex, or custom RAG pipelines
  • Hands-on experience managing relational (PostgreSQL) and graph databases across AWS and GCP cloud environments
  • Proficiency in consuming Apache Avro payloads, streaming Kafka events (AWS MSK), and parsing structured/unstructured code and JSON artifacts
  • Practical understanding of Prompts-as-Code patterns, few-shot prompt optimization, and agent tool specification
  • Experience with Infrastructure-as-Code (Terraform) primitives, Kubernetes (EKS/GKE), Docker, and pull-based GitOps workflows
  • Exposure to HashiCorp Vault Transit encryption, OIDC keyless authentication, and zero-trust workload identities
  • Familiarity with OpenTelemetry instrumentation and tracking vector search query latencies and LLM inference performance in Datadog or Grafana

Benefits

Comp & perks
  • Generous PTO Policy
  • Unplugged Days supporting work-life balance
  • Flexible WFH Policy
  • Mental & Physical Wellness programs
  • Phone and Internet Reimbursement program
  • Access to Continued Career Development
  • Comprehensive Benefits and Competitive Packages
  • Paid Volunteer Time
  • Employee Resource Groups