Senior Knowledge Graph Engineer

Reposted 20 Days Ago
Be an Early Applicant
Stockholm, SWE
In-Office
Senior level
Artificial Intelligence • Information Technology • Software • Automation
The Role
Design and own a production knowledge graph from multimodal licensed sources: build ontologies, entity resolution, provenance, live updates, and graph-backed retrieval to support multi-hop agent reasoning.
Summary Generated by Built In

Why this role exists

Agents are only as good as the knowledge they can reach. Most retrieval today is flat: chunk a document, embed it, hope the nearest vector is the right answer. That breaks the moment a question needs more than one hop, needs to know which source to trust, or needs to distinguish two entities that happen to share a name. Redpine is building the layer that fixes this: a knowledge graph over the licensed data we hold across medicine, science, law, and finance, served directly to agents over our API, MCP, and CLI.

You will own that graph end to end, from raw multimodal sources to the structured, provenance-bearing knowledge that agents query in production. This is an early, high-ownership role. The decisions you make about how we model and link knowledge will shape what every agent built on Redpine can actually reason about.

Why knowledge graphs are core to agentic retrieval

Agents don't just look things up, they traverse. They follow a claim to its source, a company to its filings, a drug to its trials to the patients those trials enrolled. That kind of multi-hop reasoning needs explicit structure: typed entities, named relations, and a graph a planner can walk.

What you'll work on

  • Ontology and schema design. Define the entity types, relations, and constraints that licensed data from each domain maps into, building on the established ontologies those fields already use rather than reinventing them. You'll decide where a shared ontology helps and where a domain needs its own.

  • Entity resolution at scale. Deduplicate, canonicalize, and disambiguate entities across heterogeneous, multimodal sources, so the same thing in the world is one node in the graph.

  • Confidence and conflict. Attach confidence to every link, and define what happens when two sources disagree. Decide what the graph asserts, what it flags, and what it surfaces to the agent.

  • Provenance as a first-class property. Preserve attribution on every node and edge, back to the source document and the license that covers it. At Redpine this is not optional, it is the product.

  • Keeping the graph live. Detect when an upstream source changes and propagate that change, so the graph reflects current knowledge instead of a stale snapshot.

  • Graph-powered retrieval. Build multi-hop traversal, hybrid graph-and-vector retrieval, and the kind of structure a reranker can actually exploit.

What we're looking for

  • Deep, direct experience with knowledge graphs: graph databases, property or RDF graphs, entity resolution, and ontology design. You've shipped one, not just read about them.

  • Strong Python and a track record of building data pipelines from scratch.

  • Real experience with retrieval and RAG, ideally graph-backed, plus the judgment to know when a graph earns its keep and when it's just overhead.

  • Judgment about messy data. You can look at multi-source, contradictory input and design a schema that survives contact with reality.

  • A bias toward small, clear systems. You question whether something needs to be built before you build it.

  • A genuinely curious mind, the kind that wants to understand a domain well enough to model it honestly.


About Redpine

If models were the first wave of AI, and compute the second, we're building the data layer that comes next.

Only a small fraction of the world's data is on the open internet. The rest, high-quality, domain-specific, often critical, sits behind paywalls, in databases, or with rights holders. Redpine is building the infrastructure to unlock it.

We provide AI builders and autonomous agents with access to licensed, high-quality, multimodal data through a unified platform and API. The goal is simple: make AI systems more accurate, more useful, and grounded in real-world information.

We're backed by Nordic Ninja, Node VC, and Luminar, alongside angels from OpenAI, Spotify, and Perplexity.

Skills Required

  • Deep hands-on experience with knowledge graphs, graph databases, RDF or property graphs, entity resolution, and ontology design.
  • Strong Python skills and demonstrated experience building data pipelines from scratch.
  • Practical experience with retrieval and RAG (retrieval-augmented generation), ideally graph-backed retrieval and hybrid graph/vector systems.
  • Experience designing and preserving provenance and attribution on nodes and edges in production systems.
  • Experience building systems to detect and propagate upstream source changes to keep graphs live and current.
  • Experience building graph-powered traversal, multi-hop retrieval, and structures usable by rerankers.
  • Good judgment with messy, contradictory data and a bias toward small, clear systems; curiosity to model domains honestly.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
12 Employees
Year Founded: 2024

What We Do

Only 1% of the world's data is openly available on the Internet. We unlock the rest, to power your AI development. Redpine powers AI builders and autonomous agents with access to licensed, high-quality and multimodal data - securely and at scale. The platform empowers content owners across domains - including copyright holders and proprietary data providers - to offer controlled and compliant access to their data, unlocking value for both data owners and AI builders. Founded in 2024 and headquartered in Stockholm, Redpine’s team brings deep expertise from leading organizations in data science, machine learning, and AI product development. The company is backed by prominent investors from OpenAI, Spotify, Perplexity, Xiaomi, Sana, and more.

Similar Jobs

Tulip Logo Tulip

Marketing Manager

Enterprise Web • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
27 Locations
310 Employees

Tulip Logo Tulip

DACH Regional Sales Lead – Enterprise Manufacturing

Enterprise Web • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
27 Locations
310 Employees

Cloudflare Logo Cloudflare

Account Executive

Cloud • Information Technology • Security • Software • Cybersecurity
Remote or Hybrid
Sweden
4400 Employees

Legora Logo Legora

Principal Platform Advisor - Stockholm

Artificial Intelligence • Legal Tech • Software
In-Office
Stockholm, SWE
700 Employees

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account