Senior Knowledge Graph Engineer

Posted Yesterday
Be an Early Applicant
Stockholm, SWE
In-Office
Senior level
Artificial Intelligence • Information Technology • Software • Automation
The Role
Own Redpine’s knowledge graph end to end, including ontology and schema design, entity resolution, provenance, confidence and conflict handling, source-change propagation, and graph-powered retrieval. Build production data pipelines from multimodal sources and enable multi-hop, graph-and-vector retrieval for AI agents. This high-ownership role requires strong Python, direct knowledge graph experience, and practical expertise with retrieval and RAG systems.
Summary Generated by Built In
The role

Agents are only as good as the knowledge they can reach. Most retrieval today is flat: chunk a document, embed it, hope the nearest vector is the right answer. That breaks the moment a question needs more than one hop, needs to know which source to trust, or needs to distinguish two entities that happen to share a name. Redpine is building the layer that fixes this: a knowledge graph over the licensed data we hold across medicine, science, law, and finance, served directly to agents over our API, MCP, and CLI.

You will own that graph end to end, from raw multimodal sources to the structured, provenance-bearing knowledge that agents query in production. This is an early, high-ownership role. The decisions you make about how we model and link knowledge will shape what every agent built on Redpine can actually reason about.

What you'll do
  • Ontology and schema design. Define the entity types, relations, and constraints that licensed data from each domain maps into, building on the established ontologies those fields already use rather than reinventing them. You will decide where a shared ontology helps and where a domain needs its own.

  • Entity resolution at scale. Deduplicate, canonicalize, and disambiguate entities across heterogeneous, multimodal sources, so the same thing in the world is one node in the graph.

  • Confidence and conflict. Attach confidence to every link, and define what happens when two sources disagree. Decide what the graph asserts, what it flags, and what it surfaces to the agent.

  • Provenance as a first-class property. Preserve attribution on every node and edge, back to the source document and the license that covers it. At Redpine this is not optional, it is the product.

  • Keeping the graph live. Detect when an upstream source changes and propagate that change, so the graph reflects current knowledge instead of a stale snapshot.

  • Graph-powered retrieval. Build multi-hop traversal, hybrid graph-and-vector retrieval, and the kind of structure a reranker can actually exploit.

What we're looking for
  • Deep, direct experience with knowledge graphs: graph databases, property or RDF graphs, entity resolution, and ontology design. You have shipped one, not just read about them.

  • Strong Python and a track record of building data pipelines from scratch.

  • Real experience with retrieval and RAG, ideally graph-backed, plus the judgment to know when a graph earns its keep and when it is just overhead.

  • Judgment about messy data. You can look at multi-source, contradictory input and design a schema that survives contact with reality.

  • A bias toward small, clear systems. You question whether something needs to be built before you build it.

  • A genuinely curious mind, the kind that wants to understand a domain well enough to model it honestly.

Why this role

Agents don't just look things up, they traverse. They follow a claim to its source, a company to its filings, a drug to its trials to the patients those trials enrolled. That kind of multi-hop reasoning needs explicit structure: typed entities, named relations, and a graph a planner can walk.

Practical details
  • Based at Redpine HQ in central Stockholm. Office-first, with real autonomy for deep-work days and life admin.

  • Requires a valid Swedish work permit or similar eligibility. Relocation support available.

  • Competitive salary and meaningful equity.

  • Small team, high ownership. The same people design, ship, and run the systems.

About Redpine

If models were the first wave of AI, and compute the second, we are building the data layer that comes next. Only a small fraction of the world's data is on the open internet. The rest, high-quality, domain-specific, often critical, sits behind paywalls, in databases, or with rights holders. Redpine is building the infrastructure to unlock it.

 

We provide AI builders and autonomous agents with access to licensed, high-quality, multimodal data through a unified platform and API. The goal is simple: make AI systems more accurate, more useful, and grounded in real-world information.

 

You will work closely with the founding team, including Anders (ex-partner at VC, McKinsey) and David (ex-tech and product leader at Spotify, Zettle, and Lunar). We are backed by Nordic Ninja, Node VC, and Luminar, alongside angels from OpenAI, Spotify, and Perplexity.

 

If you are eager to make a meaningful impact in the AI space, we would love to hear from you. We are committed to building a diverse and inclusive team. If you are excited about this role but your experience does not align perfectly with every qualification, we encourage you to apply anyway.

Skills Required

  • Deep, direct experience building and shipping knowledge graphs
  • Experience with graph databases, property or RDF graphs, entity resolution, and ontology design
  • Strong Python skills
  • Track record of building data pipelines from scratch
  • Experience with retrieval and retrieval-augmented generation systems
  • Graph-backed retrieval experience
  • Valid Swedish work permit or similar work eligibility
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
12 Employees
Year Founded: 2024

What We Do

Only 1% of the world's data is openly available on the Internet. We unlock the rest, to power your AI development. Redpine powers AI builders and autonomous agents with access to licensed, high-quality and multimodal data - securely and at scale. The platform empowers content owners across domains - including copyright holders and proprietary data providers - to offer controlled and compliant access to their data, unlocking value for both data owners and AI builders. Founded in 2024 and headquartered in Stockholm, Redpine’s team brings deep expertise from leading organizations in data science, machine learning, and AI product development. The company is backed by prominent investors from OpenAI, Spotify, Perplexity, Xiaomi, Sana, and more.

Similar Jobs

Legora Logo Legora

Operations Associate

Artificial Intelligence • Legal Tech • Software
In-Office
Stockholm, SWE
700 Employees

Zscaler Logo Zscaler

Sales Engineer

Cloud • Information Technology • Security • Software • Cybersecurity
Easy Apply
Remote or Hybrid
Sweden
8697 Employees

Mondelēz International Logo Mondelēz International

CI Engineer

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Hybrid
Upplands Väsby, SWE
90000 Employees

Legora Logo Legora

Team Lead

Artificial Intelligence • Legal Tech • Software
In-Office
Stockholm, SWE
700 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account