Senior Backend Engineer (Data, Search, Infrastructure)

Posted 27 Days Ago
Be an Early Applicant
Hiring Remotely in Austria
Remote
60K-90K Annually
Senior level
Productivity • Software
The Role
Build and operate backend systems that ingest, process, store, and serve massive literature and user data. Develop data ingestion pipelines, optimize full-text search, handle PDF processing at scale, deploy services on AWS, and expose reliable REST APIs while ensuring data quality, deduplication, and consistency.
Summary Generated by Built In

Paperpile runs on data at scale, with a literature database of 250M+ academic papers and a growing body of user data accumulated over more than a decade. You'll work across the systems that ingest, process, store, and serve this data reliably: building pipelines, optimizing search, handling PDFs at scale, and exposing clean APIs.


Requirements
  • Strong backend engineering background with experience building and operating data-heavy systems in production.
  • Experience deploying and operating services on AWS.
  • Experience designing and maintaining data ingestion pipelines handling messy, heterogeneous sources. Comfortable with web scraping and working with third-party data sources and APIs.
  • Familiarity with Node.js and TypeScript. It’s fine if you come from a different background, such as Java or Python, but you should be comfortable working in this environment.
  • High standards for data quality. You think carefully about correctness, deduplication, and consistency.
  • Solid understanding of full-text search systems including indexing strategy, relevance tuning, and query optimization.
  • Proficient in building reliable REST APIs.

More useful experience:

  • Familiarity with academic publishing formats and data sources (PubMed, Crossref, arXiv…)
  • Experience with PDF processing pipelines (extraction, transformation, storage and delivery at scale).
  • Experience with LLM-based document processing or ML pipelines for extracting structured data from unstructured text.
  • Large scale web crawling and scraping.

Benefits
  • Base compensation €60,000–€90,000 based on the level of your experience
  • Bonus/equity program.
  • 4 weeks paid vacation + local holidays.
  • We sponsor co-working space in your city.
  • Learn and grow. Try out new things. We sponsor relevant courses, seminars, and conferences.

Skills Required

  • Strong backend engineering background with experience building and operating data-heavy systems in production
  • Experience deploying and operating services on AWS
  • Experience designing and maintaining data ingestion pipelines handling messy, heterogeneous sources
  • Comfortable with web scraping and working with third-party data sources and APIs
  • Familiarity with Node.js and TypeScript (comfortable working in this environment; Java or Python background acceptable)
  • High standards for data quality, correctness, deduplication, and consistency
  • Solid understanding of full-text search systems including indexing strategy, relevance tuning, and query optimization
  • Proficient in building reliable REST APIs
  • Familiarity with academic publishing formats and data sources (PubMed, Crossref, arXiv…)
  • Experience with PDF processing pipelines (extraction, transformation, storage and delivery at scale)
  • Experience with LLM-based document processing or ML pipelines for extracting structured data
  • Large scale web crawling and scraping experience
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
16 Employees
Year Founded: 2012

What We Do

Paperpile develops a web-based reference-management platform for researchers, universities, and institutions. Its software helps users collect, organize, read, share, and cite academic papers, manage research libraries and PDFs, and create citations and bibliographies in Google Docs. Founded by computational biologists, the company aims to simplify the workflow of collecting, managing, and writing papers through an intuitive, integrated research productivity tool.

Similar Jobs

Pfizer Logo Pfizer

Machine Learning Engineer

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
32 Locations
121990 Employees
163K-272K Annually

Pfizer Logo Pfizer

Senior Director, Innovation, Data & Analytics (IDA) Lead

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
32 Locations
121990 Employees
231K-385K Annually

Pfizer Logo Pfizer

Medical Insights, Platform Operations Lead

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
32 Locations
121990 Employees
124K-207K Annually

Pfizer Logo Pfizer

Director, Medical Insights Platform Lead

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
32 Locations
121990 Employees
177K-294K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software • Productivity
US
15 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account