Senior Data & Python Software Engineer

Reposted 2 Days Ago
Be an Early Applicant
Hiring Remotely in Poland
Remote
Senior level
Legal Tech • Software
The Role
Build and operate large-scale web data extraction and scraping systems, design ingestion and processing pipelines, ensure data quality and governance, optimize performance and reliability, implement observability, and collaborate with product and engineering teams to deliver production-grade APIs, backend services, and cloud-deployed workflows.
Summary Generated by Built In

At Ceartas, we lead the way in AI-powered brand protection, copyright law, and digital security, safeguarding the integrity of content creators, brands, and enterprises worldwide. As we scale rapidly, we're looking for a Data & Python Software engineer to drive innovation in our data pipelines and crawling technologies. In this pivotal role, you'll collaborate with our CTO and Head of Engineering, steering our Data Engineering Team toward developing groundbreaking solutions for digital security challenges.

Build Scalable Web Data Extraction Pipelines:

  • Design and develop web scraping systems that support large-scale web data extraction and brand protection workflows.

  • Ensure that data moves reliably from collection through processing to storage while maintaining performance, resilience, and operational stability at scale.

Ensure Data Quality and Governance:

  • Own data validation, consistency, and governance across ingestion, storage, and serving layers.

  • Establish clear standards for schema design, transformation logic, and monitoring to guarantee trustworthy, production-grade datasets that can be reliably consumed across the organization.

Optimize Performance and Reliability:

Continuously improve scraping system efficiency through performance tuning, cost optimization, and architectural enhancements. Implement logging, metrics, and tracing to monitor production systems, diagnose issues quickly, and maintain high reliability under growing workloads.

Responsibilities:

  • Design, build, and maintain high-performance web scraping systems as well backend services and data pipelines supporting web data extraction and brand protection use cases

  • Implement and maintain scraping focused APIs and other data services that power internal products and external integrations

  • Build reliable ingestion, processing, and storage workflows for large-scale web data

  • Handle cleaning of web data and ensure data quality, validation, and governance across ingestion, storage, and serving layers

  • Optimize scraping systems for performance, scalability, reliability, and cost efficiency

  • Monitor, debug, and improve scraping system reliability using observability tools (logging, metrics, tracing)

  • Collaborate closely with product and engineering teams to deliver features from design through full end-to-end production deployment

  • Take independent ownership of systems in production, including maintenance, iteration and performance management

Core Technical Requirements:

  • Experience with web scraping

  • Strong SQL skills

  • Strong Python experience

  • Experience with PostgreSQL or similar relational databases

  • Experience designing and building scalable APIs and backend services (e.g. FastAPI, Django, or similar frameworks)

  • Experience designing efficient, scalable data models and database schemas

  • Hands-on experience deploying and operating systems in the cloud (AWS, GCP, or Azure)

  • Experience working with Docker and containerized environments

Preferred Technical Requirements:

  • Experience with workflow orchestration tools such as Airflow

  • Experience with browser-based automation tools (Playwright, Selenium, or similar)

  • Experience with DBT or analytics-focused data transformation workflows

  • Experience building or operating high-concurrency systems and task queues

  • Experience designing and deploying cloud-native workflows on AWS

  • Familiarity with CI/CD pipelines and production deployment practices

  • Experience working in a high-growth, early-stage startup environment

  • Experience - University education in a technical field such as Computer Science, Engineering or similar. Masters level preferred. 4+ years ( or 2 year+ in a early stage startup)

What We Offer:

  • 25 paid days off per year, plus public holidays, and your birthday!

  • A culture of innovation, continuous learning, and collaboration.

  • A vibrant, inclusive workplace that celebrates diversity and equal opportunity.

  • Access to the latest technology

  • Employee recognition and reward programs.

  • Opportunity for growth within a fast-moving, forward-thinking startup.

Skills Required

  • Experience with web scraping
  • Strong SQL skills
  • Strong Python experience
  • Experience with PostgreSQL or similar relational databases
  • Experience designing and building scalable APIs and backend services (e.g. FastAPI, Django)
  • Experience designing efficient, scalable data models and database schemas
  • Hands-on experience deploying and operating systems in the cloud (AWS, GCP, or Azure)
  • Experience working with Docker and containerized environments
  • Experience with workflow orchestration tools such as Airflow
  • Experience with browser-based automation tools (Playwright, Selenium, or similar)
  • Experience with DBT or analytics-focused data transformation workflows
  • Experience building or operating high-concurrency systems and task queues
  • Familiarity with CI/CD pipelines and production deployment practices
  • University education in a technical field (Computer Science, Engineering or similar); Masters preferred; 4+ years experience (or 2+ years in an early-stage startup)
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Dublin, Dublin
15 Employees
Year Founded: 2021

What We Do

The world's top creators, agencies, and brands rely on Ceartas DMCA to prevent the unauthorized use of their creative work. With Ceartas DMCA, you can easily monitor your content and manage your rights lines. Our proprietary technology keeps you safe from copyright infringement, protecting you and the value of your creations with lightning-quick takedowns at a price to suit every need

Similar Jobs

Remote or Hybrid
2 Locations
1100 Employees
Remote or Hybrid
2 Locations
1100 Employees

Capco Logo Capco

Talent Acquisition Specialist

Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
Remote or Hybrid
Poland
6000 Employees

Taskrabbit Logo Taskrabbit

Customer Support Advocate (Spanish)

eCommerce • Information Technology • Sharing Economy • Software
Easy Apply
Remote or Hybrid
Poland
450 Employees
83K-83K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account