What you'll do:
Build and maintain reliable data pipelines and ETL/ELT workflows
Develop and optimize data models for analytics and internal tools
Work with team members to deliver clean, trusted datasets
Support core data platform tools like Spark and AWS (S3, SNS, SQS, ECS/Fargate, EMR)
Monitor data pipelines for quality, performance, and reliability
Write clear documentation and contribute to test coverage and CI/CD processes
Help shape our data lakehouse architecture and platform roadmap
What you'll bring:
Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience
2–4 years of experience in data engineering or a backend data-related role
Strong skills in Java, Scala, or another backend programming language
Python (PySpark/pandas) skills
Experience with SQL and distributed data systems (e.g., Spark, Kafka, SQS)
Familiarity with NoSQL stores like Cassandra, HBase, or similar
Understanding of data modeling for analytics and reporting
Proficient in English with strong communication skills — able to explain or demo work to non-engineers
Self-driven — picks up new technologies and frameworks with little guidance
Debugs methodically; breaks complex problems into smaller steps
Open to feedback — iterates and improves through code reviews
Uses AI-assisted engineering tools (Claude Code, Codex, Cursor, etc.) as part of daily workflow
Reliable internet connection to sustain video, audio, and screen sharing
It'd be great if you had:
Experience with dbt, Databricks, or real-time data pipelines
Familiarity with cloud infrastructure tools like Terraform or CloudFormation
Interest in data governance, ML pipelines, or compliance standards
Personal projects or open source contributions demonstrating initiative
Why you'll love working here:
Work on data that supports meaningful software security outcomes
Modern tools in a cloud-first, open-source-friendly environment
A team that values clarity, learning, and autonomy
Skills Required
- Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience
- 2–4 years of experience in data engineering or a backend data-related role
- Strong skills in Java, Scala, or another backend programming language
- Python skills, including PySpark or pandas
- Experience with SQL and distributed data systems such as Spark, Kafka, or SQS
- Familiarity with NoSQL stores such as Cassandra, HBase, or similar
- Understanding of data modeling for analytics and reporting
- Proficiency in English and strong communication skills
- Ability to learn new technologies and frameworks with little guidance
- Methodical debugging and problem-solving ability
- Openness to feedback and iterative improvement through code reviews
- Experience using AI-assisted engineering tools such as Claude Code, Codex, or Cursor
- Reliable internet connection for video, audio, and screen sharing
- Experience with dbt, Databricks, or real-time data pipelines
- Familiarity with Terraform or CloudFormation
- Interest in data governance, ML pipelines, or compliance standards
- Personal projects or open source contributions demonstrating initiative
What We Do
The Sonatype journey started almost 15 years ago, just as the concept of “open source” software development was gaining steam. From our humble beginning as core contributors to Apache Maven, to supporting the world’s largest repository of open source components (Central), to distributing the world's most popular repository manager (Nexus), we’ve played a meaningful role in helping the world embrace the power of open innovation. We empower developers and security professionals with intelligent tools to innovate more securely at scale. Our platform addresses every element of an organization’s entire software development life cycle, including third-party open source code, first-party source code, and containerized code. Sonatype identifies critical security vulnerabilities and code quality issues and reports results directly to developers when they can most effectively fix them. This helps organizations develop consistently high-quality, secure software which fully meets their business needs and those of their end-customers and partners. More than 2,000 organizations, including 70% of the Fortune 100, and 15 million software developers rely on our tools and guidance to help them deliver and maintain exceptional and secure software.
Why Work With Us
We're on a mission to change how the world innovates by making software development easier. Already used by 15 million developers, we have lofty goals for our technology to be in the hands of every engineering team. And, we need you to do that. Join us!
Gallery








