Senior Staff Software Engineer - Query Execution

Posted 7 Hours Ago
Be an Early Applicant
Hiring Remotely in Portugal
Remote
Senior level
Big Data • Software
The Role
Design and improve Dremio’s distributed query execution engine, including query operators, expression evaluation, scheduling, memory management, runtime planning, and performance optimization. Diagnose correctness and scalability issues using profiling, telemetry, benchmarks, and production data. Optimize Java and JVM behavior, collaborate across engineering teams, validate AI-assisted implementations, contribute to technical direction, review code, and mentor engineers.
Summary Generated by Built In
Be Part of Building the Future

Dremio is the unified lakehouse platform for self-service analytics and AI, serving hundreds of global enterprises, including Maersk, Amazon, Regeneron, NetApp, and S&P Global. Customers rely on Dremio for cloud, hybrid, and on-prem lakehouses to power their data mesh, data warehouse migration, data virtualization, and unified data access use cases. Based on open source technologies, including Apache Iceberg and Apache Arrow, Dremio provides an open lakehouse architecture enabling the fastest time to insight and platform flexibility at a fraction of the cost.  Learn more at www.dremio.com.

About the role

Dremio’s Query Execution team builds the runtime responsible for executing analytical queries at scale. The team sits between query planning and table scans, owning the systems that turn query plans into efficient, reliable execution.

As a Senior Staff Software Engineer, you’ll work on query operators, expression evaluation, memory management, execution scheduling, runtime planning, and performance optimization. Your work will directly affect query performance, scalability, and reliability across Dremio’s cloud data platform.


What you’ll be doing

  • Build the execution engine: Design, implement, test, and support query operators, expression evaluation, execution scheduling, runtime planning, and resource management.
  • Improve runtime performance: Optimize vectorized processing, concurrency, parallelism, memory usage, data exchange, and disk spilling.
  • Diagnose difficult problems: Investigate query correctness and performance issues using query profiles, profiling tools, telemetry, benchmarks, logs, and production data.
  • Optimize Java and the JVM: Improve CPU utilization, allocation behavior, garbage collection, thread behavior, and other runtime characteristics.
  • Collaborate across the query engine: Work with Query Planner, Datalake, Product, Support, and field engineering teams on end-to-end query improvements.
  • Use AI-assisted development tools: Apply AI tools to code comprehension, debugging, test generation, documentation, and incident analysis.
  • Apply engineering judgment: Validate AI-generated code and recommendations through review, testing, benchmarking, and engineering judgment.
  • Raise the engineering bar: Participate in design and code reviews, contribute to technical direction, and mentor other engineers.

What we’re looking for

  • Software engineering experience: 5+ years developing production software, preferably in database systems, query execution, distributed systems, or related areas.
  • Programming skills: Strong Java skills; experience with C++ is a plus.
  • Computer science fundamentals: Strong foundation in data structures, algorithms, concurrency, multithreaded programming, and asynchronous programming.
  • Database and distributed systems knowledge: Understanding of database internals, query processing, distributed execution, and performance optimization.
  • Performance engineering experience: Experience profiling and optimizing CPU, memory, garbage collection, thread behavior, or other runtime characteristics.
  • Problem-solving ability: Ability to solve ambiguous technical problems and work effectively across team boundaries.
  • Engineering judgment: Ability to evaluate tradeoffs, validate solutions rigorously, and make sound decisions when working with unfamiliar or AI-assisted implementations.
  • Communication and ownership: Strong communication skills, a sense of ownership, and a commitment to delivering high-quality software.

Bonus points if you have

  • Query engines: Distributed query engines, database engines, or large-scale data processing systems.
  • Execution techniques: Vectorized execution, query operators, expression evaluation, code generation, or runtime optimization.
  • Runtime systems: JVM profiling, memory management, caching, networking, data exchange, execution scheduling, or disk spilling.
  • Data platforms: Apache Arrow, Apache Iceberg, Parquet, Avro, Spark, Hadoop, or cloud object stores.
  • Open source: Contributing to open-source projects.
What we value 

At Dremio, we hold ourselves to high standards when it comes to People, Thinking, and Action. Our Gnarlies (that's what we call our employees) communicate with clarity, drive accountability, and are respectful towards each other. We confront brutal facts and focus on results while operating with a sense of urgency and building a "flywheel". People who like to jump in and drive momentum will thrive in our #GnarlyLife.

Dremio is an equal opportunity employer supporting workforce diversity. We do not discriminate on the basis of race, religion, color, national origin, gender identity, sexual orientation, age, marital status, protected veteran status, disability status, or any other unlawful factor.

Dremio is committed to providing any necessary accommodations for individuals with disabilities within our application and interview process. To request accommodation due to a disability, please inform your recruiter.

Dremio has policies in place to protect the personal information that employees and applicants disclose to us. Please click here to review the privacy notice. 

Important Security Notice for Candidates

At Dremio, we uphold trust and transparency as paramount values in all our interactions with customers, partners, employees, and the general public. We have been targeted by individuals creating fake domains similar to ours to scam prospects and candidates. Please note that all official communications from us will be from an @dremio.com domain. If you suspect you've been targeted by a scam, it's imperative to report the incident to your local law enforcement agencies. For more information about this type of scam, please refer to Dremio's official statement here.

Dremio is not responsible for any fees related to unsolicited resumes and will not pay fees to any third-party agency or company that does not have a signed agreement with the Company.

Skills Required

  • 5+ years developing production software, preferably in database systems, query execution, distributed systems, or related areas
  • Strong Java programming skills
  • Strong foundation in data structures and algorithms
  • Experience with concurrency, multithreaded programming, and asynchronous programming
  • Understanding of database internals, query processing, distributed execution, and performance optimization
  • Experience profiling and optimizing CPU, memory, garbage collection, thread behavior, or other runtime characteristics
  • Ability to solve ambiguous technical problems and work effectively across team boundaries
  • Ability to evaluate tradeoffs, validate solutions rigorously, and exercise sound engineering judgment
  • Strong communication skills, sense of ownership, and commitment to high-quality software
  • Experience with C++
  • Experience with distributed query engines, database engines, or large-scale data processing systems
  • Experience with vectorized execution, query operators, expression evaluation, code generation, or runtime optimization
  • Experience with JVM profiling, memory management, caching, networking, data exchange, execution scheduling, or disk spilling
  • Experience with Apache Arrow, Apache Iceberg, Parquet, Avro, Spark, Hadoop, or cloud object stores
  • Open-source contribution experience
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Austin, TX
258 Employees
Year Founded: 2015

What We Do

Dremio is the Data Lake Engine. Created by veterans of open source and big data technologies, and the creators of Apache Arrow, Dremio is a fundamentally new approach to data analytics that helps companies get more value from their data, faster. Dremio makes data engineering teams more productive, and data consumers more self-sufficient. For more information, visit www.dremio.com. Founded in 2015, Dremio is headquartered in Mountain View, CA. Investors include Lightspeed Venture Partners, Redpoint, and Norwest Venture Partners. Connect with Dremio on GitHub, LinkedIn, Twitter, and Facebook.

Similar Jobs

Pfizer Logo Pfizer

Director R&D EHS Program Lead

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
36 Locations
121990 Employees
177K-294K Annually

Pfizer Logo Pfizer

Quality Assurance Manager

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
28 Locations
121990 Employees

Pfizer Logo Pfizer

Manager Archiving and Compliance

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
28 Locations
121990 Employees

HiBob Logo HiBob

Customer Experience Process Specialist - AI & Automation

HR Tech • Information Technology • Professional Services • Sales • Software
Remote or Hybrid
Portugal
1350 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account