Principal Observability Platform Engineer

Posted 5 Hours Ago
Be an Early Applicant
Sydney, New South Wales, AUS
In-Office
Expert/Leader
Fintech • Quantitative Trading
The Role
Design, build, and operate Optiver's global observability platform covering telemetry collection, ingestion, storage, query, visualization, alerting, diagnostics and service health. Build APIs, integrations, libraries, dashboards and automation to improve telemetry quality, scalability, reliability, cost-effectiveness and developer/operator experience. Collaborate with engineering, trading and ops to improve adoption and production investigation workflows and own operational quality of platform components.
Summary Generated by Built In

WHO WE ARE

Optiver is a tech-driven trading firm and leading global market maker. For over 35 years, Optiver has been improving financial markets worldwide, making them more transparent and efficient for all participants. With more than 1,400 employees in offices around the world, we’re united in our commitment to improving the market through competitive pricing, execution and thorough risk management. By providing liquidity on multiple exchanges across the world, we actively trade on 70+ exchanges, where we’re trusted to always provide accurate buy and sell pricing – no matter the market conditions.

WHAT YOU'LL DO
We are looking for a Principal Observability Platform Engineer to help evolve observability as a business-critical platform capability at Optiver. You will work on the shared platform behind metrics, logs, traces, events, alerts, dashboards, diagnostics, instrumentation and service health.

This is a platform engineering role for someone who enjoys building reliable systems used by other engineers. You will help turn a capable but heterogeneous observability foundation into a globally consistent, regionally federated platform that is reliable at scale, easy to adopt, and deeply embedded in how Optiver builds and operates production systems.

As a Senior Observability Platform Engineer, you will design, build, and operate components that help engineers, operators, trading teams, automated systems, and future agent-based workflows collect, query, understand, and act on production signals. You will work across platform and production domains: building high-scale telemetry pipelines, improving instrumentation quality, creating golden paths for adoption, and making observability more useful during real production investigations.

In this role, you will:

  • Design, build, and operate components of Optiver’s shared observability platform across telemetry collection, ingestion, storage, query, visualisation, alerting, diagnostics, and service health.

  • Build software, services, APIs, integrations, libraries, dashboards, automation, and reusable patterns that make observability easier to adopt and more reliable to operate.

  • Improve the scalability, reliability, performance, cost-effectiveness, and operational quality of high-volume telemetry systems.

  • Improve developer and operator experience through self-service workflows, golden paths, documentation, investigation tooling, and practical platform abstractions.

  • Work with engineering, infrastructure, trading systems, research, and regional operations teams to understand production debugging needs and improve observability adoption.

  • Own the reliability and operational quality of the components you build, including service health, failure modes, monitoring, incident learnings, and continuous improvement.

  • Raise the standard for telemetry quality, instrumentation, alerting, dashboards, diagnostic workflows, and service health across Optiver.

WHAT YOU'LL BRING
You are a strong engineer with experience in production systems, platform engineering, SRE, infrastructure, observability, or distributed systems. You are comfortable working on systems that need to be reliable, scalable, understandable, and useful to other engineers.

You understand that observability is not just a tooling problem. It is about signal quality, platform reliability, developer experience, production workflows, and adoption. You care about building systems that engineers trust, operators can rely on, and production teams can depend on during high-pressure situations.

You will bring:

  • Strong engineering experience in SRE, software engineering, platform engineering, infrastructure, observability, developer tooling, or distributed systems.

  • A production mindset, with the ability to reason about failure modes, debugging workflows, service reliability, operational impact, and how systems behave under pressure.

  • Technical understanding of modern observability practices across logs, metrics, traces, events, alerting, dashboards, telemetry pipelines, diagnostics, instrumentation quality, and service health.

  • Experience designing, building, or operating reliable services, platforms, pipelines, tools, or automation used by other engineering teams.

  • Good judgement in technical trade-offs across performance, scalability, reliability, complexity, cost, and maintainability.

  • A delivery mindset, with the ability to take ambiguous platform problems and turn them into practical, reliable solutions.

  • Strong preference will be given to candidates with experience on observability, SRE, infrastructure, platform, production engineering, or developer tooling teams in large-scale distributed systems environments, including telemetry pipelines, streaming systems, time-series data, log platforms, query systems, alerting systems, or production diagnostics tooling.

  • Experience with technologies such as Kafka, Grafana, ELK/OpenSearch, ClickHouse, VictoriaMetrics, InfluxDB, Telegraf, Vector, OpenTelemetry, Prometheus-style systems, or custom telemetry collectors is valued.


WHAT YOU’LL GET

  • A performance-based bonus structure unmatched anywhere in the industry. We combine our profits across desks, teams and offices into a global profit pool, fostering a truly collaborative environment.

  • The chance to work alongside diverse and intelligent peers in a rewarding environment.

  • Training, mentorship and personal development opportunities.

  • Daily breakfast, lunch and an in-house barista.

  • Gym membership plus weekly in-house chair massages.

  • Regular social events, including a company trip every two years.

  • Guided relocation, a competitive relocation package and visa sponsorship where necessary.

DIVERSITY STATEMENT

Optiver is committed to diversity and inclusion. We encourage applications from candidates of all backgrounds, and welcome requests for reasonable adjustments during the process.

Questions? Get in touch with the recruitment team at [email protected].


Skills Required

  • Strong engineering experience in SRE, platform engineering, infrastructure, observability, developer tooling, or distributed systems
  • Production mindset with ability to reason about failure modes, debugging workflows, and service reliability
  • Technical understanding of modern observability practices across logs, metrics, traces, events, alerting, dashboards, telemetry pipelines, diagnostics, instrumentation quality, and service health
  • Experience designing, building, or operating reliable services, platforms, pipelines, tools, or automation used by other engineering teams
  • Good judgement in technical trade-offs across performance, scalability, reliability, complexity, cost, and maintainability
  • Delivery mindset with ability to turn ambiguous platform problems into practical, reliable solutions
  • Experience in large-scale distributed systems, telemetry pipelines, streaming systems, time-series data, log platforms, query systems, alerting systems, or production diagnostics tooling
  • Experience with Kafka, Grafana, ELK/OpenSearch, ClickHouse, VictoriaMetrics, InfluxDB, Telegraf, Vector, OpenTelemetry, Prometheus-style systems, or custom telemetry collectors

Optiver Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Optiver and has not been reviewed or approved by Optiver.

  • Career-Linked Recognition & Rewards Pay is considered highly competitive in trading, quantitative, and engineering paths, with performance-linked bonuses materially elevating total earnings for strong contributors. High performers are described as seeing outsized upside when firm and team results are strong.
  • Healthcare Strength Health coverage is depicted as comprehensive, including medical, dental, vision, disability and life insurance, alongside HSA/FSA options and mental-health support. Some locations note unlimited therapy access through a partner platform.
  • Leave & Time Off Breadth U.S. roles are often described as offering 25 days of paid vacation plus market holidays, alongside generous parental leave. This breadth of time off is highlighted as helping to offset a high-intensity environment.

Optiver Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Amsterdam
1,600 Employees
Year Founded: 1986

What We Do

Optiver’s story began over 30 years ago, when we started business as a single trader on the floor of Amsterdam’s options exchange. Today, we are at the forefront of trading and technology as a leading global electronic market maker, focused on pricing, execution and risk management.

Why Work With Us

People at Optiver love challenges, welcome collaboration, and strive to be better tomorrow than they are today. Improving the market is an extraordinary challenge that requires a carefully crafted approach. Optiver provides a collaborative working environment to tackle these challenges. In fact, it is this way of working that sets us apart.

Gallery

Gallery

Similar Jobs

Snyk Logo Snyk

Sales Development Representative

Artificial Intelligence • Cloud • Information Technology • Security • Software • Cybersecurity • Data Privacy
Hybrid
Sydney, New South Wales, AUS
1000 Employees
42K-78K Annually

Airwallex Logo Airwallex

Account Executive

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI
In-Office
2 Locations
2300 Employees
100K-130K Annually

Morningstar Logo Morningstar

Sales Operations Analyst

Artificial Intelligence • Big Data • Enterprise Web • Fintech • Software • Financial Services
Hybrid
Sydney, New South Wales, AUS
11500 Employees

Snap Inc. Logo Snap Inc.

Client Partner

Artificial Intelligence • Cloud • Machine Learning • Mobile • Software • Virtual Reality • App development
Hybrid
Sydney, New South Wales, AUS
5000 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account