Senior Machine Learning Engineer (Platform)

Posted 4 Days Ago
Be an Early Applicant
Sydney, New South Wales, AUS
Hybrid
Senior level
Software
The Role
Build and operate end-to-end machine learning pipelines, including training, evaluation, deployment, monitoring, and model serving. Automate CI/CD, infrastructure provisioning, and reproducible environments; manage distributed GPU training across cloud and on-premises platforms; and maintain reliable, scalable inference services. Improve ML tooling, monitor production health, troubleshoot failures, and support deployment of multimodal geospatial models into customer environments.
Summary Generated by Built In

Imagine having the power to stress-test an entire power grid against a hurricane or thunderstorm before the clouds even gather. That is the reality we are creating at Neara.

We use advanced machine learning to create engineering-grade, physics enabled digital twins of electricity grids across four continents, this helps asset owners understand their biggest challenges and bring the most viable solutions to life across millions of kilometres of infrastructure.

By simulating extreme weather and structural stress at a network-wide scale, we empower the world’s largest utilities to pinpoint risks, optimise investments and build a more resilient global energy future.

Our team is a collection of brilliant minds who are fanatical about making a tangible difference in the real world, utilising AI and machine learning to accelerate everything from data classification to complex scenario analysis. We have built a special culture where innovation thrives because everyone owns the mission and we need smart, creative people to help us scale this impact to every corner of the globe.

The Senior MLOps Engineer builds and runs the pipelines, deployment systems, and observability that keep Neara's ML models training reliably and serving in production.

Neara is conducting cutting edge research, developing multi-modal spatial frontier models. You will help the team run faster, helping overcome challenges that have never been seen before in the world. These models work with a range of less researched data types, including point cloud, geospatial data, and asset data. The lack of research maturity in the geospatial domain and novel nature of the problem presents unique challenges around performance, data unification, and deployment.

The problem and role stretch beyond pure research. These models will be deployed with our global utility and new vertical customers, delivering real value and increased climate resilience for critical infrastructure. Your role will be critical in getting models out of the lab and into customer environments - reliably, repeatably, and economically.

WHAT YOU'LL DO
  • Build and operate ML pipelines - Own the training, evaluation, and deployment pipelines end-to-end, from data ingestion through to models running in customer environments.

  • Automate the path to production - Build CI/CD for models, artefact and model registries, reproducible environments, and infrastructure as code that takes an experiment to a deployed service without manual steps.

  • Keep production models healthy - Implement monitoring, alerting, and drift and data quality checks, and act as a first responder when something breaks.

  • Run distributed training infrastructure - Manage GPU clusters and scheduling, keep jobs efficient and utilisation high, and troubleshoot failures across cloud, on-prem, and neocloud environments.

  • Ship reliable serving infrastructure - Deploy and scale inference services that handle spiky load across regions and customers, with attention to latency, cost, and data residency requirements.

  • Remove friction for the ML team - Improve the tooling and workflows ML engineers use daily, and document the standards that keep things consistent as the team grows.

WHAT YOU'LL BRING
  • Solid hands-on experience building and operating ML training pipelines, model serving, and monitoring systems in production.

  • Strong Python, and working knowledge of PyTorch (or equivalent) - enough to debug a training job, not necessarily to design the model.

  • Practical experience with distributed training and GPU infrastructure, including scheduling, resource management, and diagnosing throughput or memory issues.

  • Experience in the cloud (AWS, GCP, or Azure), container orchestration (Kubernetes, Docker), and infrastructure as code.

  • Experience with production model monitoring, data quality frameworks, and preparing training data for ML readiness.

  • Sound engineering judgement - you write maintainable code, think about failure modes, and know when a quick fix is fine and when it isn't.

  • Bonus: CUDA or kernel-level optimisation experience, exposure to point cloud or geospatial data, or experience supporting deployments into regulated or air-gapped customer environments.

WHAT'S IN IT FOR YOU?
  • Competitive salary

  • Meaningful ESOP

  • Fully Flexible Work Environment. We have a fully stocked office (and an impressive snack collection) in Redfern.

  • Regular office events

  • The real benefit is working on a genuinely complex, innovative and industry-leading product, making a genuine difference in the world around us

To apply, please use the online application link below. Neara values diversity, belonging and equal employment opportunities. We encourage individuals from all backgrounds to apply.

No agencies or third-party service providers, please.
#LI-AP1

Skills Required

  • Hands-on experience building and operating production ML training pipelines, model serving, and monitoring systems
  • Strong Python skills
  • Working knowledge of PyTorch or an equivalent ML framework
  • Experience with distributed training and GPU infrastructure, including scheduling, resource management, and troubleshooting throughput or memory issues
  • Experience with AWS, GCP, or Azure
  • Experience with Kubernetes and Docker
  • Experience with infrastructure as code
  • Experience with production model monitoring and data quality frameworks
  • Experience preparing training data for ML readiness
  • Sound engineering judgment and ability to write maintainable code
  • CUDA or kernel-level optimization experience
  • Exposure to point cloud or geospatial data
  • Experience supporting deployments into regulated or air-gapped customer environments
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Redfern
80 Employees
Year Founded: 2016

What We Do

At Neara, we're helping utilities future-proof their infrastructure. We create 3D network-wide models that reflect and simulate how utility assets behave in their real-world environment in any scenario, empowering you to prepare your network for anything — from systematic risks to severe weather and a clean energy future, so you can protect your assets, teams, and communities. Our customers identify and reduce risks 9x faster across a combined territory of >650k square miles and >7.9m assets. Contact the Neara team for a free trial or demo at [email protected]

Similar Jobs

Atlassian Logo Atlassian

Senior Machine Learning Engineer

Cloud • Information Technology • Productivity • Security • Software • App development • Automation
In-Office or Remote
Sydney, New South Wales, AUS
11000 Employees

Atlassian Logo Atlassian

Software Engineer

Cloud • Information Technology • Productivity • Security • Software • App development • Automation
In-Office or Remote
Sydney, New South Wales, AUS
11000 Employees

NinjaOne Logo NinjaOne

Sales Development Representative

Information Technology • Productivity • Software • Infrastructure as a Service (IaaS)
Hybrid
Sydney, New South Wales, AUS
2000 Employees

Cloudflare Logo Cloudflare

Data Centre Selection Manager

Cloud • Information Technology • Security • Software • Cybersecurity
Hybrid
2 Locations
4400 Employees

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software • Productivity
US
15 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account