Software Engineer, Machine Lifecycle

Posted 16 Days Ago
Be an Early Applicant
New York, NY, USA
In-Office
150K-250K Annually
Mid level
Fintech
The Role
Design and build a zero-touch machine lifecycle pipeline: automate out-of-band firmware/BIOS management, network boot OS provisioning, Ansible-based configuration, CI/CD-driven GitOps reconciliation, automated validation/burn-in, lifecycle state machine, and instrumentation for fleet-scale datacenter operations.
Summary Generated by Built In

Tower Research Capital is a leading quantitative trading firm founded in 1998. Tower has built its business on a high-performance platform and independent trading teams. We have a 25+ year track record of innovation and a reputation for discovering unique market opportunities.
Tower is home to some of the world’s best systematic trading and engineering talent. We empower portfolio managers to build their teams and strategies independently while providing the economies of scale that come from a large, global organization. 

Engineers thrive at Tower while developing electronic trading infrastructure at a world class level. Our engineers solve challenging problems in the realms of low-latency programming, FPGA technology, hardware acceleration and machine learning. Our ongoing investment in top engineering talent and technology ensures our platform remains unmatched in terms of functionality, scalability and performance.

At Tower, every employee plays a role in our success. Our Business Support teams are essential to building and maintaining the platform that powers everything we do — combining market access, data, compute, and research infrastructure with risk management, compliance, and a full suite of business services. Our Business Support teams enable our trading and engineering teams to perform at their best.

At Tower, employees will find a stimulating, results-oriented environment where highly intelligent and motivated colleagues inspire each other to reach their greatest potential.

Summary:
This role owns the journey of every machine in our fleet: from the moment a server is racked, cabled, and powered on, to the moment it is fully configured, validated, and available for users. Your mission is to make that journey zero-touch.

You will design and build the automation pipeline that takes a machine through discovery, firmware and BIOS configuration, OS installation, configuration management, health validation and burn-in, and finally handoff into production, treating each stage as code that lives in Git, runs through CI/CD, and can be reviewed, tested, and rolled back like any other software.

The guiding principle is GitOps for physical infrastructure: the desired state of the fleet is declared in a repository, and automation continuously reconciles reality against it. A new machine shows up as a commit; a decommission is a deletion; drift is detected and corrected by the pipeline, not by a person with a checklist.

Responsibilities:

  • Design and build the end-to-end machine lifecycle pipeline: from power-on and network boot through OS install, configuration, validation, and production handoff.
  • Automate hardware bring-up via out-of-band management (BMC, Redfish, IPMI): firmware updates, BIOS settings, boot order, and inventory discovery.
  • Automate OS provisioning with network boot (PXE / UEFI HTTP boot) and unattended installation, so no one ever installs a machine by hand.
  • Write and maintain the Ansible and Python that configure machines into their final roles, replacing manual runbooks with reviewed, versioned code.
  • Apply GitOps and CI/CD principles to the fleet: desired state in Git, changes through merge requests, pipelines that test and apply them, and reconciliation that catches drift.
  • Build automated validation and burn-in: health checks, stress tests, and acceptance criteria a machine must pass before users ever see it.
  • Model the lifecycle as a state machine (new, provisioning, validating, in-service, needs-repair, decommissioned) with clear, automated transitions and an auditable history.
  • Instrument the pipeline with metrics and logging so we always know where a machine is in its lifecycle, and where the process is slow or failing.
  • Work with the HPC and datacenter teams to fold their hard-won operational knowledge into the automation, one stage at a time.

Qualifications:

  • A smart, curious engineer who learns fast and is genuinely excited by the challenge of automating physical infrastructure end to end. This matters more to us than any specific line on your resume.
  • Strong Python for building automation, tooling, and services, not just scripts.
  • Hands-on Ansible experience: writing playbooks and roles you would be happy to code-review, not just run.
  • A solid grasp of CI/CD principles: pipelines, testing, staged rollouts, and the discipline of driving change through version control.
  • An automation-first, GitOps mindset: you believe infrastructure state belongs in Git, and that any task done by hand twice should be code.
  • Working knowledge of Linux: comfortable with the boot process, system services, and debugging when a machine does not come up the way it should. Depth here is a real plus, but interest and trajectory count.
  • Sound engineering judgment: you design workflows that fail safely, retry sensibly, and leave an audit trail.
  • Clear communication and the patience to turn tribal operational knowledge into reliable, documented automation.

Nice to Have:

  • Experience with bare-metal provisioning tooling such as MAAS, Tinkerbell, Foreman, Ironic, or a home-grown equivalent.
  • Familiarity with out-of-band management: BMCs, Redfish, IPMI, and vendor variants like iDRAC or iLO.
  • Exposure to hardware validation and burn-in: stress testing, firmware qualification, or failure prediction at fleet scale.
  • Experience with GitOps tooling or declarative infrastructure management in general.
  • Prior work in datacenter, HPC, or large-fleet environments where machines number in the hundreds or thousands.

Anticipated annual base salary range $150,000-$250,000, plus eligible for discretionary bonus.

Tower’s headquarters are in the historic Equitable Building, right in the heart of NYC’s Financial District and our impact is global, with over a dozen offices around the world.

 At Tower, we believe work should be both challenging and enjoyable. That is why we foster a culture where smart, driven people thrive – without the egos. Our open concept workplace, casual dress code, and well-stocked kitchens reflect the value we place on a friendly, collaborative environment where everyone is respected, and great ideas win.

Our benefits include:

  • Generous paid time off policies
  • Savings plans and other financial wellness tools available in each region
  • Hybrid working opportunities
  • Free breakfast, lunch, and snacks daily
  • In-office wellness experiences and reimbursement for select wellness expenses (e.g., gym, personal training and more)
  • Company-sponsored sports teams and fitness events (JPM Corporate Challenge, Cycle for Survival, Wall Street Rides FAR and more)
  • Volunteer opportunities and charitable giving
  • Social events, happy hours, treats, and celebrations throughout the year
  • Workshops and continuous learning opportunities

At Tower, you’ll find a collaborative and welcoming culture, a diverse team and a workplace that values both performance and enjoyment. No unnecessary hierarchy. No ego. Just great people doing great work – together.

Tower Research Capital is an equal opportunity employer.

Skills Required

  • Strong Python for building automation, tooling, and services
  • Hands-on Ansible experience writing playbooks and roles
  • Solid grasp of CI/CD principles, pipelines, testing, and staged rollouts
  • GitOps mindset and experience driving infrastructure changes through version control
  • Working knowledge of Linux, boot process, system services, and debugging
  • Experience automating out-of-band management (BMC/Redfish/IPMI) for firmware and BIOS management
  • Ability to design workflows that fail safely, retry sensibly, and leave an audit trail
  • Clear communication and ability to convert tribal operational knowledge into documented automation
  • Experience with bare-metal provisioning tooling (MAAS, Tinkerbell, Foreman, Ironic)
  • Familiarity with vendor OOB variants (iDRAC, iLO) and Redfish/IPMI specifics
  • Exposure to hardware validation, burn-in, stress testing, or firmware qualification at fleet scale
  • Prior work in datacenter, HPC, or large-fleet environments (hundreds to thousands of machines)
  • Experience with GitOps tooling or declarative infrastructure management

Tower Research Capital Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Tower Research Capital and has not been reviewed or approved by Tower Research Capital.

  • Strong & Reliable Incentives Pay is considered competitive versus broader tech and finance markets, with strong base and bonus potential in quant research/trading and specialized engineering. Feedback suggests total compensation can be particularly attractive early in career and for revenue‑linked roles.
  • Leave & Time Off Breadth Paid time off is described as generous, with policies highlighting substantial annual vacation. Feedback suggests formal allowances are a notable part of the package.
  • Wellbeing & Lifestyle Benefits Complimentary in‑office meals, regular social events, donation matching, and professional development workshops are highlighted as everyday perks. Feedback suggests these cultural and lifestyle benefits are a consistent strength across materials.

Tower Research Capital Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: NEW YORK, NY
1,056 Employees
Year Founded: 1998

What We Do

Founded in 1998 by Mark Gorton, Tower Research Capital is a trading and technology company that has built some of the fastest, most sophisticated electronic trading platforms in the world.

Similar Jobs

inKind Logo inKind

Account Executive

eCommerce • Fintech • Food • Mobile • Social Impact
Easy Apply
Remote or Hybrid
USA
170 Employees
100K-160K Annually

EliseAI Logo EliseAI

Information Technology Manager

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Real Estate
In-Office
New York City, NY, USA
550 Employees
150K-190K Annually

Legora Logo Legora

Director of GTM Systems & Architecture

Artificial Intelligence • Legal Tech • Software
In-Office
New York City, NY, USA
700 Employees
204K-300K Annually

Hinge Logo Hinge

Director of UX Research, Dating Outcomes & Platform

Artificial Intelligence • Machine Learning • Mobile • Social Impact • Software • App development
Hybrid
New York, NY, USA
305 Employees
100K-287K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account