Tower Research Capital, LLC

Software Engineer, Machine Lifecycle

Tower Research Capital, LLC$150K — $250K *
Enterprise Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • Engineering background with a strong enthusiasm for automating physical infrastructure.
  • Proficient in Python for developing robust automation tools, not just scripting.
  • Hands-on experience with Ansible for writing and reviewing playbooks and roles.
  • Solid understanding of CI/CD practices around pipelines and version control for infrastructure.
  • Familiar with GitOps principles, advocating for code-driven infrastructure management.
  • Working knowledge of Linux systems and troubleshooting skills during the boot process.
  • Demonstrates sound engineering judgment in designing resilient workflows.
  • Excellent communication skills for documenting processes and training teams.

Responsibilities

  • Design and build a comprehensive machine lifecycle automation pipeline.
  • Automate hardware bring-up processes utilizing out-of-band management tools.
  • Implement OS provisioning via network boot and unattended installations.
  • Develop and maintain Ansible and Python configurations to eliminate manual runbooks.
  • Employ GitOps and CI/CD methods to manage infrastructure changes and corrections.
  • Create automated validation and burn-in tests for machine readiness.
  • Model and automate the machine lifecycle stages with clear transitions and audit history.
  • Track and log pipeline metrics to identify slow points in the lifecycle process.
  • Collaborate with HPC and datacenter teams to integrate operational knowledge into automation.

Benefits

  • Flexible paid time off policies to promote work-life balance.
  • Access to savings plans and financial wellness resources.
  • Options for hybrid work arrangements.
  • Daily free meals and snacks to enhance workplace experience.
  • Wellness programs and reimbursements for fitness-related expenses.
  • Opportunities to participate in company-sponsored sports and fitness events.
  • Engagement in volunteer efforts and charitable contributions.
  • Regular social events and celebrations to foster team spirit.
  • Access to workshops and continual learning opportunities to support career growth.
Full Job Description
Summary:

This role owns the journey of every machine in our fleet: from the moment a server is racked, cabled, and powered on, to the moment it is fully configured, validated, and available for users. Your mission is to make that journey zero-touch.

You will design and build the automation pipeline that takes a machine through discovery, firmware and BIOS configuration, OS installation, configuration management, health validation and burn-in, and finally handoff into production, treating each stage as code that lives in Git, runs through CI/CD, and can be reviewed, tested, and rolled back like any other software.

The guiding principle is GitOps for physical infrastructure: the desired state of the fleet is declared in a repository, and automation continuously reconciles reality against it. A new machine shows up as a commit; a decommission is a deletion; drift is detected and corrected by the pipeline, not by a person with a checklist.

Responsibilities:
  • Design and build the end-to-end machine lifecycle pipeline: from power-on and network boot through OS install, configuration, validation, and production handoff.
  • Automate hardware bring-up via out-of-band management (BMC, Redfish, IPMI): firmware updates, BIOS settings, boot order, and inventory discovery.
  • Automate OS provisioning with network boot (PXE / UEFI HTTP boot) and unattended installation, so no one ever installs a machine by hand.
  • Write and maintain the Ansible and Python that configure machines into their final roles, replacing manual runbooks with reviewed, versioned code.
  • Apply GitOps and CI/CD principles to the fleet: desired state in Git, changes through merge requests, pipelines that test and apply them, and reconciliation that catches drift.
  • Build automated validation and burn-in: health checks, stress tests, and acceptance criteria a machine must pass before users ever see it.
  • Model the lifecycle as a state machine (new, provisioning, validating, in-service, needs-repair, decommissioned) with clear, automated transitions and an auditable history.
  • Instrument the pipeline with metrics and logging so we always know where a machine is in its lifecycle, and where the process is slow or failing.
  • Work with the HPC and datacenter teams to fold their hard-won operational knowledge into the automation, one stage at a time.


Qualifications:
  • A smart, curious engineer who learns fast and is genuinely excited by the challenge of automating physical infrastructure end to end. This matters more to us than any specific line on your resume.
  • Strong Python for building automation, tooling, and services, not just scripts.
  • Hands-on Ansible experience: writing playbooks and roles you would be happy to code-review, not just run.
  • A solid grasp of CI/CD principles: pipelines, testing, staged rollouts, and the discipline of driving change through version control.
  • An automation-first, GitOps mindset: you believe infrastructure state belongs in Git, and that any task done by hand twice should be code.
  • Working knowledge of Linux: comfortable with the boot process, system services, and debugging when a machine does not come up the way it should. Depth here is a real plus, but interest and trajectory count.
  • Sound engineering judgment: you design workflows that fail safely, retry sensibly, and leave an audit trail.
  • Clear communication and the patience to turn tribal operational knowledge into reliable, documented automation.


Nice to Have:

  • Experience with bare-metal provisioning tooling such as MAAS, Tinkerbell, Foreman, Ironic, or a home-grown equivalent.
  • Familiarity with out-of-band management: BMCs, Redfish, IPMI, and vendor variants like iDRAC or iLO.
  • Exposure to hardware validation and burn-in: stress testing, firmware qualification, or failure prediction at fleet scale.
  • Experience with GitOps tooling or declarative infrastructure management in general.
  • Prior work in datacenter, HPC, or large-fleet environments where machines number in the hundreds or thousands.


Anticipated annual base salary range $150,000-$250,000, plus eligible for discretionary bonus.

Tower's headquarters are in the historic Equitable Building, right in the heart of NYC's Financial District and our impact is global, with over a dozen offices around the world.

Our benefits include:
  • Generous paid time off policies
  • Savings plans and other financial wellness tools available in each region
  • Hybrid working opportunities
  • Free breakfast, lunch, and snacks daily
  • In-office wellness experiences and reimbursement for select wellness expenses (e.g., gym, personal training and more)
  • Company-sponsored sports teams and fitness events (JPM Corporate Challenge, Cycle for Survival, Wall Street Rides FAR and more)
  • Volunteer opportunities and charitable giving
  • Social events, happy hours, treats, and celebrations throughout the year
  • Workshops and continuous learning opportunities

About Tower Research Capital, LLC

Tower Research Capital, LLC is a quantitative trading firm that was founded in 1998. The company uses advanced technology and algorithms to trade in multiple asset classes across global markets. Tower Research Capital, LLC is headquartered in New York City and has offices in North America, Europe, and Asia. The company is known for its innovative approach to trading and its use of cutting-edge technology to analyze market data and make trading decisions. Tower Research Capital, LLC is a privately held company and does not disclose its financial information to the public.
Learn more about Tower Research Capital, LLC
Size
1,000 employees
Industry
Founded
1998

Similar Jobs

More Jobs at Tower Research Capital, LLC

More Enterprise Technology Jobs

Find similar Software Engineer, Machine Lifecycle jobs: