Senior Linux Infrastructure Engineer

Posted 6 Days Ago
Be an Early Applicant
Somerville, MA, USA
Hybrid
Senior level
Information Technology • Software
The Role
Own and operate Northern Light’s private-colocation Linux infrastructure, including bare-metal servers, virtual machines, core services, databases, monitoring, backups, vulnerability management, and hardware lifecycle. Write Ansible automation, maintain security and reliability, lead incident response and disaster recovery activities, coordinate vendors and remote hands, and support infrastructure modernization. The role includes occasional after-hours maintenance and periodic travel to the Somerville, Massachusetts office and datacenter.
Summary Generated by Built In

About Northern Light

Northern Light provides the world's most sophisticated machine learning-powered competitive intelligence platform for market research. For over 25 years, we've been helping Fortune 1000 enterprises make smarter, faster, and more informed decisions through our award-winning SinglePoint knowledge management platform. Our clients include global leaders across technology, pharmaceuticals, telecommunications, and life sciences who depend on us to transform fragmented data into strategic clarity.

We're a company that takes pride in our compulsive drive to provide exceptional client support. We wake up each day ready to tackle the challenges of knowledge management and we never stand still. Our recent innovations include generative AI capabilities, machine learning insights, and advanced competitive intelligence automation.

The Opportunity

Northern Light is seeking a Senior Linux Infrastructure Engineer to take hands-on ownership of the Linux infrastructure behind our platform's compute-heavy backend, which runs on our own hardware in a private colocation cage in Somerville, MA. You will be the primary owner of that datacenter environment: the person who keeps it reliable and secure today, and who leads its next phase, including a hardware refresh across the fleet and deeper automation of the environment.

On our Platform team you will own Linux infrastructure, working closely with the engineering, security, and operations teams, including the team that runs our customer-facing frontend in AWS. Our philosophy is to buy our platforms rather than build them: where a supported, vendor-backed product exists, we run it and follow the vendor's best practices instead of maintaining our own substitute. Automation on top of those platforms is very much your work, and we want someone who writes it well and knows how to get the most out of a vendor relationship.

What You'll Own

  • A fleet of ~60 HPE ProLiant DL360 Gen9/Gen10 bare-metal servers and ~100 virtual machines running RHEL-family Linux (currently Oracle Linux 9).
  • Virtualization: ~50 VMs on VMware, moving to a new platform this year, and a KVM/libvirt environment running the remaining ~50 VMs, which serve as Kubernetes nodes managed by DevOps.
  • Core infrastructure services: BIND, LDAP, mail relay, SFTP, and Foreman.
  • Self-hosted applications and database servers the product teams depend on: GitLab Self-Managed Premium, MariaDB, PostgreSQL, and MongoDB. These are yours at the system level; application and DBA-level tuning stay with the teams that use them.
  • Operational tooling: inventory (NetBox), monitoring (LogicMonitor / Site24x7 / Grafana), vulnerability management (Tenable One), automation hub (Ansible Automation Platform), backups (Veeam).
  • The physical environment: a six-rack private colo cage in Somerville, MA, including spare-parts inventory and hardware lifecycle tracking.
  • Routine network operations, such as wiring and top-of-rack switch port configuration. Major network reconfiguration and network device patching sit with our network engineering partner.

What You'll Do

  • Keep the fleet current and consistent: patch, harden, and upgrade Linux servers on a controlled cadence through managed repositories and staged rollouts, and install, upgrade, and maintain the applications and database servers to vendor guidance — largely through Ansible roles and playbooks you write and the Ansible Automation Platform you operate.
  • Run the vulnerability management cycle: scheduled Tenable scans, triage of findings, remediation prioritization, patching or documented mitigation, and compliance reporting that stands up to customer security reviews and audits.
  • Own infrastructure backups (policy, platform administration, restore testing) and partner with the application team on disaster recovery planning and exercises.
  • Keep monitoring and logging reliable and free of noise: complete coverage of the fleet, alerts that fire on real problems and not on everything else, and log collection you can trust when investigating an incident.
  • Lead incident response for infrastructure issues, run post-incident reviews, drive corrective actions to closure, and keep the resulting SOPs, runbooks, and infrastructure diagrams current and clear to technical and non-technical readers alike.
  • Run the datacenter as a remotely operated facility: accurate NetBox records, labeled cabling, iLO/out-of-band access to everything, spare parts on the shelf, and clear work orders for remote hands.
  • Work with vendors and our colo provider on hardware lifecycle and capacity, and coordinate scheduled maintenance windows and network changes with our network partner.
  • Participate in occasional scheduled after-hours maintenance (historically 3–5 times per year).

What We Require

  • Substantial production Linux systems engineering experience, typically 7+ years, including several years where you were accountable for the reliability and security of the environment, whether alone or as a senior member of a small team. A BS or MS in Computer Science, Computer Engineering, or Information Technology is a plus, but practical experience matters more to us.
  • Deep RHEL-family Linux skills (RHEL, Oracle Linux, Rocky, Alma, CentOS): systemd, kernel and performance tuning, storage (LVM, RAID, NFS), and networking (bonding, VLANs, firewalld/iptables).
  • Production experience administering a virtualization platform (VMware vSphere, KVM/libvirt, OpenShift Virtualization, Proxmox, Hyper-V, or similar).
  • Hands-on Ansible authoring: you have written and maintained roles and playbooks, not only run them.
  • Experience running a patch and vulnerability management program in production: scheduled scanning with Tenable/Nessus, Qualys, Rapid7, or similar, interpreting results, and driving remediation across a fleet.
  • Experience installing and operating server applications and database servers at the system level (packaging, storage, TLS, access control, backups) from vendor documentation, including the judgment to plan a database upgrade that can be rolled back and to treat a backup as unproven until it has been restored.
  • Experience with enterprise server hardware (HPE ProLiant or equivalent): out-of-band management (iLO/IPMI), firmware, diagnostics, and component replacement, and comfort doing occasional physical work in a datacenter.
  • Experience designing or materially improving highly available, redundant infrastructure, and a track record of leading incident response and writing useful root-cause analyses.
  • Solid networking fundamentals: enough to make routine switch changes yourself, and to diagnose and scope switch, firewall, and load-balancer issues well enough to hand them to network engineers.
  • Strong documentation habits and clear written and spoken communication.

Prior experience with the specific products named above is a plus but not required; we expect a strong engineer to pick them up here.

Job Details

Job Type: Full-Time

Location: Remote, within a two- to three-hour drive of Somerville, MA; local candidates are especially welcome. The role is home-based, and you will come to our Somerville office and datacenter for hands-on datacenter projects and for planning and brainstorming sessions with the team, typically a couple of consecutive days at a time, every few weeks. Travel for on-site days is reimbursed with pre-approval. Eastern Time working hours.

Datacenter work: You decide how much of it you do yourself. The colo offers a remote-hands service that can handle routine tasks such as drive and component swaps, cabling, and receiving shipments, provided the planning, documentation, and work orders behind them are in good order. When you do work in the cage yourself, expect elevated noise levels and variable temperatures.

Work authorization requirements: Must be authorized to work in the United States (unfortunately, we cannot sponsor visas).

Why Join Northern Light

  • Join a company shaping the future of competitive and market intelligence.
  • Work with a collaborative, high-performing team that values creativity, experimentation, and measurable results.
  • Competitive salary, benefits, and professional growth opportunities.
  • Own a real production environment end to end, on supported enterprise platforms with vendor backing.
  • Regular team offsites and opportunities for cross-functional collaboration.

Working at Northern Light

Northern Light is based in the Boston, MA area. Northern Light is proud to provide equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local laws. Northern Light SinglePoint, LLC participates in E-Verify and will provide the federal government with Form I-9 information to confirm employment eligibility for all new hires. For more information, visit www.e-verify.gov.

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Skills Required

  • Substantial production Linux systems engineering experience, typically 7 or more years
  • Several years accountable for the reliability and security of a Linux environment
  • Deep RHEL-family Linux experience, including systemd, kernel and performance tuning, storage, and networking
  • Production administration experience with a virtualization platform such as VMware vSphere, KVM/libvirt, OpenShift Virtualization, Proxmox, or Hyper-V
  • Hands-on experience authoring and maintaining Ansible roles and playbooks
  • Production experience operating patch and vulnerability management programs using Tenable, Nessus, Qualys, Rapid7, or similar tools
  • Experience installing and operating server applications and database servers at the system level
  • Experience planning database upgrades with rollback procedures and validating backups through restoration
  • Experience with enterprise server hardware, out-of-band management, firmware, diagnostics, and component replacement
  • Comfort performing occasional physical work in a datacenter
  • Experience designing or materially improving highly available and redundant infrastructure
  • Track record leading incident response and writing root-cause analyses
  • Solid networking fundamentals, including routine switch changes and troubleshooting switch, firewall, and load-balancer issues
  • Strong technical documentation and written and spoken communication skills
  • Bachelor’s or master’s degree in Computer Science, Computer Engineering, or Information Technology
  • Prior experience with the specifically named infrastructure products
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Boston, MA
48 Employees
Year Founded: 1996

What We Do

Northern Light has been providing knowledge management platforms for competitive intelligence and market research insights to global enterprises since 1996. Northern Light’s current clients include Fortune 1000 leaders across multiple industries such as information technology, pharmaceuticals, telecommunications, and life sciences. Northern Light has over 200,000 users of its strategic research portals worldwide. Headquartered in Boston, Massachusetts, Northern Light has repeatedly been recognized as one of KMWorld’s “AI 50” – the companies empowering intelligent knowledge management – and has won the KMWorld Readers’ Choice Award multiple times.

Similar Jobs

NVIDIA Logo NVIDIA

Systems Engineer

Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
In-Office or Remote
5 Locations
21960 Employees
184K-357K Annually

Liberty Mutual Insurance Logo Liberty Mutual Insurance

Inside Sales Representative

Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Remote or Hybrid
12 Locations
40000 Employees
45K-85K Annually
Hybrid
4 Locations
289097 Employees

PwC Logo PwC

Pricing and Revenue Consulting Senior Manager - Consumer Markets Sector

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
58 Locations
370000 Employees
124K-280K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account