Back to search
Northern Light Linkedin · Posted yesterday

Senior Linux Infrastructure Engineer

Somerville, Massachusetts, United States

Linkedin
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

About Northern Light

Northern Light provides the world's most sophisticated machine learning-powered competitive intelligence platform for market research. For over 25 years, we've been helping Fortune 1000 enterprises make smarter, faster, and more informed decisions through our award-winning SinglePoint knowledge management platform. Our clients include global leaders across technology, pharmaceuticals, telecommunications, and life sciences who depend on us to transform fragmented data into strategic clarity.

We're a company that takes pride in our compulsive drive to provide exceptional client support. We wake up each day ready to tackle the challenges of knowledge management and we never stand still. Our recent innovations include generative AI capabilities, machine learning insights, and advanced competitive intelligence automation.

The Opportunity

Northern Light is seeking a Senior Linux Infrastructure Engineer to take hands-on ownership of the Linux infrastructure behind our platform's compute-heavy backend, which runs on our own hardware in a private colocation cage in Somerville, MA. You will be the primary owner of that datacenter environment: the person who keeps it reliable and secure today, and who leads its next phase, including a hardware refresh across the fleet and deeper automation of the environment.

On our Platform team you will own Linux infrastructure, working closely with the engineering, security, and operations teams, including the team that runs our customer-facing frontend in AWS. Our philosophy is to buy our platforms rather than build them: where a supported, vendor-backed product exists, we run it and follow the vendor's best practices instead of maintaining our own substitute. Automation on top of those platforms is very much your work, and we want someone who writes it well and knows how to get the most out of a vendor relationship.

What You'll Own

  • A fleet of ~60 HPE ProLiant DL360 Gen9/Gen10 bare-metal servers and ~100 virtual machines running RHEL-family Linux (currently Oracle Linux 9).
  • Virtualization: ~50 VMs on VMware, moving to a new platform this year, and a KVM/libvirt environment running the remaining ~50 VMs, which serve as Kubernetes nodes managed by DevOps.
  • Core infrastructure services: BIND, LDAP, mail relay, SFTP, and Foreman.
  • Self-hosted applications and database servers the product teams depend on: GitLab Self-Managed Premium, MariaDB, PostgreSQL, and MongoDB. These are yours at the system level; application and DBA-level tuning stay with the teams that use them.
  • Operational tooling: inventory (NetBox), monitoring (LogicMonitor / Site24x7 / Grafana), vulnerability management (Tenable One), automation hub (Ansible Automation Platform), backups (Veeam).
  • The physical environment: a six-rack private colo cage in Somerville, MA, including spare-parts inventory and hardware lifecycle tracking.
  • Routine network operations, such as wiring and top-of-rack switch port configuration. Major network reconfiguration and network device patching sit with our network engineering partner.

What You'll Do

  • Keep the fleet current and consistent: patch, harden, and upgrade Linux servers on a controlled cadence through managed repositories and staged rollouts, and install, upgrade, and maintain the applications and database servers to vendor guidance — largely through Ansible roles and playbooks you write and the Ansible Automation Platform you operate.
  • Run the vulnerability management cycle: scheduled Tenable scans, triage of findings, remediation prioritization, patching or documented mitigation, and compliance reporting that stands up to customer security reviews and audits.
  • Own infrastructure backups (policy, platform administration, restore testing) and partner with the application team on disaster recovery planning and exercises.
  • Keep monitoring and logging reliable and free of noise: complete coverage of the fleet, alerts that fire on real problems and not on everything else, and log collection you can trust when investigating an incident.
  • Lead incident response for infrastructure issues, run post-incident reviews, drive corrective actions to closure, and keep the resulting SOPs, runbooks, and infrastructure diagrams current and clear to technical and non-technical readers alike.
  • Run the datacenter as a remotely operated facility: accurate NetBox records, labeled cabling, iLO/out-of-band access to everything, spare parts on the shelf, and clear work orders for remote hands.
  • Work with vendors and our colo provider on hardware lifecycle and capacity, and coordinate scheduled maintenance windows and network changes with our network partner.
  • Participate in occasional scheduled after-hours maintenance (historically 3–5 times per year).

What We Require

  • Substantial production Linux systems engineering experience, typically 7+ years, including several years where you were accountable for the reliability and security of the environment, whether alone or as a senior member of a small team. A BS or MS in Computer Science, Computer Engineering, or Information Technology is a plus, but practical experience matters more to us.
  • Deep RHEL-family Linux skills (RHEL, Oracle Linux, Rocky, Alma, CentOS): systemd, kernel and performance tuning, storage (LVM, RAID, NFS), and networking (bonding, VLANs, firewalld/iptables).
  • Production experience administering a virtualization platform (VMware vSphere, KVM/libvirt, OpenShift Virtualization, Proxmox, Hyper-V, or similar).
  • Hands-on Ansible authoring: you have written and maintained roles and playbooks, not only run them.
  • Experience running a patch and vulnerability management program in production: scheduled scanning with Tenable/Nessus, Qualys, Rapid7, or similar, interpreting results, and driving remediation across a fleet.
  • Experience installing and operating server applications and database servers at the system level (packaging, storage, TLS, access control, backups) from vendor documentation, including the judgment to plan a database upgrade that can be rolled back and to treat a backup as unproven until it has been restored.
  • Experience with enterprise server hardware (HPE ProLiant or equivalent): out-of-band management (iLO/IPMI), firmware, diagnostics, and component replacement, and comfort doing occasional physical work in a datacenter.
  • Experience designing or materially improving highly available, redundant infrastructure, and a track record of leading incident response and writing useful root-cause analyses.
  • Solid networking fundamentals: enough to make routine switch changes yourself, and to diagnose and scope switch, firewall, and load-balancer issues well enough to hand them to network engineers.
  • Strong documentation habits and clear written and spoken communication.

Prior experience with the specific products named above is a plus but not required; we expect a strong engineer to pick them up here.

Job Details

Job Type: Full-Time

Location: Remote, within a two- to three-hour drive of Somerville, MA; local candidates are especially welcome. The role is home-based, and you will come to our Somerville office and datacenter for hands-on datacenter projects and for planning and brainstorming sessions with the team, typically a couple of consecutive days at a time, every few weeks. Travel for on-site days is reimbursed with pre-approval. Eastern Time working hours.

Datacenter work: You decide how much of it you do yourself. The colo offers a remote-hands service that can handle routine tasks such as drive and component swaps, cabling, and receiving shipments, provided the planning, documentation, and work orders behind them are in good order. When you do work in the cage yourself, expect elevated noise levels and variable temperatures.

Work authorization requirements: Must be authorized to work in the United States (unfortunately, we cannot sponsor visas).

Why Join Northern Light

  • Join a company shaping the future of competitive and market intelligence.
  • Work with a collaborative, high-performing team that values creativity, experimentation, and measurable results.
  • Competitive salary, benefits, and professional growth opportunities.
  • Own a real production environment end to end, on supported enterprise platforms with vendor backing.
  • Regular team offsites and opportunities for cross-functional collaboration.

Working at Northern Light

Northern Light is based in the Boston, MA area. Northern Light is proud to provide equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local laws. Northern Light SinglePoint, LLC participates in E-Verify and will provide the federal government with Form I-9 information to confirm employment eligibility for all new hires. For more information, visit www.e-verify.gov.

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search