Back to search
Listing unavailableData & AnalyticsOn-site

Lab Administrator

DDN · On-site

This listing is no longer verified as available.

MeritLog keeps this source-backed description for reference. Availability is not verified, and there is no application link here.

Last seen by MeritLog September 9, 2026Source: AshbySource version: ashby-public-job-posting-v1

Source: the employer's Ashby job board. Open the original listing for current details. Availability is not verified for this retained page.

Job details

Work model
On-site
Salary
Not listed by source
Location
Columbia Office

What the role asks for

What you'd do

  • Install, configure, and maintain lab hardware (servers, storage arrays, NICs, switches) and software (OS images, drivers, monitoring tools)
  • Manage user access, accounts, and permissions across lab systems
  • Monitor system health, performance, and capacity (CPU, memory, storage, network utilization)
  • Troubleshoot hardware/software issues and coordinate with vendors for support/RMAs
  • Maintain documentation: network diagrams, asset inventories, configuration baselines, SOPs
  • Implement and enforce security policies (patching, firewall rules, access controls)
  • Manage backups, disaster recovery procedures, and data retention policies
  • Support researchers/engineers with environment setup for experiments, benchmarks, or testing (e.g., provisioning compute nodes, storage volumes, network configs)
  • Track licensing, warranties, and hardware lifecycle (procurement to decommissioning)
  • Coordinate lab scheduling/resource allocation if shared across teams
  • Strong Linux administration experience (Ubuntu/RHEL/CentOS)
  • Networking fundamentals (TCP/IP, VLANs, bonding/LACP, basic troubleshooting)
  • Experience with storage systems (SAN/NAS, parallel filesystems like Lustre/GPFS a plus)
  • Scripting ability (Bash, Python) for automation
  • Familiarity with virtualization/containerization (KVM, Docker) is a plus
  • Experience with monitoring tools (Prometheus/Grafana, Nagios, Zabbix)
  • Understanding of hardware components (CPUs, NUMA architecture, PCIe, NICs/RDMA)
  • Good documentation and communication skills

What they're asking for

  • Experience with HPC/AI infrastructure (InfiniBand, RDMA, GPU clusters)SkillPreferred
  • Experience with configuration management (Ansible, Puppet)SkillPreferred

Parsed by MeritLog from the employer’s own posting. The full description follows below.

Job description

Role Summary Responsible for the setup, maintenance, security, and day-to-day operation of a technical/research lab environment (servers, storage systems, networking equipment, and lab-issued workstations), ensuring high availability, performance, and compliance with organizational policies. Key Responsibilities - Install, configure, and maintain lab hardware (servers, storage arrays, NICs, switches) and software (OS images, drivers, monitoring tools) - Manage user access, accounts, and permissions across lab systems - Monitor system health, performance, and capacity (CPU, memory, storage, network utilization) - Troubleshoot hardware/software issues and coordinate with vendors for support/RMAs - Maintain documentation: network diagrams, asset inventories, configuration baselines, SOPs - Implement and enforce security policies (patching, firewall rules, access controls) - Manage backups, disaster recovery procedures, and data retention policies - Support researchers/engineers with environment setup for experiments, benchmarks, or testing (e.g., provisioning compute nodes, storage volumes, network configs) - Track licensing, warranties, and hardware lifecycle (procurement to decommissioning) - Coordinate lab scheduling/resource allocation if shared across teams Required Skills/Qualifications - Strong Linux administration experience (Ubuntu/RHEL/CentOS) - Networking fundamentals (TCP/IP, VLANs, bonding/LACP, basic troubleshooting) - Experience with storage systems (SAN/NAS, parallel filesystems like Lustre/GPFS a plus) - Scripting ability (Bash, Python) for automation - Familiarity with virtualization/containerization (KVM, Docker) is a plus - Experience with monitoring tools (Prometheus/Grafana, Nagios, Zabbix) - Understanding of hardware components (CPUs, NUMA architecture, PCIe, NICs/RDMA) - Good documentation and communication skills Nice to Have - Experience with HPC/AI infrastructure (InfiniBand, RDMA, GPU clusters) - Experience with configuration management (Ansible, Puppet)

Keep exploring

Available Data & Analytics roles

These current listings are available to explore now.

Search all jobs

Privacy choices

Analytics and advertising stay off unless you allow them. Private data stays out.

Read the privacy notice