About the role
What will you do at Holtec Europe?
The Senior Systems Engineer will be responsible for managing and supporting Holtec's enterprise systems both on-premise and in the cloud. A core function of this role includes HPC Cluster Administration supporting a high-performance computing environment used for scientific, engineering, and analytics workloads.
In this role, you will
- administer Linux-based infrastructure, manage software environments, optimize cluster performance, secure sensitive data, and partner directly with users to translate workload requirements into reliable, scalable compute solutions.
- Location: Camden, NJ
- Primary
Responsibilities
- Design, implement, and maintain complex enterprise systems (on-premise / cloud) to support mission-critical workloads, ensuring high availability and scalability.
- Administer and maintain Linux-based HPC cluster infrastructure supporting diverse research and engineering workloads.
- Install, configure, and manage scientific and engineering software, compilers, libraries, and module-based software environments.
- Tune applications and system configurations to improve performance, throughput, and resource utilization.
- Experience with configuring High Availability and redundancy within Linux for network interfaces and storage.
- Manage HPC job scheduling platforms such as Altair PBS, Slurm, or LSF, including queues, priorities, reservations, and policy enforcement.
- Work directly with users and stakeholders to capture computational, storage, and workflow requirements and translate them into system configurations.
- Provision and support pre-processing and post-processing systems, including data staging and workflow integration.
- Architect and manage Azure Cloud environments, including Azure Gov Cloud, to meet compliance and security requirements for government and regulated workloads.
- Deploy, integrate, and support infrastructure solutions that support business stakeholder AI workloads.
- Implement and maintain security controls for sensitive workloads, including access control, encryption, auditing, and compliance alignment.
- Troubleshoot cluster, scheduler, software, and job submission issues across compute, storage, and user environments.
- Automate administrative tasks and standardize operations using scripting and infrastructure management tools.
- Create and maintain technical documentation, configuration records, operational procedures, and user guidance.
Requirements
- 10+ years of experience administering Linux-based compute environments including High Performance Clusters.
- Expert in Linux administration (preferably RHEL or Rocky Linux).
- Experience in Microsoft setting (Azure / Entra / Active Directory / Windows Administration)
- Experience with HPC workload managers and schedulers, preferably Altair PBS; familiarity with Slurm or LSF is also valuable.
- Proficiency in software installation, compilation, dependency management, and container technologies.
- High level of expertise with scripting and automation skills using Bash, Python, or similar tools.
- Experience supporting virtual machines and workstations, including GPU-related configuration.
- Knowledge of security practices for protecting sensitive data and applications in research, engineering, or regulated environments.
- Experience collaborating with engineering, research, or analytics teams in high-demand compute environments.
- Experience with Infrastructure code (Bicep or Terraform) and configuration management tools (ie. Ansible or Puppet) for automation preferred
- Background in scientific computing, parallel computing, or application optimization preferred
- Familiarity with compliance and security frameworks such as NIST or CIS preferred
- A Generation Ahead by Design™Investing in people who power the future.
- We believe great work deserves great rewards. Our benefits go beyond the basics, reflecting our commitment to supporting your whole self…your health, your future, and your career growth. This includes:
Compensation
Annual compensation based on skills and experience ranging from $135,000 - $155,000
Benefits
Industry‑leading medical, dental, and vision coverage with a generous employer premium contribution and Day 1 eligibility!
FLEX Program – Hybrid work opportunities available for eligible roles, where business needs and role requirements allow
401(k) retirement plan with up to a 5% company match and immediate vesting to support long‑term financial security
Paid time off and 11 paid holidays to support rest, balance, and recharge
Wellness program offering rewards for participation in company‑supported health initiatives
Commuter benefits supporting mass transit and commuter parking options
Company‑paid life and AD&D insurance for added peace of mind
Education assistance to support continued education and professional growth
Employee support and voluntary benefits, including an Employee Assistance Program and optional coverage such as disability, legal, identity theft, and insurance programs
Benefits
eligibility and offerings may vary based on role, location, and employment status
If you’re driven to solve meaningful challenges and make a lasting impact, we invite you to build your career at Holtec—a generation ahead by design™.
Which skills does this role require?
Make your next move
Build a shortlist and prepare
Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.
- Build a focused shortlist before you applyCompare role requirements with your experience and give each application a clear reason.
- Practice explaining your experience in an interviewRehearse your answers before meeting the hiring team.
Other roles to compare
Review the responsibilities and requirements before adding an opening to your shortlist.
Senior Principal Engineer, Enterprise Networking Architecture, Automation and AI
Equinix · Toronto, Ontario, Canada
Orchestration Workload Engineer - ACE - AI Factory
Roche · Kaiseraugst, Aargau, Switzerland
Senior SRAM Circuit Design Engineer - AI & HPC (7708)
TSMC · San Jose, California, United States
AI Workflow Engineer
Scout Motors Inc. · Charlotte, North Carolina, United States
AI Outcome Customer Engineer, Forward Deployed Engineering
Google · Atlanta, Georgia, United States
AI Risk Engineer
Bright Vision Technologies · Columbus, Ohio, United States
Role information can change. Confirm current details on the original application page.
