HPC Platform Engineer
NorthMark Strategies · Dallas–Fort Worth, TX
📍 Dallas, TXvia workdayFirst listed here 2026-08-21
Apply on company site ↗
Career Moonshot pulls this listing straight from the employer's hiring system — no recruiter middleman, no reposts. Applying takes you directly to NorthMark Strategies.
THE COMPANY
NorthMark Compute & Cloud (NMC²) is backed by dedicated leadership and investment, with a clear mission as it operates at the bleeding edge of technology. Its goal is to scale and enhance the high-performance computing (HPC) and cloud infrastructure that supports its clients' research, production, and delivery, enabling breakthroughs that shape the industries of tomorrow. Its engineers build critical infrastructure to eliminate friction in scientific research, simulations, analysis, and decision-making, accelerating discovery and driving faster innovation.
THE POSITION
As an HPC Platform Engineer, you will be responsible for implementing and supporting state-of-the-art datacenter infrastructure solutions that support high-performance computing and scientific research. You will collaborate with cross-functional teams, including researchers, system administrators, network engineers, and data scientists, to understand their requirements and create efficient and scalable solutions.
Your expertise in HPC technologies and emerging trends will be instrumental in driving innovation and optimizing performance within the datacenter environment. You will contribute to the advancement of scientific research and innovation by designing and optimizing cutting-edge infrastructure, and you will be trusted to own systems end to end, from architecture and deployment through steady-state operation and capacity planning.
The role is based out of the Victory Commons office in Dallas, TX. As a senior engineer on the team, you will provide technical leadership and mentorship to junior engineers and help set the standards other groups build against.
RESPONSIBILITIES
Develop and refine datacenter architecture blueprints and guidelines considering performance, scalability, security, and efficiency, and design and implement solutions for compute, storage, networking, and cooling infrastructure that align with HPC requirements.
Continuously evaluate and enhance the infrastructure to maximize HPC performance and resource utilization, identifying and addressing potential bottlenecks and performance gaps using industry best practices and cutting-edge technologies.
Collaborate with system administrators and engineers to ensure seamless integration and deployment of HPC systems, overseeing hardware and software installation, configuration, and testing activities.
Stay current with emerging HPC technologies, tools, and methodologies, conducting research and feasibility studies on new hardware and software solutions, evaluating vendor offerings, and providing recommendations for procurement.
Monitor and analyze performance metrics to identify issues and implement necessary optimizations, troubleshooting complex system problems with technical teams to ensure efficient resolution and minimal impact on operations.
Collaborate with security teams to design and implement robust security measures within the infrastructure, ensuring compliance with relevant industry standards and regulations, such as HIPAA or GDPR, in data handling and storage.
Create comprehensive technical documentation, including architectural diagrams, standard operating procedures, and configuration guidelines, and prepare regular reports on performance, capacity planning, and future infrastructure requirements.
Collaborate effectively with cross-functional teams to foster a culture of knowledge sharing and innovation, providing technical leadership and mentorship to junior team members as they adopt best practices and grow their skill sets.
REQUIREMENTS
Bachelor's Degree or equivalent experience.
Minimum 5 years of experience as an HPC engineer or in a similar role, with a strong focus on engineering and optimization.
In-depth knowledge of HPC technologies, including parallel computing, distributed storage systems, job scheduling, InfiniBand and Ethernet networking, GPU acceleration, and job scheduling frameworks.
Familiarity with industry-standard tools and software used in HPC environments, such as Slurm, PBS Pro, Lustre, GPFS, OpenStack, and containerization technologies (e.g., Docker, Kubernetes).
Experience with automation tools including Python, Ansible, and Puppet or Chef.
Experience with monitoring tools including Prometheus, Ganglia, Nagios, SNMP, and Telegraf.
ZFS and NiFi are a plus, as is experience with CFD (Computational Fluid Dynamics) workloads and associated HPC optimization.
Familiarity with security protocols and compliance requirements in the context of datacenter operations.
Strong problem-solving and analytical skills, with the ability to identify and resolve complex technical issues.
Excellent communication and interpersonal skills, a detail-oriented mindset with a strong focus on documentation and adherence to standards, and the ability to adapt to a fast-paced and rapidly evolving technological landscape.
It is impossible to list every requirement for, or responsibility of, any position. Similarly, we cannot identify all the skills a position may require since job responsibilities and the Company’s needs may change over time. Therefore, the above job description is not comprehensive or exhaustive. The Company reserves the right to adjust, add to or eliminate any aspect of the above description. The Company also retains the right to require all employees to undertake additional or different job responsibilities when necessary to meet business needs.
Must be legally authorized to work in the United States without the need for employer sponsorship, now or at any time in the future.
Benefits & Perks:
Company-Paid Lunch Stipend : Lunch is provided via GrubHub
Company-Paid Benefits: 100% Employer-Paid Medical in our High Deductible Health Plan, Dental and Vision benefits for employees and their families, 16 weeks of Paid Parental Leave, Employee Assistance Program, Life insurance, Short-Term Disability and Long-Term Disability
401(k): Company will match 100% of
More Dallas–Fort Worth, TX jobs
Dallas–Fort Worth, TX jobs · Browse all locations