Senior HPC Hardware Engineer
NorthMark Strategies - Dallas-Fort Worth Metroplex
Hiring: Senior HPC Hardware Engineer Company: NorthMark Strategies Location: Dallas-Fort Worth Metroplex Job Posted Time: 2026-09-11 13:14:33 Target Skills & Keywords : Ansible, Chef, Docker, GDPR, HIPAA, Kubernetes, Prometheus, Puppet, Python About the job Experience: •5 years of experience as an HPC engineer or in a similar role, with a strong focus on engineering and optimization. •Minimum 5 years of experience as an HPC engineer or in a similar role, with a strong focus on engineering and optimization. Required Skills: •Develop and refine datacenter architecture blueprints and guidelines considering performance, scalability, security, and efficiency, and design and implement solutions for compute, storage, networking, and cooling infrastructure that align with HPC requirements. •Continuously evaluate and enhance the infrastructure to maximize HPC performance and resource utilization, identifying and addressing potential bottlenecks and performance gaps using industry best practices and cutting-edge technologies. •Partner cross-functionally with system administrators and engineers to ensure seamless integration and deployment of HPC systems, overseeing hardware and software installation, configuration, and testing activities. •Stay current with emerging HPC technologies, tools, and methodologies, conducting research and feasibility studies on new hardware and software solutions, evaluating vendor offerings, and providing recommendations for procurement. •Monitor and analyze performance metrics to identify issues and implement necessary optimizations, troubleshooting complex system problems with technical teams to ensure efficient resolution and minimal impact on operations. •Partner cross-functionally with security teams to design and implement robust security measures within the infrastructure, ensuring compliance with relevant industry standards and regulations, such as HIPAA or GDPR, in data handling and storage. •Create comprehensive technical documentation, including architectural diagrams, standard operating procedures, and configuration guidelines, and prepare regular reports on performance, capacity planning, and future infrastructure requirements. •Collaborate effectively with cross-functional teams to foster a culture of knowledge sharing and innovation, providing technical leadership and mentorship to junior team members as they adopt best practices and grow their skill sets. Qualifications: •Bachelor's Degree or equivalent experience. •In-depth knowledge of HPC technologies, including parallel computing, distributed storage systems, job scheduling, InfiniBand and Ethernet networking, GPU acceleration, and job scheduling frameworks. •Operational familiarity with industry-standard tools and software used in HPC environments, such as Slurm, PBS Pro, Lustre, GPFS, OpenStack, and containerization technologies (e.g., Docker, Kubernetes). •ZFS and NiFi are a plus, as is experience with CFD (Computational Fluid Dynamics) workloads and associated HPC optimization. •Operational familiarity with security protocols and compliance requirements in the context of datacenter operations. •Strong problem-solving and analytical skills, with the ability to identify and resolve complex technical issues. •Excellent communication and interpersonal skills, a detail-oriented mindset with a strong focus on documentation and adherence to standards, and the ability to adapt to a fast-paced and rapidly evolving technological landscape. •Must be legally authorized to work in the United States without the need for employer sponsorship, now or at any time in the future. Compensation: •Company-Paid Lunch Stipend: Lunch is provided via GrubHub Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!