Role OverviewSupport enterprise AI mission systems by designing, developing, and optimizing GPU clusters, with deep focus on operating systems, hardware, GPU platforms, and high-speed networking in a secure customer environment.
What You Will Do
Design, configure, and maintain GPU clusters, collaborate with a multidisciplinary team, work with AI/ML engineers, optimize GPU drivers, analyze GPU performance, and build debugging tools.
Why It Might Be a Fit
The ideal candidate will have strong expertise with Linux distributions, excellent problem-solving skills, and the ability to collaborate within a team.
Requirements
- Active TS/SCI with ability to obtain a CI Polygraph
- Bachelor's degree with a minimum of ten years of experience in the category field
- Experience managing NVIDIA GPU data center platforms, including DGX, HGX, H200, H100, and L4s
- Knowledge of enterprise server components, including storage/network controllers, HBAs, and SSDs
- Strong expertise with Linux distributions, including RHEL, Ubuntu, Oracle, and Rocky
- Meet DoD 8570.11 IAT Level II certification requirements at a minimum; IAT Level III is also acceptable
- U.S. citizenship is required due to the nature of the government contracts supported
]]>