
Posted 23 days ago
HPC Infrastructure & Cluster Engineer
AI Summary
Manages a dedicated customer compute cluster for DoW/Joint Force operations, handling Linux administration, hardware monitoring, workload orchestration, and InfiniBand GPU networking to ensure high availability and performance for AI/ML workloads.
About this role
Job Overview:
We are seeking an Infrastructure & Cluster Engineer to manage the administration, health, and performance of the foundational compute environmen. In this role, you will be responsible for the end-to-end administration of a dedicated customer compute cluster. Your primary mission is to ensure a highly available, secure, and optimized hardware foundation. By maintaining a robust infrastructure, you will directly contribute to the critical technology integration and performance engineering efforts, ensuring a highly reliable platform for integrating and executing complex customer workloads.
Here, your work is more than a job- it's a journey in innovation. With opportunities to work on high-impact projects, access to the latest technologies, and a culture that thrives on creativity and collaboration, INflow Federal is where your expertise can truly make a difference.
Specific Duties and Responsibilities:
Required Skills:
Preferred Skills:
- Familiarity with parallel file systems and high-throughput storage architectures.
- Prior experience engineering or managing high-speed GPU-to-GPU communication topologies.
Skills
Explore related jobs
More jobs at INflow Federal
Browse these categories
Market data for this role
All reports →- SeriesRole reportsOne role family at a time: how many openings, what changed this week, who is hiring, what it pays.
- SeriesSalary reportsWhat employers publish in job postings, by level and workplace. Not self-reported pay.
- Market overviewState of tech hiring, September 2026: up 4.8%Tech hiring rose 4.8% month over month in September 2026, with 411,122 new listings. Customer support and account executive roles led the growth.