About this opportunity
Hamilton Barnes lists this Storage Engineer (AI Infrastructure) - Hosting opportunity in workfromhome, California. Review the employer’s description below for duties, qualifications and application requirements.
Job description
Looking for a role with plenty of growth opportunities?
Join a rapidly scaling AI cloud infrastructure provider building next-generation GPU platforms for large-scale AI training, experimentation, and inference, significantly expanding operations across the United States alongside continued international growth.
This is a great opportunity for a Senior Storage Engineer to take ownership of the high-performance storage layer underpinning large GPU clusters and AI workloads at scale. The ideal candidate will work closely with infrastructure, networking, and platform engineering teams to design and optimize storage environments capable of supporting massive throughput, low latency, and highly parallelized AI training workloads.
Responsibilities:
Design, deploy, and operate large-scale high-performance storage platforms supporting AI and HPC workloads
Manage and optimize distributed storage environments for large GPU training and inference clusters
Work closely with platform, compute, and networking teams to ensure end-to-end infrastructure performance
Troubleshoot storage bottlenecks, latency issues, throughput constraints, and data flow inefficiencies
Contribute to storage architecture strategy, scalability planning, and operational best practices
Automate storage provisioning, monitoring, and lifecycle management processes
Support performance tuning across parallel file systems, object storage, and AI data pipelines
Implement observability and capacity planning solutions for petabyte-scale environments
Skills/Must Have:
Deep experience in storage engineering within HPC, AI infrastructure, hyperscale, or large-scale data center environments
Deep hands-on expertise with VAST Data storage platforms strongly preferred
Strong understanding of high-performance distributed storage architectures and parallel file systems
Experience supporting GPU-intensive AI/ML workloads and high-throughput data environments
Strong Linux systems administration skills
Experience troubleshooting performance across storage, networking, and compute layers
Familiarity with NFS, RDMA, NVMe-oF, InfiniBand, and modern storage networking concepts
Automation and scripting experience using Python, Bash, or Ansible preferred
Strong understanding of scalability, resiliency, and data protection in enterprise storage environments
Benefits:
Stock options
Remote working options and allowance
Salary:
Circa $200,000 base salary
#J-18808-Ljbffr
Worksite address
workfromhome, CA, 94199, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.