JobsSenior Storage Production Engineer - DGX Cloud
Senior Storage Production Engineer - DGX Cloud
NVIDIASenior Storage Production Engineer - DGX Cloud
NVIDIALocation
remote, Santa Clara, CA
Type
Full-time
Posted
6/15/2026
Compensation
$176,000 - $333,500 per year
Undergraduate with 5+ Years of Experience
Approval 99.2%·Filings 1,781·New hires 873·
👑 Elite Sponsor
·FY 2025Job description
The Production Engineer role at NVIDIA focuses on designing, implementing, and maintaining large-scale storage systems to ensure high efficiency, reliability, and availability. The team works on optimizing storage architectures for AI/ML workloads and automating storage operations. Professionals in this role are expected to have expertise in various storage technologies and practices. The position emphasizes proactive monitoring, automation, and continuous improvement of storage services.
Requirements
- BS degree or equivalent experience in Computer Science, Storage Systems, or a related technical field with 8+ years of practical experience.
- Experience with distributed and high-performance storage solutions, including clustered and parallel file systems.
- Solid understanding of block, file, and object storage technologies.
- Experience with storage networking protocols such as NFS, SMB, iSCSI, S3, Fibre Channel, RDMA, and NVMe over Fabrics.
- Expertise in algorithms, data structures, complexity analysis, software design, and automating maintenance of large-scale Linux-based storage systems.
- Experience in one or more programming languages such as C/C++, Java, Python, Go, NodeJS, and Bash.
- Hands-on experience with infrastructure configuration management tools like Ansible, Chef, Puppet, and Terraform.
Responsibilities
- Design, implement, and support large-scale storage clusters, ensuring scalability, high availability, and data integrity.
- Develop and maintain storage monitoring, logging, and alerting systems to ensure proactive detection and resolution of performance issues.
- Work with AI/ML workloads to improve storage architectures for low-latency access and high-throughput performance.
- Improve the lifecycle of storage services from inception and design to deployment, operation, and continuous optimization.
- Maintain production storage infrastructure by supervising availability, latency, and system health.
- Optimize storage efficiency through compression, deduplication, tiering strategies, and intelligent workload placement.
- Scale storage systems sustainably using AI/ML-driven automation and dynamic data migration techniques.
- Ensure data security and compliance by implementing encryption, access controls, and auditing mechanisms for storage systems.
- Practice sustainable incident response and blameless root cause analysis.
Benefits
- Employees at NVIDIA are often offered comprehensive, day-one benefits—including medical, dental, and vision coverage with HSA support, life and disability insurance, an Employee Assistance Program, and a 401(k) with auto-enrollment. Many roles also have generous time off and holidays, donation matching (up to $10,000), and a wide menu of extras like FSAs, commuter benefits, legal and identity-theft protection, pet insurance, and wellness discounts. Optional programs can include student-loan and home-purchase support, plus family care resources and expert medical services.
Is this posting expired or inaccurate?
