Перейти к основному содержимому
DDN

Senior Staff Engineer - AI Data Path

УдалённоUnited States только
Опубликовано
Роль
Бэкенд
Опыт
Стафф
Занятость
Полная занятость
Зарплата не указанаОценка модели · $185k–$235k/годСредняя уверенность · 80 сопоставимых вакансийМедиана $200k · Основание: роль, уровень, география, требования, тип занятости и формат работы
Проверьте доступность

Доступно для: US only. Укажите, откуда вы работаете, чтобы проверить доступность.

Коротко по делу

Senior Staff-level engineer with 12+ years in storage, distributed systems, or performance engineering. Must have deep distributed storage, Linux I/O, NVMe/SSD, RDMA/InfiniBand, GPU data-path knowledge, and Python and/or C/C++ skills. Focus is AI inference data movement and storage optimization using NVIDIA NIXL/Infinia-style infrastructure.

Ключевые навыки

NVIDIA NIXLDistributed StorageGPU Data Path

Обязательные навыки

LinuxNVMeSSDRDMA/InfiniBandPython/C/C++

Желательные навыки

GPUDirect StorageRAGVector Search

Чем предстоит заниматься

  • Lead the design and implementation of high-performance data movement pipelines using NVIDIA NIXL across GPU, CPU, and storage tiers.
  • Architect and drive integration of DDN Infinia with GPU-accelerated inference platforms for large-scale, real-time AI workloads.
  • Own end-to-end optimization of I/O paths between GPU memory and storage using technologies such as NVIDIA GPUDirect Storage, RDMA, and NVMe-over-Fabrics.
  • Define and implement multi-tier storage architectures optimized for inference latency, throughput, and scalability.
  • Lead development of advanced KV cache management strategies, including offloading, prefetching, and persistence across distributed storage layers.
  • Partner with AI/ML engineering teams to optimize inference performance in frameworks such as PyTorch and TensorFlow.
  • Establish benchmarking frameworks and lead performance tuning efforts for storage and data movement in production inference environments.
  • Diagnose and resolve complex system bottlenecks across storage, networking, and GPU subsystems.
  • Influence architecture decisions for distributed inference systems, ensuring scalability, resilience, and efficient data locality.
  • Drive engineering excellence through best practices in observability, performance monitoring, automation, and reliability engineering.
  • Mentor junior engineers and provide technical leadership across cross-functional teams.
  • Lead the evolution of storage systems into GPU-native data layers for AI inference.
  • Build next-generation distributed AI infrastructure using NIXL and Infinia.
  • Drive performance breakthroughs in real-time LLM inference at scale.
  • Design storage architectures for large-scale AI datasets and retrieval systems.

Что требуется

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
  • 12+ years of experience in storage systems, distributed systems, or performance engineering.
  • Proven track record of architecting and delivering large-scale, high-performance infrastructure systems.
  • Deep expertise in distributed storage architectures such as object storage, scalable file systems, or cloud-native storage platforms.
  • Strong understanding of Linux I/O stack, filesystem internals, and storage protocols.
  • Extensive hands-on experience with NVMe, SSD optimization, and high-performance storage environments.
  • Strong experience with RDMA, InfiniBand, or other high-speed data transfer technologies.
  • Solid understanding of GPU computing concepts and CPU–GPU data movement patterns.
  • Proficiency in Python and/or C/C++, with advanced debugging, profiling, and performance tuning skills.
  • Demonstrated ability to optimize latency-sensitive, high-throughput production systems.
  • Preferred: Strong understanding of AI inference systems, LLM serving architectures, and KV cache optimization.
  • Preferred: Background in high-performance computing (HPC) or hyperscale distributed environments.
  • Preferred: Expertise in caching strategies, memory tiering, and data locality optimization.
  • Preferred: Experience designing disaggregated compute and storage architectures.

DDN

DDN is positioned as NVIDIA’s storage and data intelligence partner for AI factories and the NVIDIA AI Data Platform.

Data Storageddnet.org/

Что говорят о компании

4.0/ 5

Детали

Способ откликаDom
Зарплата не указана