Director, Engineering – Release Engineering, DevOps & SRE
- Role
- Engineering Management
- Experience
- Lead
- Employment
- Full-time
Open to US only. Set where you work from to check your eligibility.
No BS summary
Hands-on engineering leader for Release Engineering, DevOps/SecDevOps, and SRE. Needs 12+ years in DevOps/Release Engineering/SRE, 5+ years managing engineering teams, and deep CI/CD, IaC, Kubernetes/Docker, DevSecOps, and observability experience. Remote role in North Carolina.
Core skills
Required skills
Optional skills
We're looking for a hands-on engineering leader to build and own the Release Engineering, SecDevOps, and Site Reliability Engineering (SRE) functions for Infinia. This is a foundational role: you'll define how our software is built, secured, released, and kept running at enterprise scale.
If you thrive at the intersection of infrastructure automation, operational excellence, and team building — and you want your work to directly shape the reliability and velocity of a market-leading storage platform — this is your role.
WHAT YOU'LL DO
RELEASE ENGINEERING
- Own the end-to-end release pipeline — build systems, artifact management, versioning, and release gating
- Drive automation of build, test, and packaging workflows to maximize developer velocity
- Define release cadence, branching strategies, and code freeze processes with engineering leadership
DEVOPS AND SECDEVOPS
- Architect and operate scalable CI/CD pipelines and developer toolchains for on-prem and cloud
- Embed security into the SDLC — SAST, DAST, dependency scanning, secrets management
- Champion automation-first DevOps practices from commit to production delivery
SITE RELIABILITY ENGINEERING
- Define SLOs, SLIs, and error budgets; own incident response and post-mortem culture
- Drive observability — logging, metrics, distributed tracing — across the platform
- Influence reliability and operability early, at design and code-review stages
- Lead capacity planning and infrastructure scaling decisions
LEADERSHIP
- Build, mentor, and grow teams across all three disciplines in a global, distributed org
- Communicate roadmap and risks to senior leadership; partner cross-functionally with product, QA, and security
WHAT YOU BRING
- 12+ years in DevOps, Release Engineering, or SRE — with 5+ years managing engineering teams
- Deep experience with CI/CD platforms (GitHub Actions, Jenkins, GitLab CI, Tekton, or similar)
- Infrastructure-as-code fluency: Terraform, Ansible, Pulumi, or equivalent
- Container orchestration expertise: Kubernetes and Docker in production
- Practical DevSecOps background — you've shipped security tooling, not just talked about it
- SRE chops: SLO/SLI design, on-call frameworks, observability stacks (Prometheus, Grafana, OpenTelemetry)
- Strong communicator — you can translate complex technical tradeoffs for exec and non-technical audiences
BONUS POINTS
- Experience with storage, distributed systems, or infrastructure software
- Familiarity with HPC, AI/ML infrastructure, or enterprise data platforms
- Background scaling DevOps in high-growth, globally distributed environments
- Exposure to SOC 2, ISO 27001, or FedRAMP compliance requirements
WHY DDN
- Work on infrastructure that underpins some of the world's most demanding AI and research workloads
- Greenfield opportunity to shape Release Engineering, DevOps, and SRE from the ground up for Infinia
- Collaborative, engineering-first culture that values autonomy, technical depth, and continuous learning
- Competitive compensation, remote-first flexibility, and a team that genuinely enjoys solving hard problems
What you'll do
- Build and own the Release Engineering, SecDevOps, and Site Reliability Engineering functions for Infinia.
- Define how software is built, secured, released, and kept running at enterprise scale.
- Own the end-to-end release pipeline, including build systems, artifact management, versioning, and release gating.
- Drive automation of build, test, and packaging workflows to maximize developer velocity.
- Define release cadence, branching strategies, and code freeze processes with engineering leadership.
- Architect and operate scalable CI/CD pipelines and developer toolchains for on-prem and cloud.
- Embed security into the SDLC, including SAST, DAST, dependency scanning, and secrets management.
- Champion automation-first DevOps practices from commit to production delivery.
- Define SLOs, SLIs, and error budgets.
- Own incident response and post-mortem culture.
- Drive observability, including logging, metrics, and distributed tracing, across the platform.
- Influence reliability and operability early at design and code-review stages.
- Lead capacity planning and infrastructure scaling decisions.
- Build, mentor, and grow teams across Release Engineering, DevOps/SecDevOps, and SRE in a global, distributed organization.
- Communicate roadmap and risks to senior leadership.
- Partner cross-functionally with product, QA, and security.
What they require
- 12+ years in DevOps, Release Engineering, or SRE, with 5+ years managing engineering teams.
- Deep experience with CI/CD platforms such as GitHub Actions, Jenkins, GitLab CI, Tekton, or similar.
- Infrastructure-as-code fluency with Terraform, Ansible, Pulumi, or equivalent.
- Container orchestration expertise with Kubernetes and Docker in production.
- Practical DevSecOps background shipping security tooling.
- SRE experience with SLO/SLI design, on-call frameworks, and observability stacks such as Prometheus, Grafana, and OpenTelemetry.
- Strong communication skills, including ability to translate complex technical tradeoffs for executive and non-technical audiences.
- Preferred: Experience with storage, distributed systems, or infrastructure software.
- Preferred: Familiarity with HPC, AI/ML infrastructure, or enterprise data platforms.
- Preferred: Background scaling DevOps in high-growth, globally distributed environments.
- Preferred: Exposure to SOC 2, ISO 27001, or FedRAMP compliance requirements.
Benefits
- Offers equity.
- Offers bonus.
- Work on infrastructure that underpins demanding AI and research workloads.
- Greenfield opportunity to shape Release Engineering, DevOps, and SRE from the ground up for Infinia.
- Collaborative, engineering-first culture that values autonomy, technical depth, and continuous learning.
- Competitive compensation.
- Remote-first flexibility.
- Team that enjoys solving hard problems.
DDN is positioned as NVIDIA’s storage and data intelligence partner for AI factories and the NVIDIA AI Data Platform.
What people say about this company
4.0/ 5