SysOps Engineer — Hardware & Secure Infrastructure
- Role
- DevOps
- Experience
- Senior
- Employment
- Full-time
Open to US only. Set where you work from to check your eligibility.
No BS summary
Senior SysOps engineer with 6+ years operating physical infrastructure and building/debugging servers in datacenter environments. Must know Linux systems administration, networking fundamentals, secure/regulated deployments, IaC/config management, and Python or shell. US-anywhere remote role with regular travel to data centers and client sites.
Core skills
Required skills
Optional skills
The role RAAD is not a cloud-only company. We design and build our own servers from the ground up — GPU processing nodes, storage arrays, and edge appliances — and we deploy them everywhere our clients need them: our own racks, public cloud for burst capacity, and hardened on-prem installations inside secure client environments where data can never leave the building. Energy, defense-adjacent, and critical-infrastructure clients choose us because we can run the entire stack on hardware we control. We're hiring a SysOps engineer with deep hardware experience to help scale this footprint as we add racks, regions, and secure deployments. What you'll do Spec, build, burn-in, and deploy our server platforms — GPU compute, high-density storage, and networking — from component selection to racked-and-serving. Run our colocation and on-prem fleet: provisioning (PXE/Redfish), firmware and BIOS management, out-of-band access, capacity planning, and hardware lifecycle. Stand up secure on-prem installations at client sites: air-gapped and restricted-network deployments, disk encryption and key management, physical and network hardening to client security requirements. Operate the hybrid layer — our own metal for baseline load, cloud for burst — including networking, VPN/WireGuard meshes, and observability across both. Automate everything twice-done: Ansible/Terraform, image building, self-healing where hardware allows and runbooks where it doesn't. Carry real operational ownership, including participating in an on-call rotation for infrastructure that clients bet their projects on. What you'll bring 6+ years operating physical infrastructure: you've built servers with your hands, debugged failed DIMMs and flaky NICs, and know your way around a datacenter cage. Strong Linux systems administration and network fundamentals (VLANs, BGP basics, firewalls). Experience with secure or regulated environments — restricted networks, encryption at rest, compliance-driven hardening. Fluency in infrastructure-as-code and configuration management; comfort writing Python or shell to close gaps. GPU cluster experience (drivers, CUDA stack, thermals, power budgeting) is a strong plus.
What you'll do
- Spec, build, burn-in, and deploy server platforms for GPU compute, high-density storage, and networking
- Run colocation and on-prem fleet operations, including provisioning, firmware and BIOS management, out-of-band access, capacity planning, and hardware lifecycle
- Stand up secure on-prem installations at client sites with air-gapped and restricted-network deployments
- Manage disk encryption, key management, and physical and network hardening to client security requirements
- Operate hybrid infrastructure across owned metal, public cloud burst capacity, networking, VPN/WireGuard meshes, and observability
- Automate repeated work with Ansible, Terraform, image building, self-healing systems, and runbooks
- Participate in an on-call rotation for infrastructure operations
What they require
- 6+ years operating physical infrastructure, including building servers by hand, debugging failed DIMMs and flaky NICs, and working in datacenter cages
- Strong Linux systems administration and network fundamentals
- Experience with secure or regulated environments, restricted networks, encryption at rest, and compliance-driven hardening
- Fluency in infrastructure-as-code and configuration management
- Comfort writing Python or shell to close gaps
- GPU cluster experience with drivers, CUDA stack, thermals, and power budgeting is a strong plus
Benefits
- Strong base salary plus meaningful equity
- Health, dental, and vision coverage for employee and dependents, tailored to country
- Remote-first work from anywhere in the role's region
- Async-friendly, documentation-driven culture
- Gear and home office budget for top-spec hardware and workspace
- Flexible PTO plus local public holidays
- Annual learning budget
- Team offsites
- Field time with operations
RAAD designs and builds its own servers, including GPU processing nodes, storage arrays, and edge appliances, and deploys them across its own racks, public cloud, and hardened on-prem installations for clients in energy, defense-adjacent, and critical infrastructure environments.