Skip to main content
Astreya

IT Infrastructure Operations Engineer II

RemoteIndia only
Published
Role
DevOps
Experience
Senior
Employment
Full-time
Salary not disclosed
Check eligibility

Open to IN only. Set where you work from to check your eligibility.

No BS summary

Mid-level (L2) IT infrastructure operations engineer with 5+ years' enterprise experience. Must have hands-on Dell PowerEdge (iDRAC/Redfish) and Cisco (IOS/NX-OS) skills, plus monitoring/dashboarding and on-call shift availability. Remote, India location requirement.

Core skills

Dell PowerEdgeCisco IOS/NX-OS

Required skills

iDRACRedfishDell OpenManageRAID/PERCBIOS/firmware updatesCisco routersCisco switchesIOSNX-OSCisco CLImonitoring and logging toolsAnsiblePythonCMDB systemsPXEWindows ServerLinuxDell Lifecycle ControllerOpenManagenetwork monitoringfirmware lifecycle managementfirmware/BIOS/driver updateshardware break/fixCMDB

Optional skills

Dell Server certificationsITIL Foundationconfiguration backup toolsdocumentation/diagramming platformsiDRAC Virtual Mediabandwidth/latency testing utilities

What you'll do

  • Provide advanced troubleshooting and fault isolation for escalated server and network incidents using iDRAC, Redfish, and Cisco CLI tools.
  • Execute firmware, BIOS, and driver updates on Dell PowerEdge servers following standardized procedures.
  • Perform IOS/NX-OS firmware and software updates on Cisco routers and switches, adhering to change management protocols and conducting post-update validation.
  • Manage hardware break/fix procedures for server infrastructure, coordinating with Dell support for warranty claims, parts ordering, and scheduling on-site technician dispatch.
  • Conduct regular network health audits and performance analysis, identify bottlenecks and recommend optimization measures.
  • Collaborate with the SRE team to enhance monitoring dashboards and refine alerting thresholds.
  • Mentor and provide technical guidance to L1 engineers, conduct knowledge transfer sessions, and assist with complex ticket resolution.
  • Participate in blameless post-mortems following major incidents and contribute to root cause analysis and preventative actions.
  • Maintain and update operational runbooks, network diagrams, and technical documentation.
  • Support hardware lifecycle management activities including equipment provisioning, asset tracking, and vendor coordination.
  • Provide 24x7 on-call support for critical escalations.
  • Collaborate with the FTE IT Team Lead on capacity planning and provide data-driven insights on infrastructure utilization.
  • Provide advanced troubleshooting and fault isolation for escalated server and network incidents using iDRAC, Redfish, and Cisco CLI.
  • Perform IOS/NX-OS firmware and software updates on Cisco routers and switches, adhere to change management and conduct post-update validation.
  • Manage hardware break/fix procedures for server infrastructure, coordinate with Dell support for warranty claims, parts ordering, and on-site technician dispatch.
  • Conduct regular network health audits and performance analysis; identify bottlenecks and recommend optimizations.
  • Participate in blameless post-mortems after major incidents; contribute to root cause analysis and implement preventative actions.
  • Support hardware lifecycle management including provisioning, asset tracking, and vendor coordination for returns/repairs.

What they require

  • 5+ years of hands-on experience in enterprise IT infrastructure operations.
  • Strong proficiency with Dell PowerEdge server administration, including hardware troubleshooting, iDRAC/Redfish management, and firmware lifecycle management.
  • Solid experience with Cisco networking equipment (routers, switches), including IOS/NX-OS configuration, troubleshooting, and upgrade procedures.
  • Working knowledge of monitoring and logging tools, with ability to create dashboards, configure alerts, and analyze performance metrics.
  • Excellent problem-solving abilities with experience in incident management, root cause analysis, and implementing corrective actions in production environments.
  • Ability to work rotating shifts in a 24x7 global support model.
  • Industry certifications such as Dell Server certifications or ITIL Foundation.
  • Excellent problem-solving abilities with demonstrated experience in incident management, root cause analysis, and implementing corrective actions in production environments.

Astreya is a global IT managed services provider dedicated to creating technology solutions that are reliable and human-centered. We partner with the world’s most innovative organizations to optimize modern IT environments across the digital workplace, cloud, data, AI, and enterprise platforms. Our people are our strength. Across regions and roles, Astreyans bring deep expertise, curiosity, and a commitment to excellence. We believe great technology starts with great humans, which is why we invest in learning, collaboration, and career growth. At Astreya, transparency, accountability, and trust guide how we work with each other and our customers. We foster an inclusive culture where diverse perspectives are valued and everyone has the opportunity to make an impact. Join us to build solutions that matter and shape the future of IT together. Onwards and upwards!

IT ServicesEnterprise
Salary not disclosed