Software Engineering SMTS (Storage & Data Protection)

Airkit
Airkit

Software Engineering

Hyderabad, Telangana, India

Posted on Jul 21, 2026

Description

Storage & Data Protection Engineer - Hyderabad

About the Role

The Salesforce Technology Services business unit within the Information Technology organization is responsible for deploying new capabilities within infrastructure and enterprise applications to 50,000+ Salesforce employees and business partners. Our infrastructure includes corporate offices, co-lo data centers, and public clouds (AWS for now), while our enterprise applications include SaaS, containerized, virtualized and bare metal deployments.

We are hiring a Storage & Data Protection Engineer to join our global IT Infrastructure team based in Hyderabad. This is a new position created to strengthen our APAC-region operational coverage and support the team's growing automation and cloud transformation agenda.

This is a hands-on operations role at its core — you will own day-to-day storage and backup responsibilities for enterprise platforms in a production environment. At the same time, you will be expected to grow into automation tooling, contribute to AI-assisted operations initiatives, and help modernize how the team delivers infrastructure services.

If you are a solid infrastructure engineer who wants to operate at the level of a reliable, skilled practitioner and gradually build automation and cloud skills in a team that is actively investing in those areas — this role is for you.
What You Will Do

Storage Operations — Day to Day

  • Provision, manage, and decommission storage resources on Dell EMC PowerMax including LUN creation, host mapping, masking views, and capacity management
  • Administer NetApp AFF arrays — create and manage NFS and SMB volumes and shares, manage qtrees and permissions, and support capacity expansion requests
  • Support PowerScale (Isilon) OneFS operations — NFS/SMB share provisioning, access zone management, and basic capacity monitoring
  • Manage Cisco MDS SAN fabric — support zone creation and modification, port allocation, and basic fabric troubleshooting under senior guidance
  • Handle day-to-day storage provisioning tickets and service requests within defined SLA windows, ensuring requests are completed accurately and documented
  • Monitor storage health using platform dashboards (UNISPHERE, NetApp System Manager, BlueXP) — flag anomalies, capacity thresholds, and performance alerts to the team
  • Participate in storage lifecycle activities — firmware updates, controller replacements, and technology refresh projects alongside senior engineers
Data Protection Operations — Day to Day
  • Administer enterprise backup platforms — Cohesity (physical clusters and Helios SaaS), Rubrik, and Veritas NetBackup — including policy creation, schedule management, and backup monitoring
  • Perform and validate restore operations — file-level, VM-level, and application-consistent restores ensuring recovery meets defined RTO and RPO targets
  • Monitor daily backup job health — investigate and resolve failed or missed jobs within defined SLA windows, escalating complex issues to senior team members
  • Manage SnapMirror replication relationships on NetApp and SRDF replication status on PowerMax — monitor lag, investigate alerts, and escalate replication anomalies
  • Support Iron Mountain DRMS tape management operations — monthly/quarterly data restoration integrity checks, vault requests, and inventory reconciliation
  • Contribute to DR test execution — follow runbooks, document test outcomes, and flag discrepancies for remediation
  • Manage vendor support cases with Cohesity, Rubrik, DellEMC, and NetApp — gathering diagnostic data, tracking case progress, and coordinating resolution
Automation & AI Operations — Growth Area
  • Learn and apply the team's automation tooling — Python scripting, Ansible playbooks, and REST API interactions with storage platforms (Dell EMC, NetApp, Cohesity APIs)
  • Contribute to GitOps workflows — raise pull requests for configuration changes, follow peer review processes, and use GitHub Actions pipelines for validated deployments
  • Support the team's Storage Operations AI Incident Agent pilot — help test automated remediation workflows, validate outputs, and contribute to runbook documentation used by AI-assisted triage
  • Write and maintain operational runbooks that form the knowledge base for both human engineers and AI-assisted operations
  • Gradually take on Terraform and Ansible tasks under lead engineer guidance, contributing to infrastructure-as-code maturity for storage provisioning workflows
  • Use AI-assisted coding tools (Cursor AI, GitHub Copilot) to accelerate script writing, documentation generation, and automation task delivery
Cloud & Hybrid Infrastructure Support
  • Help administer Cohesity Cloud Edition on AWS — monitor backup jobs, support restore requests, and assist with IAM configuration under senior guidance
  • Support VMware vSphere storage tasks — datastore management, VMDK operations, and storage vMotion assistance during maintenance activities
  • Develop awareness of cloud-native storage services across AWS, Azure, and GCP as part of the team's ongoing hybrid cloud expansion
Operational Excellence & Collaboration
  • Participate in the on-call rotation — provide first-response for storage and data protection incidents following documented runbooks
  • Document all changes, incidents, and resolutions accurately in the ticketing system — maintain high-quality records that support audit compliance and team knowledge sharing
  • Collaborate with US-based senior engineers, SRE, and platform teams on cross-regional incidents, projects, and knowledge transfer sessions
  • Attend and contribute to team knowledge sessions — learning from senior engineers across PowerMax, PowerScale, Cohesity, and automation topics
  • Proactively flag recurring issues, capacity risks, or process inefficiencies to the team — contributing to continuous improvement even in a junior capacity
Required
  • Experience: 3–6 years of hands-on experience in an enterprise storage or infrastructure engineering role
  • Block Storage: Hands-on experience with at least one block storage platform — Dell EMC PowerMax or equivalent (NetApp, HP, Hitachi) — including LUN provisioning, host connectivity, performance, capacity, and basic troubleshooting
  • Storage Networking: Hands-on experience with storage networking switches — Cisco MDS, Cisco/Brocade NetApp Cluster Interconnect — including zoning, port allocation, performance, and basic troubleshooting
  • File Storage: Hands-on experience with NetApp ONTAP (NFS/SMB/pNFS) — volume and share management, SnapMirror, performance, capacity, and basic troubleshooting
  • Data Protection: Hands-on experience with enterprise backup platforms — Cohesity, Rubrik, Veritas NetBackup, or Commvault — including policy management, configuration, job monitoring, and restore execution
  • Scripting: Intermediate scripting ability in Python and Bash for automating repetitive tasks, parsing logs, or generating reports
  • IaC: Exposure to Ansible or Terraform — even if limited to running existing playbooks or modules
  • Process: Experience working from runbooks and standard operating procedures in a production environment
  • Troubleshooting: Strong troubleshooting mindset — methodically diagnose storage and backup issues, gather diagnostic data, and escalate appropriately
  • Communication: Good written and spoken English — ability to communicate clearly in a global, distributed team environment
Desired Skills
  • Demonstrable experience with cloud storage — AWS S3, EBS, EFS, FSx, Azure Blob — even on a small scale or lab environment
  • Experience with PowerScale (Isilon) OneFS administration
  • Familiarity with Git and GitHub — committing changes, raising pull requests, working in a version-controlled environment
  • Experience with VMware vSphere storage — datastores, VMDK management, or storage vMotion
  • Active or completed certifications: AWS Cloud Practitioner, NetApp NCSA, Dell EMC Proven Professional, or Cohesity/Rubrik certification
  • Experience using PagerDuty or equivalent on-call and alerting platforms
  • Exposure or theoretical knowledge of AI tools including Cursor and Claude CLI
Why Join This Team

Real production ownership from day oneYou will not be running tests in a sandbox. From month one you will be working on live production infrastructure that matters, with proper mentoring and guardrails to help you succeed.
A team that invests in your growthStructured development plans, certification support, peer knowledge sessions, and a team that genuinely cares about moving engineers forward — not just filling seats.
Work on the future of Infra OpsThe team is actively building AI-assisted incident response, IaC-driven storage automation, and cloud-native data protection. You will contribute to this work as your skills grow.Global team, local presenceYou will work closely with experienced engineers across the US, APAC, and EMEA — getting exposure to global infrastructure and cross-regional collaboration from a Hyderabad/Bangalore base.