Staff Site Reliability Engineer
- Seniority
- Staff Principal
- Work model
- Hybrid
- Location
- New York, New York, United States
- Posted
- 6d ago
<p></p> <p class="p8i6j01 paragraph"><strong>Role Overview</strong><br>You’re a seasoned Site Reliability Engineer who loves owning complex infrastructure, making things run faster, safer, and with less manual effort. In this Staff‑level role, you’ll design and operate VMware‑based private cloud platforms that power mission‑critical SaaS products used by customers around the world. You’ll work across Linux, Windows Server, networking, storage, and automation frameworks to increase reliability, reduce toil, and modernize a global datacenter environment. You’ll have the scope to set technical direction, build automation at scale, and mentor engineers while staying hands‑on with VMware vSphere, F5/AVI load balancers, and hybrid Active Directory.</p> <p class="p8i6j01 paragraph"><strong>Here’s a breakdown of what you’ll do (not all of it, just the important stuff)</strong></p> <ul class="p8i6j07 p8i6j02"> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Lead the architecture, deployment, and ongoing optimization of VMware vSphere–based private cloud infrastructure across multiple global datacenters.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Design and build automation using PowerShell/PowerCLI, Ansible, Python, and CI/CD tools to streamline provisioning, configuration, and compliance.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Administer, harden, and troubleshoot Linux (RHEL/CentOS/Ubuntu) and Windows Server environments that host enterprise and SaaS workloads.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Integrate and manage Active Directory for authentication, access control, and service accounts across hybrid on‑prem and cloud environments.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Partner with network and security teams to manage firewalls, VPNs, storage, and load balancers (F5 BIG‑IP, AVI/NSX Advanced Load Balancer) for highly available services.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Document architectures and runbooks, participate in on‑call and change management, and mentor engineers while influencing long‑term reliability and automation strategy.</p> </li> </ul> <p class="p8i6j01 paragraph"><strong>These are the essentials you’ll need to get an interview</strong></p> <ul class="p8i6j07 p8i6j02"> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">10+ years of experience in systems or infrastructure engineering, including operating large‑scale enterprise or SaaS datacenter environments.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Deep hands‑on expertise with VMware vSphere (ESXi, vCenter, DRS, HA, vMotion, distributed switches) in production.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Strong Linux administration skills (RHEL/CentOS/Ubuntu), including performance tuning, system hardening, and advanced troubleshooting.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Solid experience with Windows Server and Active Directory (Group Policy, DNS, authentication and access integrations).</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Proven track record building and maintaining automation using PowerShell/PowerCLI, Ansible, Python, or similar tools, plus familiarity with Git or other version control.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Good understanding of storage (SAN/NAS), TCP/IP networking, DNS, VPNs, firewalls, and production monitoring/alerting.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">A collaborative, problem‑solving mindset with the ability to lead complex incidents, communicate clearly, and operate in an on‑call, high‑availability environment.</p> </li> </ul> <p class="p8i6j01 paragraph"><strong>It would be great if you had these too, but we’ll support you if you don’t</strong></p> <ul class="p8i6j07 p8i6j02"> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Experience with enterprise storage and compute platforms such as Pure Storage or Cisco UCS.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Familiarity with Terraform, Jenkins, Azure DevOps, or similar tools for infrastructure as code and CI/CD automation.</p> </li> <li class="p8i6j0a"> <p class="p8i6j01 paragraph">Exposure to security hardening and compliance frameworks such as CIS benchmarks, NIST, or ISO 27001.</p> </li> </ul> <p></p><div class="content-pay-transparency"><div class="pay-input"><div class="title">U.S pay range </div><div class="pay-range"><span>$131,000</span><span class="divider">—</span><span>$164,000 USD</span></div></div></div><div class="content-conclusion"><p> </p> <p><strong>About Us</strong></p> <p><span data-olk-copy-source="MessageBody">Diligent is the AI leader in governance, risk and compliance (GRC) SaaS solutions, helping more than 1 million users and 700,000 board members to clarify risk and elevate governance. The Diligent One Platform gives practitioners, the C-Suite and the board a consolidated view of their entire GRC practice so they can more effectively manage risk, build greater resilience and make better decisions, faster. </span></p> <div><span data-teams="true">At Diligent, we're building the future with people who think boldly and move fast. Whether you're designing systems that leverage large language models or part of a team reimaging workflows with AI, you'll help us unlock entirely new ways of working and thinking. Curiosity is in our DNA, we look for individuals willing to ask the big questions and experiment fearlessly - those who embrace change not as a challenge, but as an opportunity. The future belongs to those who keep learning, and we are building it together. At Diligent, you’re not just building the future - you’re an agent of positive change, joining a global community on a mission to make an impact.</span></div> <p>Learn more at diligent.com or follow us on <u><a id="OWA170dad44-6a69-e1e1-d720-601cd59ec77c" class="x_x_x_OWAAut