Senior System Admin
Ferguson
Job Description
Senior Systems Administrator:
The Senior Systems Administrator is responsible for supporting, administering, and optimizing enterprise infrastructure across multi-cloud, hybrid, and on-premises environments. This role provides advanced technical administration and escalation-level support for Azure, AWS, GCP, Active Directory, VMware, Windows and Linux systems, enterprise networking, storage, security, monitoring, automation, backup, and disaster recovery. The position requires strong troubleshooting skills, automation experience, and the ability to coordinate across infrastructure, networking, security, and application teams to maintain stable, secure, and resilient technology services.
Key Responsibilities:
- Support multi-cloud and hybrid infrastructure operations across Azure, AWS, and GCP, along with on-premises Active Directory, VMware, Windows/Linux systems, networking, storage, security, and enterprise backup platforms.
- Administer, monitor, and troubleshoot Azure infrastructure across multiple subscriptions and environments, including Virtual Machines, Storage Accounts, App Services, networking, backup, monitoring, security, and lifecycle operations.
- Administer Microsoft Entra ID and Azure RBAC, including App Registrations, Enterprise Applications, Service Principals, managed identities, built-in and custom roles, role assignments, and least-privilege access controls.
- Develop and maintain PowerShell automation for Azure administration, RBAC auditing, resource discovery, reporting, virtual machine operations, disaster recovery, and repeatable infrastructure tasks.
- Administer and troubleshoot Azure networking services, including virtual networks, peering, network security groups, user-defined routes, Application Gateway, Load Balancers, private connectivity, DNS, firewall paths, health probes, and application traffic flows.
- Administer and support Azure Kubernetes Service, including Kubernetes RBAC, Roles, ClusterRoles, RoleBindings, ClusterRoleBindings, namespaces, access controls, and integrations with Azure Key Vault using Secrets Store CSI Driver, SecretProviderClass, and managed or workload identities.
- Lead troubleshooting of complex production issues involving networking, routing, DNS, identity, permissions, compute, storage, virtualization, operating systems, performance, and application connectivity.
- Support Azure governance, monitoring, and security using Management Groups, subscriptions, Azure Policy, Azure Monitor, Log Analytics, Network Watcher, diagnostic settings, and Microsoft Defender for Cloud.
- Design, automate, and support disaster recovery and resiliency solutions, including Azure Backup, cross-region virtual machine recovery, snapshots, storage recovery, and infrastructure reconstruction using PowerShell, ARM templates, and Bicep.
- Administer and troubleshoot on-premises Active Directory, Windows Server, Linux/Ubuntu, VMware, and NetBackup environments in support of hybrid identity, virtualization, patching, backup, recovery, and infrastructure operations.
- Support AWS and GCP cloud environments and Microsoft Defender for Cloud integrations, including AWS IAM, CloudFormation, GCP IAM, service accounts, and Workload Identity Federation.
- Develop and maintain technical documentation, operational procedures, disaster recovery plans, knowledge-base articles, and runbooks while providing escalation-level support across infrastructure teams.
Required Qualifications:
- Advanced experience administering enterprise infrastructure in hybrid and multi-cloud environments.
- Strong working knowledge of Microsoft Azure infrastructure services, identity, networking, monitoring, security, backup, and governance capabilities.
- Hands-on experience with Microsoft Entra ID, Azure RBAC, managed identities, service principals, and least-privilege access models.
- Experience supporting Windows Server, Linux/Ubuntu, Active Directory, VMware, enterprise backup, and disaster recovery technologies.
- Proficiency with PowerShell automation and infrastructure scripting for administration, auditing, reporting, recovery, and repeatable operational tasks.
- Strong troubleshooting skills across identity, networking, compute, storage, virtualization, operating systems, performance, and application connectivity.
- Ability to create clear technical documentation, operational procedures, runbooks, and disaster recovery plans.
- Strong collaboration and communication skills with the ability to coordinate effectively across infrastructure, networking, security, and application teams.