Live opening · Posted 22 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
The candidate will have responsibilities across the following functions:
Hybrid and Colocation Infrastructure:
Support and maintain infrastructure across colocation facilities and cloud environments.
Strong hands-on experience in Windows Server administration, including patching, troubleshooting, performance tuning, and lifecycle management.
Administer VMware vSphere clusters, ESXi hosts, and vCenter, including VM provisioning, resource management, and day-to-day operational support.
Manage enterprise storage systems including PURE Storage arrays: provisioning, performance monitoring, and snapshot lifecycle management.
Perform hardware and software troubleshooting across physical and virtualized environments.
Cloud and Platform Operations:
Administer and support workloads in Azure and Google Cloud Platform.
Manage cloud resources including virtual machines, storage, networking, and platform services.
Perform system upgrades, patching, and backup operations across cloud and on-premises environments.
Ensure compliance with internal IT policies and relevant industry standards.
Support identity and access management in Azure Active Directory and related cloud services.
Patch Management and Security Vulnerability Response:
Execute and maintain patch management processes across Windows server environments using Ansible, PatchMyPC, WSUS, or equivalent tooling.
Manage patching cadences for production, test, and ancillary environments, coordinating maintenance windows and pre/post-patch validation.
Respond to security vulnerability findings from tools such as CrowdStrike or Qualys, triaging, scheduling, and remediating identified exposures.
Maintain exclusion and deferral tracking for systems that cannot be patched on standard cadence; escalate unresolved exposures per InfoSec policy.
Support out-of-band and zero-day patch response processes as directed by security leadership.
Automation and Configuration Management:
Build and maintain automation workflows using Ansible, PowerShell, Python, or Bash to reduce manual operational effort.
Support Infrastructure as Code practices using Terraform or ARM templates for repeatable environment provisioning.
Contribute to GitLab CI/CD pipelines for infrastructure automation and deployment tasks.
Assist with configuration and secret management using HashiCorp Vault or equivalent tooling.
Identify repetitive operational tasks and implement scripted or automated solutions.
Monitoring, Observability and Incident Response:
Monitor system performance, availability, and capacity across cloud and on- premises environments using Datadog, SigNoz, or equivalent platforms.
Configure and maintain observability dashboards and alerting pipelines using OpenTelemetry, Datadog, or SigNoz.
Respond to alerts and incidents in a timely manner; perform root cause analysis and implement corrective actions.
Identify trends and proactively address performance or capacity issues before they impact operations.
Support incident management processes and contribute to post-incident reviews and documentation.
Security, Compliance and Governance:
Collaborate with the security team to maintain system integrity, access controls, and compliance posture.
Support governance and audit requirements including GDPR, ISO 27001 and internal controls.
Assist with identity and access management tasks including Okta and Azure AD administration.
Contribute to security hardening activities and compliance remediation efforts as directed.
Disaster Recovery and Resilience:
Support and participate in disaster recovery testing and business continuity procedures.
Maintain snapshot-based recovery configurations and cross-site replication for critical workloads.
Follow and contribute to documented DR runbooks; identify gaps and suggest improvements.
Collaboration and Documentation:
Work effectively with development, cloud engineering, IT support, and operations teams to support system integration and service delivery.
Create and maintain clear operational documentation, runbooks, and procedural guides.
Communicate clearly on system status, incident updates, and project progress across teams.
Participate in team standups, planning, and retrospectives, contributing operational insights to improve team processes.
Requirements:
Bachelor's degree in Engineering, Computer Science, or a related field, or equivalent practical experience.
5-7 years of experience in IT systems administration, cloud operations, or data center engineering.
Strong hands-on experience in Windows Server administration, including patching, troubleshooting, performance tuning, and lifecycle management.
Hands-on experience with VMware vSphere, ESXi, and vCenter VM administration, resource pools, and basic cluster operations.
Working knowledge of Azure administration of virtual machines, networking, storage, Azure AD, and core platform services.
Experience with patch management tooling such as Ansible, SCCM, or WSUS in a production environment.
Proficiency in at least one scripting language: PowerShell, Python, or Bash.
Familiarity with monitoring platforms such as Datadog, SigNoz, SolarWinds, or equivalent.
Identify opportunities to reduce manual effort by applying automation and AI-driven solutions to operational workflows.
Understanding of security vulnerability management scanning, triage, and remediation workflows.
Solid troubleshooting and problem-solving skills across infrastructure, networking, and cloud environments.
Strong written and verbal communication skills; ability to produce clear documentation and runbooks.
Experience with Azure or Google Cloud Platform (GCP) Compute Engine, Cloud Storage, networking, and IAM.
Strong hands-on experience in Windows Server administration, including patching, troubleshooting, performance tuning, and lifecycle management.
Familiarity with Infrastructure as Code tools such as Terraform or ARM templates.
Experience with Ansible for configuration management and patching automation.
Exposure to observability tooling such as OpenTelemetry, Datadog, or SigNoz.
Familiarity with enterprise storage platforms such as PURE Storage, including snapshot and replication management.
Experience with HashiCorp Vault or equivalent secret management solutions.
Exposure to GitLab CI/CD for infrastructure automation workflows.
Familiarity with Okta or Azure Active Directory for identity and access management.
Experience with low-code automation tools such as Power Automate or Tines.
Relevant certifications: AZ-104 (Azure Administrator), VMware VCP, Google Cloud Associate, or Ansible equivalent are a plus.
Experience
5-9 yrs
More openings worth a look
Recently tracked roles with full details and direct application links.