🌟 We're Hiring: Senior Systems Engineer! 🌟
The Senior Systems Engineer operates, supports and continuously improves infrastructure and cloud services delivered to Ethan customers. The role combines Level 2/3 operational engineering, strong service ownership, evidence-based delivery and an automation-first approach. The role is expected to move beyond reactive support by removing repeat work, improving monitoring, building reusable automation and delivering measurable customer and service outcomes.
Key Responsibilities
- Own and resolve complex incidents and service requests across Windows Server, identity, virtualisation, backup, storage, monitoring and cloud platforms.
- Participate in major incident response, root cause analysis, problem management and permanent corrective actions.
- Plan and implement approved changes, including risk assessment, testing, rollback planning, validation and evidence capture.
- Perform service health, patching, lifecycle, capacity, backup and recovery assurance activities.
- Maintain accurate ticket updates, closure evidence, timesheets, configuration records, runbooks and knowledge articles.
- Support customer onboarding, technical handover and ongoing operational readiness.
- Participate in the after-hours on-call roster and follow escalation procedures when required.
Innovation, Automation and Continuous Improvement
- Identify recurring work that can be automated, standardised or eliminated.
- Design, test, peer review, document and support reusable automation using PowerShell, APIs, Azure Automation, Logic Apps, Git-based workflows or equivalent tools.
- Use infrastructure-as-code and configuration-as-code approaches where appropriate, with version control and controlled deployment.
- Improve monitoring through alert tuning, actionable service health views and automated remediation.
- Identify safe and practical artificial intelligence opportunities for operations, knowledge retrieval, triage and reporting.
- Document measurable benefits, including time saved, tickets avoided, risk reduced, improved recovery or better consistency.
- Deliver or contribute to one to two validated automation, AI, monitoring, reporting, documentation or process improvements per quarter.
Performance Outcomes
- Operate: support agreed service levels, major incidents, problem management and authorised change.
- Prove: maintain current work notes and provide meaningful closure, validation and customer outcome evidence.
- Improve: deliver measurable automation, AI and service improvement contributions.
- Sustain: maintain runbooks, knowledge, onboarding artefacts and accurate configuration records.
Additional Responsibilities
- Achieve agreed personal and team performance outcomes.
- Undertake reasonable duties delegated by the Team Leader or business leader.
- Conduct all activities in accordance with Ethan policies, security requirements, quality standards and customer obligations
Essential Technical Skills:
- Strong troubleshooting capability across Windows Server 2019/2022/2025, Active Directory, Group Policy, DNS, DHCP, certificate services and hybrid identity.
- Hands-on experience with Microsoft Azure, Microsoft Entra ID, Microsoft 365, Azure Arc or comparable hybrid cloud services.
- Experience with VMware vSphere/vCenter and/or Hyper-V. Exposure to Kubernetes or Tanzu is desirable where relevant.
- Practical automation experience using PowerShell and REST APIs. Experience with Git, Azure Automation, Logic Apps, Bicep, Terraform or equivalent is highly regarded.
- Experience with infrastructure monitoring, log analysis, alert tuning, service health reporting and automated remediation.
- Working knowledge of enterprise backup and recovery, disaster recovery, storage, replication and restore testing.
- Understanding of privileged access, least privilege, server hardening, patching and vulnerability remediation.
- Sound practical understanding of incident, problem, change, knowledge and configuration management.
Experience:
- Demonstrated experience in a Level 2/3 infrastructure, systems or cloud engineering role, preferably in a managed services or multi-customer environment.
- A track record of personally owning complex operational issues through to a documented and validated outcome.
- Examples of automation or engineering improvements delivered, including the problem, approach, controls, measurable benefit and ongoing support model.
- Experience delivering planned changes, migrations, upgrades or platform improvements with testing and rollback plans.
- Experience contributing to major incident resolution, root cause analysis and permanent corrective actions.
- Evidence of clear technical documentation, knowledge sharing and disciplined service management records.
- Ability to explain technical risks, options and outcomes to technical and non-technical stakeholders.
- Collaborative approach to peer review, mentoring, knowledge transfer and continuous improvement.
- High attention to detail and respect for security, privacy, safety and company policies.
Certifications:
Relevant certifications are desirable and must be supported by practical capability. Equivalent commercial experience will be considered.
- Microsoft Certified: Azure Administrator Associate or a current role-aligned Microsoft cloud certification.
- Current Microsoft Applied Skills relevant to Windows Server, Azure Arc, automation, identity or security.
- VMware, backup, storage, Kubernetes, Linux, AWS or networking certification aligned to the supported environment.
- ITIL 4 Foundation or demonstrated experience applying service management practices.
- Tertiary qualification in information technology, engineering or a related discipline, or equivalent commercial experience.
Current hands-on capability and evidence of delivery are more important than legacy certification titles.
Eager to advance your career? 🚀 Apply now and be part of our innovative team!