SRE Automation Engineer (Site Reliability Engineering)
- Posted 9 hours ago
- Be among the first 10 applicants
Job Description
About The Role
We are looking for an SRE Automation Engineer to join HAVI's Supply Chain Technology team.
You will be responsible for building automation solutions that improve system reliability, reduce operational toil and increase platform resilience. The role focuses heavily on automation, self-healing capabilities, incident response and improving deployment reliability across enterprise platforms.
You will work closely with SRE, Platform Engineering and Application Development teams in a global environment.
What You'll Do
You'll be joining a global SRE environment where your work will directly contribute to automation, resilience and reliability engineering across enterprise technology platforms.
We are looking for an SRE Automation Engineer to join HAVI's Supply Chain Technology team.
You will be responsible for building automation solutions that improve system reliability, reduce operational toil and increase platform resilience. The role focuses heavily on automation, self-healing capabilities, incident response and improving deployment reliability across enterprise platforms.
You will work closely with SRE, Platform Engineering and Application Development teams in a global environment.
What You'll Do
- Design and implement automation to eliminate repetitive manual operational tasks
- Build and maintain auto-remediation and self-healing mechanisms
- Develop automation for incident response and runbook execution
- Improve CI/CD pipeline reliability and deployment safeguards
- Integrate observability tooling programmatically to improve detection and recovery
- Standardise automation frameworks across platforms
- Reduce alert noise through automated correlation and response
- Work with SRE and Platform Engineering teams to embed resilience and reliability patterns into system design
- 3+ years of experience in Automation Engineering, DevOps or SRE
- Strong programming/scripting experience with Python, Go, Bash or PowerShell
- Experience building automation frameworks and operational tooling
- Knowledge of Infrastructure as Code, such as Terraform, ARM or CloudFormation
- Experience with CI/CD and deployment automation
- Understanding of observability tools and telemetry integration
- Experience designing auto-remediation workflows
- Cloud platform experience, with Azure preferred
- Strong troubleshooting, analytical and systems-thinking skills
You'll be joining a global SRE environment where your work will directly contribute to automation, resilience and reliability engineering across enterprise technology platforms.
More Info
Key Skills
Operational tooling
Auto-remediation workflows
Telemetry integration
Observability tools
CI/CD
