The Site Reliability Engineer (Azure) will be responsible for ensuring the reliability, availability, and performance of cloud-based systems in the Transport & Distribution industry. This role involves implementing best practices, automating processes, and resolving system issues to maintain high service standards in a fast-paced environment.
Client Details
Our client haven't just been in the business of supply chain; they have been in the business of re-imagining connections between people and the products they love to create a better future
Description
- Build CI/CD pipelines by introducing automation, reliability controls, and deployment safeguards. Integrate and optimize observability and monitoring tools to strengthen system visibility, detection, and recovery capabilities.
- Build and maintain self-healing and auto-remediation capabilities to minimize manual intervention and accelerate issue resolution.
- Design, develop, and implement automation solutions to improve system reliability, operational efficiency, and platform resilience.
Profile
- Strong proficiency in programming and scripting languages such as Python, Go, Bash, or PowerShell.
- Experience with CI/CD platforms, deployment automation, and cloud-native environments.
- Understanding of auto-remediation, self-healing architectures, and reliability engineering principles.
- Strong analytical, troubleshooting, problem-solving, and systems-thinking capabilities.
Job Offer
- The People! The Work! The Environment
- Competitive Salary: we offer a salary that reflects your experience, and expertise
- Work-Life Balance: Hybrid Working - 3 days in office, 2 days from home
- Professional Growth: access to ongoing training
- Inclusive Environment: we prioritize diversity and inclusion, ensuring everyone feels welcome and valued.
To apply online please click the Apply button below. For a confidential discussion about this role please contact Arshanaa Lechumi on +60323024087.