Search Jobs

Search by job, company or skills

Senior Site Reliability Engineer

Senior Site Reliability Engineer

Mahindra Satyam
  • Posted 10 hours ago
  • Be among the first 10 applicants

Job Description

Tech Mahindra represents the connected world, offering innovative and customer-centric information technology experiences, enabling Enterprises, Associates, and the Society to Rise. It has 150,000+ professionals working for 1000+ Global Customers (including Fortune 500 companies) in 90 Countries. We're part of the esteemed Mahindra group, headquartered in India. Under a new CEO, Tech Mahindra is committed to a transformative journey with Scale @ Speed as our guiding principle.

Job Summary:

We are seeking a highly skilled Site Reliability Engineer (SRE) with 8+years of experience to join our dynamic team in Petaling Jaya. The ideal candidate will possess a deep understanding of SRE principles and practices, ensuring the reliability, availability, and performance of our systems. You will work closely with development and operations teams to implement best practices in system reliability and automation.

Responsibilities:

  • Design, implement, and maintain scalable and reliable systems and services.
  • Monitor system performance and reliability, proactively identifying and resolving issues.
  • Develop and maintain automation tools for deployment, monitoring, and incident response.
  • Collaborate with development teams to ensure that reliability is built into the software development lifecycle.
  • Implement and manage incident response processes, including post mortem analysis.
  • Participate in on call rotations and provide support for production systems.
  • Continuously improve system architecture and operational processes.
  • Document processes, systems, and best practices for knowledge sharing.

Mandatory Skills:

  • Strong knowledge of Site Reliability Engineering (SRE) principles and practices.
  • Proficiency in scripting and programming languages such as Python, Go, or Ruby.
  • Experience with cloud platforms (AWS, Azure, GCP) and container orchestration (Kubernetes, Docker).
  • Solid understanding of networking, security, and system architecture.
  • Experience with monitoring and logging tools (Prometheus, Grafana, ELK stack).
  • Strong problem solving skills and the ability to work under pressure.

More Info

Job Type:
Industry:
Employment Type:

Key Skills

About Company

Similar Jobs

8-10 yrs
Malaysia, Kuala Lumpur
Skills:
AWS, Elk, Hadoop, Kafka, Datadog, Hive, Docker, Terraform, Teamcity, AWS CloudFormation, Spark, Kubernetes, Python, Go, GitHub Actions