Search by job, company or skills

Site Reliability Engineer

  • Posted 11 hours ago
  • Be among the first 10 applicants

Job Description

Tech Mahindra represents the connected world, offering innovative and customer-centric information technology experiences, enabling Enterprises, Associates, and the Society to Rise. It has 150,000+ professionals working for 1000+ Global Customers (including Fortune 500 companies) in 90 Countries. We're part of the esteemed Mahindra group, headquartered in India. Under a new CEO, Tech Mahindra is committed to a transformative journey with Scale @ Speed as our guiding principle.

Site Reliability Engineer

Qualifications

  • Minimum of 5+ years of experience in a major services organization, with large-scale data or production/project management and oversight experience.
  • 2+ years of experience in supporting solutions preferably in the big data world or managed similar roles

Desired Technical Skills

  • Big Data & eco systems knowledge, distributed storage like MINIO, ADLS using S3, Iceberg tables
  • Linux shell scripting
  • Support the data loading using Spark engine, knowledge on Kubernetes, secrets management
  • Data schedulers like Ariflow etc
  • Maintaining and troubleshooting the Environment with dashboard and other tools like Grafana
  • Manage desired alerts and monitoring dashboards from SRE perspective.
  • Ability to work 24*7 shifts on need basis

a) Relationship Management

  • Ability to manage senior relationships across all the Business and Functional areas
  • Ability to develop cooperative and constructive working relationships
  • Ability to handle complaints, settle disputes and resolve conflicts and negotiate with others
  • Collaborative team player orientation towards work relationships, strong culture awareness

b) Production Support and Decision-making

  • Highly developed skills in managing 24x7 production support comprising of Incident, Problem, Change management and other ITIL phases of production support
  • Ability to break down complex problems and projects into manageable goals
  • Ability to get to the heart of the problem and make sound and timely decisions to resolve problems
  • High severity incident management and situation management skills
  • Knowledge on ADO and SNOW workflow management

c) Development

  • Ability to develop technical skills in coaching, mentoring, and teaching on the job to co-staff
  • Effectiveness in building trust, respect and cooperation among teams

More Info

Job Type:
Industry:
Employment Type:

About Company

Job ID: 152968437

Similar Jobs

Malaysia, Kuala Lumpur

Skills:

CloudformationLoad BalancersWindowsLambdaGrafanaRDSHttpLinuxNode.jsSecurity GroupsECSSqsTerraformS3VpcDnsElkRoute 53Distributed SystemsSesFirewallsCloud NetworkingRubyPrometheusTlsPythonBashEc2CloudwatchElastiCache

Malaysia, Kuala Lumpur

Skills:

JavaGrafanaLoad BalancingArmCloudformationGcpTerraformAzure DevOpsDnsNew RelicElkDatadogAWSPrometheusTcp IpKubernetesPythonBashAzureDockerJenkinsFirewallsGoLinux Unix systemsGitLab CIGitHub ActionsEFK

Beware of Scammers

We don’t charge money for job offers