Search by job, company or skills

Senior Data Engineer

  • Posted 2 hours ago
  • Be among the first 10 applicants

Job Description

Job Description: Big Data Platform Engineer

Role Summary

We are seeking a highly motivated and skilled Big Data Platform Engineer to design, develop, and maintain scalable data processing platforms and pipelines. The ideal candidate should have strong expertise in Python, Apache Spark, SQL, and experience working in Linux-based environments with exposure to DevOps practices. The role involves building reliable, high-performance data solutions that support enterprise-scale analytics and business-critical applications.

Key Responsibilities

Data Engineering & Development

  • Design, develop, and maintain scalable data pipelines using Python and Apache Spark.
  • Implement efficient ETL/ELT processes for large-scale structured and unstructured datasets.
  • Develop and optimize complex SQL queries, data models, and transformations.
  • Ensure data quality, integrity, and reliability across the platform.

Platform Operations

  • Work with Linux-based environments for deployment, troubleshooting, and performance tuning.
  • Develop and maintain shell scripts for automation and operational tasks.
  • Monitor and optimize Spark jobs for performance, scalability, and resource utilization.

DevOps & Automation

  • Implement CI/CD pipelines and deployment automation.
  • Participate in infrastructure provisioning, monitoring, and release management activities.
  • Collaborate with DevOps teams to improve platform reliability and operational efficiency.

Collaboration & Governance

  • Work closely with Data Architects, Product Owners, and Business Stakeholders.
  • Participate in code reviews and ensure adherence to engineering best practices.
  • Create and maintain technical documentation and operational runbooks.

Mandatory Skills

Core Big Data Platform Skills

  • Strong programming experience in Python
  • Hands-on expertise with Apache Spark (PySpark preferred)
  • Strong SQL development and query optimization skills
  • Apache Airflow for workflow orchestration and scheduling
  • Kubernetes (AKS/SKE) for container orchestration and deployment
  • MinIO / S3 Compatible Object Storage

Technical Competencies

  • Data Pipeline Development
  • ETL/ELT Processing
  • Data Modelling
  • Performance Tuning
  • Version Control (Git)
  • Agile/Scrum Delivery Model

Good to Have Skills

  • Basic Linux Administration
  • Shell Scripting (Bash/KSH)
  • Understanding of DevOps practices and CI/CD pipelines
  • Docker and Containerization Concepts
  • Cloud Platform Experience (Azure/AWS)
  • Monitoring & Logging Tools

Desired Experience

  • 6 to 8 years of experience or 8 to 12 years or 12+ years of experience in Data Engineering or Big Data Platform Development.
  • Experience working with enterprise-scale data platforms.
  • Experience in distributed computing environments.

Strong analytical and problem-solving skills.

More Info

Job Type:
Industry:
Function:
Employment Type:

About Company

Job ID: 153700477

Similar Jobs

Remote, Bengaluru, India

Skills:

GcpApache SparkDatabricksTableauPythonSqlAWSAirflowBitbucket Pipelinesdbt

Malaysia, Kuala Lumpur

Skills:

JavaData CleaningPower BiScalaTableauData WarehouseSqlELTData LakeData GovernancePythonEtlLooker

Malaysia, Kuala Lumpur

Skills:

Aws LambdaGithubSAPAws RedshiftPower BiAWS AuroraPysparkAWS GlueSqlELTGitAws RdsLinuxPythonEtl

Malaysia, Kuala Lumpur

Skills:

MinIO - S3Apache SparkSqlDevopsLinuxShell scriptingPythonAws S3GKEApache Delta LakeAKSSKEEKSApache Hudi

Kuala Lumpur

Skills:

Kubernetes (SKE/AKS)MinIO - S3Apache AirflowDevopsLinuxShell ScriptingPythonSparkSqlJavaCore Big Data Platform

Beware of Scammers

We don’t charge money for job offers