Algotale-Data Engineer(Lumiq)

Mumbai, IN Until 8/22/2026 3+ years exp First posted April 20, 2026 Last posted April 20, 2026
Job description

Data Engineer – AWS/GCP | SQL | PySpark

šŸ“ Location: Mumbai
šŸ¢ Company: Algotale
šŸ§‘ā€šŸ’» Experience: 3+ Years

About Algotale

Algotale is a technology-driven organization focused on building scalable data solutions, cloud-native architectures, and advanced analytics platforms. We help businesses transform data into actionable insights through modern data engineering practices.

Role Overview

We are looking for a skilled Data Engineer with strong expertise in AWS/GCP, SQL, and PySpark to design, build, and optimize scalable data pipelines and cloud-based data platforms. The ideal candidate should have hands-on experience in distributed data processing and cloud environments.

Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT pipelines

  • Build and optimize data architectures on AWS and/or GCP

  • Develop data processing frameworks using PySpark

  • Write complex and optimized SQL queries for large datasets

  • Work with structured and unstructured data from multiple sources

  • Implement data quality, validation, and governance frameworks

  • Collaborate with data analysts, data scientists, and cross-functional teams

  • Monitor and troubleshoot production data systems

  • Ensure best practices in performance tuning and cost optimization in cloud environments

Required Skills & Qualifications

  • 3+ years of experience in Data Engineering

  • Strong hands-on experience with AWS (S3, Redshift, Glue, EMR, Lambda) or GCP (BigQuery, Dataflow, Cloud Storage, Dataproc)

  • Strong proficiency in SQL

  • Hands-on experience with PySpark

  • Experience with data warehousing concepts and dimensional modeling

  • Familiarity with workflow orchestration tools (e.g., Airflow)

  • Understanding of distributed data processing systems

  • Experience with version control tools like Git

Good to Have

  • Experience with real-time data processing (Kafka or similar tools)

  • Knowledge of CI/CD pipelines

  • Exposure to Infrastructure as Code (Terraform, CloudFormation)

  • Basic understanding of machine learning pipelines

About this role

Summary

Design, develop, and optimize scalable data pipelines and cloud data platforms.

Job title

Data Engineer

Experience level

3+ years

Minimum experience

3+ years exp

Industry

software

Location requirements

Mumbai, IN, remote not allowed

Salary

Not specified

Management role

No

Skills & keywords

Required skills

AWS (S3, Redshift, Glue, EMR, Lambda)GCP (BigQuery, Dataflow, Cloud Storage, Dataproc)SQLPySparkdata warehousingworkflow orchestrationGit

Preferred skills

real-time data processingCICDTerraformmachine learning pipelines

Specializations

AWSGCPSQLPySpark
Locations

Structured locations inferred from the posting.

Mumbai, Maharashtra, India

On-site City