Data Engineer
RemoteRemoteINR 800k–1400kmidFull Time
- Posted
- today
- Source
- Himalayas
- Field
- Engineering, Data & Analytics
Skills
GCPAirflowPythonScalaAzureCI/CDSparkAWSETLGit
Description
Job Title: Data Engineer (PySpark / Scala / Python)
Location: Remote
Job Description:
We are hiring a Data Engineer with strong hands-on experience in PySpark, Scala, and Python. You must have solid expertise in Apache Spark, as it will be the core technology used for building and managing large-scale data processing pipelines.
Experience with cloud platforms like Google Cloud Platform (GCP), Microsoft Azure, or AWS is a plus.
Required Skills:
-
Strong hands-on experience with Apache Spark
-
Proficient in PySpark
-
Experience in Scala and Python
-
Knowledge of ETL processes and data pipeline design
-
Understanding of distributed data processing
-
Familiarity with version control tools like Git
-
Basic knowledge of cloud platforms (GCP, AWS, or Azure)
Nice to Have:
-
Experience with cloud-native data tools (e.g., Dataproc, Glue, EMR, BigQuery)
-
Familiarity with workflow/orchestration tools like Airflow or Cloud Composer
-
Experience with CI/CD for data engineering
-
Exposure to both structured and unstructured data
Originally posted on Himalayas
JobMatch aggregates public listings. Always apply through the original posting.