About this role
Posted:
Read in Mongolian ↗
Google Translate opens in a new tab. Automatic translation may contain errors; check important requirements against the original listing.
Source: Himalayas · Worldwide
Screened as open to applicants in Mongolia using published location requirements. This is not direct employer confirmation; working hours and qualification requirements still apply.
Job Title: Data Engineer (PySpark / Scala / Python)
Location: Remote
Job Description:
We are hiring a Data Engineer with strong hands-on experience in PySpark, Scala, and Python. You must have solid expertise in Apache Spark, as it will be the core technology used for building and managing large-scale data processing pipelines.
Experience with cloud platforms like Google Cloud Platform (GCP), Microsoft Azure, or AWS is a plus.
Required Skills:
• Strong hands-on experience with Apache Spark
• Proficient in PySpark
• Experience in Scala and Python
• Knowledge of ETL processes and data pipeline design
• Understanding of distributed data processing
• Familiarity with version control tools like Git
• Basic knowledge of cloud platforms (GCP, AWS, or Azure)
Nice to Have:
• Experience with cloud-native data tools (e.g., Dataproc, Glue, EMR, BigQuery)
• Familiarity with workflow/orchestration tools like Airflow or Cloud Composer
• Experience with CI/CD for data engineering
• Exposure to both structured and unstructured data
Originally posted on Himalayas