Posted today · be early
Data Engineer
Pontoonglobal
WorldwideremotePosted today
Skill Required
Data-EngineerBig-Data-EngineerETL-DeveloperData-Pipeline-EngineerData-Engineer-JobsData-EngineeringData-Engineering-SpecialistData-Engineering-JobsData-Engineering-PositionsData Engineerdata engineeringAzureETLGCPEngineeringBigQuerybuildingPythonAirflowSparkdesignScalaCI/CDCloudDesign PatternsAWSGitandCICDFulltime
Key highlights
- Remote position
- Proficiency in PySpark required
- Experience in Scala and Python required
- Knowledge of ETL processes and data pipeline design required
Role overview
We are hiring a Data Engineer with strong hands-on experience in PySpark, Scala, and Python. Apache Spark will be the core technology used for building and managing large-scale data processing pipelines.
Responsibilities
- Build and manage large-scale data processing pipelines using PySpark, Scala, and Python
- Use Apache Spark as the core technology for data processing
- Design and implement ETL processes and data pipelines
- Work with distributed data processing
- Utilize version control tools like Git
Requirements
- Strong hands-on experience with Apache Spark
- Proficient in PySpark
- Experience in Scala and Python
- Knowledge of ETL processes and data pipeline design
- Understanding of distributed data processing
- Familiarity with version control tools like Git
- Basic knowledge of cloud platforms (GCP, AWS, or Azure)
Nice to have
- Experience with cloud-native data tools (e.g., Dataproc, Glue, EMR, BigQuery)
- Familiarity with workflow/orchestration tools like Airflow or Cloud Composer
- Experience with CI/CD for data engineering
- Exposure to both structured and unstructured data
Additional details
- Job Title: Data Engineer (PySpark / Scala / Python)
- Location: Remote
- Originally posted on Himalayas