Retour à la recherche
Logo de J&M Group
J&M GroupSource d’offres vérifiée

Data Engineer

Offre en anglais
  • Toronto, ON
  • Sur place
  • Publié 8 sept. 2026
  • 1 poste

Ouvre un site externe

Connectez-vous pour enregistrer ce poste
Type d’emploi
Temps plein
Niveau d’expérience
Intermédiaire · 2+ ans
Postuler avant le
7 mars 2027
Langue de l’offre
anglais
Heures de travail
40 heures par semaine
Niveau d’expérience
Entry level
Mode de candidature
La candidature directe est offerte

Résumé du poste

Design, develop, and optimize scalable data pipelines and real-time processing solutions using PySpark, Spark, and Kafka. Collaborate with cross-functional teams to implement ETL/ELT workflows and ensure data quality and governance.

Détails du poste

An experienced Data Engineer with strong expertise in Big Data technologies to design, develop, and support enterprise-scale data platforms. The ideal candidate should possess hands-on experience in PySpark, Apache Spark, Kafka, Hadoop ecosystem components, and Apache NiFi, with a strong understanding of data ingestion, transformation, and real-time processing frameworks. Key Responsibilities Design, develop, and optimize scalable data pipelines using PySpark, Spark, Hadoop, and Apache NiFi. Build and maintain batch and real-time data processing solutions. Develop and support Kafka-based streaming applications and event-driven architectures. Create and optimize ETL/ELT workflows for large-scale structured and unstructured datasets. Develop complex SQL queries for data extraction, transformation, validation, and troubleshooting. Implement data ingestion solutions from databases, APIs, files, and streaming sources. Monitor, troubleshoot, and enhance the performance of Spark jobs and data pipelines. Collaborate with architects, business analysts, and development teams to deliver high-quality data solutions. Support platform upgrades, deployments, testing, certification, and production releases. Ensure data quality, governance, security, and operational excellence across data platforms. Mandatory Skills PySpark Apache Spark (Spark SQL, DataFrames) Apache Kafka Hadoop Ecosystem (HDFS, Hive, YARN) Apache NiFi SQL Python Preferred Skills Spark Streaming Airflow / Oozie Hive Scala Jenkins, Bitbucket, Git JIRA, Confluence Cloud Platforms (GCP/AWS/Azure) Data Warehousing concepts and Dimensional Modeling

Ce que vous ferez

Design, develop, and optimize scalable data pipelines and real-time processing solutions using PySpark, Spark, and Kafka. Collaborate with cross-functional teams to implement ETL/ELT workflows and ensure data quality and governance.

Exigences

Requires strong expertise in Big Data technologies including the Hadoop ecosystem, Apache NiFi, and Python. Candidates should be proficient in SQL and experienced in building event-driven architectures and data ingestion solutions.

Compétences indiquées

  • SQL · Souhaitée
  • Git · Souhaitée
  • Python · Souhaitée

Autres compétences pertinentes

Relevées dans la description du poste. Confirmez les exigences importantes ci-dessus.

  • PySpark
  • Apache Spark
  • Apache Kafka
  • Hadoop
  • Apache NiFi
  • SQL
  • Python
  • Spark Streaming
  • Airflow
  • Oozie
  • Hive
  • Scala
  • Jenkins
  • Bitbucket
  • Git
  • Cloud Platforms

Domaines d’emploi

  • Data & Analytics
  • Technology
  • Software
  • Engineering
  • Consulting

D’autres postes auxquels postuler directement

Des possibilités semblables publiées par des employeurs qui recrutent sur Jobs.ca, sans formulaire externe.

Voir tous les postes à candidature simplifiée