Skip to main content
ResumeKart
← Back to Jobs

Data Engineer – Kafka / PySpark / Hadoop

AcestackRegent Park, City of Toronto
Full-time3-7
👁️ 11 views📝 0 applicationsPosted 9/15/2026Expires 10/16/2026
Tailor Resume for This JobCheck ATS Score

Get alerts for roles like this

More Data Engineer – Kafka / PySpark / Hadoop roles in Regent Park, City of Toronto — straight to your inbox. No account needed.

Applying to this role? Tailor your résumé to this job description in one click, then download it clean — no watermark, no subscription.

Job Description

• Job Title: Data Engineer Kafka / PySpark / Hadoop • Location: Toronto, ON • Work Model: Onsite • Job Type: Full Time (FTE) Job Description We are seeking an experienced Data Engineer with strong hands-on expertise in Kafka, PySpark, Python, and Hadoop to design, develop, and support scalable batch and real-time data pipelines. The ideal candidate will have strong experience working with large-scale distributed data processing environments and enterprise data integration solutions. Key Responsibilities • Design, develop, and maintain scalable batch and real-time data pipelines . • Develop data processing applications using Python and PySpark/Apache Spark . • Build and support Kafka-based data ingestion and streaming pipelines . • Work with Hadoop and related technologies to process large volumes of data. • Develop and maintain ETL/ELT pipelines for data ingestion, transformation, cleansing, and integration. • Perform data validation, reconciliation, and quality checks. • Troubleshoot pipeline failures, data discrepancies, and performance issues. • Optimize Spark/PySpark jobs and SQL queries for performance and scalability. • Monitor data pipelines and resolve production issues. • Collaborate with data architects, developers, analysts, and business teams. • Participate in Agile development, testing, deployment, and production support activities. Required Skills • Strong hands-on experience with Python for data engineering and automation. • Strong expertise in PySpark / Apache Spark . • Hands-on experience with Apache Kafka for real-time data ingestion and streaming. • Strong experience with the Hadoop ecosystem and distributed data processing. • Strong SQL skills and experience working with large datasets. • Experience developing and maintaining ETL/ELT data pipelines . • Strong understanding of distributed computing and data processing concepts. • Experience with data ingestion, transformation, cleansing, and integration. • Strong troubleshooting and performance optimization skills. Good to Have • Hive • Databricks • AWS, Azure, or GCP • Git and CI/CD • Unix/Linux • Airflow or Autosys • Relational and NoSQL databases

Required Skills

PythonPySparkApache KafkaHadoopSQLETLELTdata ingestiondata transformationdata cleansingdata integrationtroubleshootingperformance optimizationdistributed computingdata processingHiveDatabricksAWSAzureGCPGitCI/CDUnixLinuxAirflowAutosysrelational databasesNoSQL databases

Partner picks for Data Engineer – Kafka / PySpark / Hadoop in Regent Park

Matched to the skills this page calls for and the candidate's location.

Partner
  • UdemyVerified partner
    Partner course provider

    A marketplace of instructor-created courses across technology, business and creative skills.

    covers pythoncovers awscovers sql
  • Partner course provider

    University and industry courses, professional certificates and online degrees.

    covers pythoncovers sql

Partners are ResumeKart affiliates or institutes it works with; ResumeKart may earn a commission when a candidate enrols. Placement is decided by relevance, not payment. How ResumeKart earns

The best-paying roles in your field. Every week. Free.

Join 10,000+ professionals getting job alerts and salary insights in their inbox

We respect your privacy. Unsubscribe anytime with one click.