Back to job search
T
TekRekVerified Job Source

Staff Data Platform Engineer

Own the long-term buildout of the data platform, driving architectural direction for real-time ingestion and analytics. Design and optimize scalable storage and streaming pipelines for petabyte-scale IoT and time-series data.

  • On-site
  • Toronto, ON
  • Posted Jul 30, 2026
  • Apply by Aug 29, 2026
  • 1 position

Job summary

About the Company TekRek is partnered with a high-growth product company building the data platform behind a modern telematics API, supporting real-time ingestion, storage, enrichment, and analytics for high-volume IoT and time-series workloads at petabyte scale. The Role This is not a traditional data engineering seat; you will own the long-term buildout of the data platform powering the API and analytics layer, drive architectural direction, and partner with product, engineering, and leadership to deliver scalable abstractions and reusable components. What You’ll Do Own projects enhancing data replication, storage, enrichment, and reporting capabilities Build and optimize streaming and batch pipelines that support the core product and API Design scalable storage solutions for petabytes of IoT and time-series data Develop and maintain real-time ingestion systems to support growing data volumes Implement distributed tracing, data lineage, and observability to improve monitoring and troubleshooting What You Bring (Must Have) 8 plus years of experience in platform engineering or data engineering 6 plus years designing and optimizing data pipelines at TB to PB scale Strong Java proficiency with a focus on clean, maintainable production code Strong system design skills for big data and real-time workflows Experience with lakehouse and streaming ecosystems such as Iceberg or Delta, plus Kafka or Flink or Spark Tech Stack Languages: Java, Python Framework: Spring Boot Storage: AWS S3, DynamoDB, Apache Iceberg, Redis Streaming: AWS Kinesis, Kafka, Flink ETL: AWS Glue, Spark Serverless: SQS, EventBridge, Lambda, Step Functions Infrastructure as Code: AWS CDK CI/CD: GitHub Actions

What you’ll do

Own the long-term buildout of the data platform, driving architectural direction for real-time ingestion and analytics. Design and optimize scalable storage and streaming pipelines for petabyte-scale IoT and time-series data.

Requirements

Requires over 8 years of platform or data engineering experience, with at least 6 years focusing on TB to PB scale pipelines. Must be proficient in Java and experienced with lakehouse and streaming ecosystems like Iceberg, Kafka, or Flink.

Other relevant skills

Identified from the job description. Confirm important requirements above.

  • Java
  • Python
  • System Design
  • Big Data
  • Data Pipeline Optimization
  • Apache Iceberg
  • Apache Kafka
  • Apache Flink
  • Apache Spark
  • AWS S3
  • DynamoDB
  • Redis
  • AWS Kinesis
  • AWS Glue
  • AWS CDK
  • GitHub Actions

Job areas

  • Data & Analytics
  • Software
  • Technology
  • Engineering
  • Transportation

Additional details

Minimum experience
10+ years
Apply by
Aug 29, 2026
Posting language
English
Working hours
40 hours per week
Seniority
Mid-Senior level
Application method
Direct apply is available