Data Engineer managing ingestion pipelines in cloud-native data ecosystem for analytics, marketing, and reporting use cases. Ensuring data quality and governance throughout the process.
Responsibilities
Design, build, and operate **robust ingestion pipelines** for batch and near-real-time data using AWS-native services
Implement **CDC-based ingestion patterns** for databases, SaaS platforms, and external partners
Standardize ingestion frameworks for files, APIs, event streams, and cross-account data sharing
Define and maintain **raw and staging data models** that preserve source fidelity and lineage
Partner with source system owners to define ingestion SLAs, contracts, schemas, and change management strategies
Ensure ingestion pipelines meet **data quality, observability, and reliability standards**
Implement metadata capture, schema evolution handling, and data validation at ingestion time
Automate infrastructure using **AWS CDK** and integrate CI/CD pipelines via CodeCommit and CodePipeline
Optimize ingestion workflows for scalability, cost efficiency, and fault tolerance
Support Agile delivery and collaborate closely with offshore engineering teams
Requirements
Bachelor’s or Master’s degree in Computer Science, Engineering, or a quantitative field
5+ years of experience in data engineering, with a strong focus on **data ingestion and integration**
Strong understanding of **change data capture (CDC)** concepts and ingestion patterns
Hands-on experience with AWS services such as **Lambda, Step Functions, MWAA, Glue, Redshift**
Experience building ingestion pipelines for **APIs, files, databases, and event-based systems**
Proficiency in **Python** and familiarity with data serialization formats (JSON, Parquet, Avro)
Experience implementing infrastructure as code using **AWS CDK**
Working knowledge of CI/CD, version control, and automated deployments
Strong collaboration skills and comfort working with distributed offshore teams
Detail-oriented, proactive, and ownership-driven mindset
Senior Data Engineer building scalable, AI - powered data infrastructure for CloudBlue, HostPapa’s cloud commerce platform. Developing real - time pipelines, APIs, and production analytics systems.
Senior Fabric Data Engineer modernizing enterprise data for Canadian IT consulting clients. Building Microsoft Fabric pipelines, curated datasets, security controls, and analytics - ready models.
Staff Data Engineer building scalable pipelines and lakehouse architecture for Sonatype, a software supply chain security company. Driving trusted analytics, ML, and business intelligence data with Databricks, Spark, and modern cloud technologies.
Senior Fabric Data Engineer modernizing learning - platform data for Cornerstone Galaxy. Building Microsoft Fabric pipelines, curated datasets, governance, and analytics - ready integrations.
GCP Data Platform Engineer maintaining and optimizing production data platforms for Innodata, a global data engineering and AI services company. Supporting Airflow pipelines, GCP infrastructure, reliability, and troubleshooting.
Senior Data Engineer building finance data pipelines and models for Instacart’s grocery delivery platform. Owning financial data infrastructure supporting accounting, billing, invoicing, and reporting.
Senior Data Architect modernizing Alberta justice data marts into an integrated enterprise data warehouse. Designing ETL, Power BI testing, data models, governance, and reporting architecture.
Senior Data Architect integrating Alberta court data marts into an enterprise data warehouse. Designing ETL, data models, governance practices, and Power BI validation reports.
Data & Analytics Engineer building scalable data platforms for Ledgebrook, an insurance company. Designing pipelines, governance, and cloud data systems while partnering with technical and insurance teams.