Geospatial Data Engineer at GHGSat integrating geospatial data and optimizing AI/ML pipelines. Supporting climate impact mission through data systems and analytics in a hybrid role.
Responsibilities
Design, implement, and optimise scalable geospatial data and AI/ML pipelines.
Integrate new data sources, including satellite and terrestrial, both public and proprietary.
Re-engineer and validate existing pipelines, ensuring high-quality and performance standards.
Blend and process various geospatial data sources to create artifacts for exploratory analysis and insights.
Build scripts and automations for geospatial data processing, using tools like QGIS, GeoPandas, Rasterio, Xarray and rioxarray.
Conduct geospatial analysis and contribute to mapping and visualization.
Data testing and quality control of geospatial datasets.
Contribute to the automation of testing, deployment, and monitoring of data pipelines and AI/ML models using DBT, Airflow, Docker, and AWS services.
Work collaboratively with the Analytics team, Subject Matter Experts, and cross-teams to prototype new data solutions.
Explore applications of AI/ML for geospatial data and integrate emerging technologies where possible.
Present findings and recommendations to both technical and non-technical stakeholders, fostering a data-driven culture.
Communicate complex geospatial data insights in a clear, accessible manner to support informed outcomes.
Requirements
2-4 years of experience in data engineering, with specific expertise in geospatial data processing and analysis.
Proficiency in SQL and geospatial databases e.g., PostgresSQL/PostGIS
Experience with Airflow, DBT, and dashboarding tools such as Grafana
Proficiency in Python and experience with libraries like Pandas, pytest, NumPy, sqlalchemy.
Comfortable with cloud infrastructure (AWS preferred), containerization tools (Docker), and version control (Git).
Experience with geospatial packages such as GeoPandas, Rasterio, and QGIS is beneficial.
Knowledge of AI/ML concepts applied to geospatial data is a plus.
Knowledge of ClearML is beneficial.
Benefits
Competitive salary + stock options for all full-time employees
Staff Data Engineer building scalable pipelines and lakehouse architecture for Sonatype, a software supply chain security company. Driving trusted analytics, ML, and business intelligence data with Databricks, Spark, and modern cloud technologies.
Senior Fabric Data Engineer modernizing learning - platform data for Cornerstone Galaxy. Building Microsoft Fabric pipelines, curated datasets, governance, and analytics - ready integrations.
GCP Data Platform Engineer maintaining and optimizing production data platforms for Innodata, a global data engineering and AI services company. Supporting Airflow pipelines, GCP infrastructure, reliability, and troubleshooting.
Senior Data Engineer building finance data pipelines and models for Instacart’s grocery delivery platform. Owning financial data infrastructure supporting accounting, billing, invoicing, and reporting.
Senior Data Architect modernizing Alberta justice data marts into an integrated enterprise data warehouse. Designing ETL, Power BI testing, data models, governance, and reporting architecture.
Senior Data Architect integrating Alberta court data marts into an enterprise data warehouse. Designing ETL, data models, governance practices, and Power BI validation reports.
Data & Analytics Engineer building scalable data platforms for Ledgebrook, an insurance company. Designing pipelines, governance, and cloud data systems while partnering with technical and insurance teams.
Data Engineer building scalable pipelines for SumerSports’ football intelligence platform. Supporting deep learning, video, LLM, analytics, and AI - driven products across sports.
Data Platform Engineer building ingestion pipelines, storage, governance, and observability systems for Movable Ink’s AI - driven marketing personalization platform. Supporting scalable, secure, multi - tenant data services.