Senior Data Engineer responsible for building and evolving core data platforms at Alpaca, a global leader in brokerage infrastructure. Working with modern data technologies and a distributed team on financial services data.
Responsibilities
Design, build, and evolve the core data platform infrastructure e.g., distributed query engines, orchestration, warehousing, cataloging, and more.
Own our lakehouse infrastructure as code, managing deployments through Terraform and Ansible on Kubernetes.
Build and maintain low-latency streaming and CDC ingestion pipelines, as well as batch ingestion paths landing in Iceberg.
Develop and scale our BI landscape so downstream teams and agents get performant, self-serve access to lakehouse data.
Enforce platform reliability best practices, including monitoring and alerting, on-call rotations, incident response, maintenance windows, runbooks, and SLAs.
Partner with DevOps, Analytics Engineering, and other stakeholders to close infrastructure gaps and support new data requirements.
Requirements
5+ years of experience in Data Engineering, including 2+ years building and operating scalable, low-latency data platforms handling > 100M events/day.
Strong hands-on experience running data infrastructure on Kubernetes, with cloud-native tooling like Docker and Helm.
Production experience with IaC: Terraform, Ansible, and ArgoCD (or equivalents).
Deep knowledge of distributed systems (storage, transactions, and query processing) with hands-on experience operating open-source query engines like Trino or Presto.
Strong experience with object storage and open table formats, specifically Apache Iceberg.
Experience with streaming and CDC systems: Kafka, Redpanda, and Debezium.
Hands-on experience with orchestration frameworks (Airflow) and ELT tools (Airbyte).
Strong working knowledge of Python and SQL for building pipelines and platform tooling.
Experience with Google Cloud Platform and its data services (GCS, Cloud Build, Cloud SQL, Dataproc, etc); or related experience with other cloud services.
Ability to thrive in a fast-paced startup environment and adapt infrastructure to rapidly changing needs.
Benefits
Competitive Salary & Stock Options
Health Benefits
New Hire Home-Office Setup: One-time USD $500
Monthly Stipend: USD $150 per month via a Brex Card
Senior Data Architect modernizing Alberta justice data marts into an integrated enterprise data warehouse. Designing ETL, Power BI testing, data models, governance, and reporting architecture.
Senior Data Architect integrating Alberta court data marts into an enterprise data warehouse. Designing ETL, data models, governance practices, and Power BI validation reports.
Data & Analytics Engineer building scalable data platforms for Ledgebrook, an insurance company. Designing pipelines, governance, and cloud data systems while partnering with technical and insurance teams.
Data Engineer building scalable pipelines for SumerSports’ football intelligence platform. Supporting deep learning, video, LLM, analytics, and AI - driven products across sports.
Data Platform Engineer building ingestion pipelines, storage, governance, and observability systems for Movable Ink’s AI - driven marketing personalization platform. Supporting scalable, secure, multi - tenant data services.
Senior Data Engineer building AWS ETL/ELT pipelines and dbt models for Tango’s cloud real - estate and facilities SaaS. Managing databases, data quality, warehousing, and observability.
Senior Data Engineer building AWS and dbt data pipelines for Tango’s cloud real - estate and facilities SaaS platform. Managing databases, warehouses, data quality, and analytics - ready datasets.
Data Engineer building streaming pipelines and analytical databases for Movable Ink’s data - activated marketing personalization platform. Developing Elixir/Python services and reliable event - data products at billions - event scale.
Senior Data Engineer needed for a 1 - year contract with an asset management client in Toronto. Requires 7+ years of experience, Snowflake, Python, SQL, and cloud platform expertise.