Movable Ink

Movable Ink

Senior Data Engineer, AI Systems

Movable Ink - Toronto · Senior

Sponsorship not specifiedDetected 47 days ago
PythonDistributed SystemsCode ReviewGitGCPCloud PlatformsDockerKubernetesCI/CDGitHub ActionsKafkaMachine LearningSparkData EngineeringMLOpsAccessibility

About the role

  • Movable Ink scales content personalization for marketers through data-activated content generation and AI decisioning.
  • The world's most innovative brands rely on Movable Ink to maximize revenue, simplify workflow and boost marketing agility.
  • Headquartered in New York City with close to 600 employees, Movable Ink serves its global client base with operations throughout North America, Central America, Europe, Australia, and Japan.

Responsibilities

  • Build, maintain, and optimize production data pipelines that power AI-driven personalization at scale across content selection, send-time optimization, subject line personalization, and frequency capping
  • Own and scale Spark-based batch pipelines, including cluster configuration, tuning, and performance optimization across GCP Dataproc
  • Build and maintain our ML Data Lake, ensuring data quality, accessibility, and efficient storage
  • Support the data needs of ML Engineers and Scientists for model development, training, and evaluation
  • Collaborate with distributed systems engineers on the platform's architectural evolution, ensuring data layer continuity throughout
  • Release features and data products that deliver measurable and tangible business value
  • The AI Systems team owns the core recommendations engine and ML platform that powers billions of AI-driven marketing decisions daily across some of the world's largest consumer brands.
  • As a Senior Data Engineer, you will own the Spark-based data pipelines and data infrastructure at the heart of this system - building, scaling, and optimizing the data layer that feeds our production ML models.

Requirements

  • 5+ years of data engineering experience
  • Deep expertise with Apache Spark, including the PySpark DataFrame API and experience solving challenging scaling problems
  • Experience with large-scale data processing, cluster configuration, optimization, and tuning (we use GCP Dataproc)
  • Experience with data storage formats (we use Parquet, Delta Lake)
  • Experience with event streaming data (we use Kafka)
  • Experience with cloud computing platforms (we use Google Cloud Platform)
  • Experience with advanced query optimization
  • Strong software development skills in Python (unit testing, git, code review, CI/CD)
  • Familiar with Software Development Lifecycle practices, such as continuous integration/continuous delivery and automated deployment (we use Docker, Kubernetes, and GitHub Actions)
  • Ability to collaborate with technical partners - you'll be working closely with ML engineers, scientists, and other teams to determine requirements and make design decisions
  • Enjoys working in a fast-paced, goal-driven environment

Compensation

  • The base pay range for this position is $144K CAD - 188K CAD/year, which can include additional bonus depending on the position ultimately offered, in addition to a full range of medical, financial, and/or other benefits.

Benefits

  • This is an opportunity to work end-to-end on large-scale data systems that touch millions of customers, on a team working at the intersection of data engineering and machine learning.

Visa & Work Authorization

  • y or expression, religion, genetic information, parental or pregnancy status, national origin, sexual orientation, age, citizenship, marital status, ethnicity, family or marital status, physical and mental ability, polit

This listing is sourced directly from Movable Ink's careers page and normalized into a canonical job model.