Mistplay
Principal Platform Engineer, ML
Toronto · Principal
Sponsorship not specifiedDetected 180 days ago
TypeScriptPythonJavaGoVue.jsTerraformCI/CDDevOpsSite Reliability EngineeringPlatform EngineeringMachine LearningTensorFlowPyTorchscikit-learnAirflowData EngineeringData ScienceIncident ResponseCommunicationCollaboration
About the role
- Mistplay est l'application de fidélité n°1 pour les joueurs mobiles.
- Notre communauté de millions de joueurs mobiles engagés utilise Mistplay pour découvrir de nouveaux jeux et gagner des récompenses.
- Les joueurs sont récompensés pour le temps et l'argent qu'ils consacrent aux jeux et peuvent échanger ces récompenses contre des cartes cadeaux.
Responsibilities
- Be the main driver and expert for designing, building, and operating:
- High accuracy low latency feature serving layer and preprocessing solutions to support online serving of the models
- Build platform abstractions and golden paths: Airflow DAG templates, CLI/SDKs, cookie-cutter repos, and CI/CD pipelines that take models from notebooks to production predictably.
- Partner with Security, SRE, and Data Engineering on private networking, policy-as-code, PII handling, least-privilege IAM, and cost-efficient architectures across environments.
- lead migrations with clear change management and minimal downtime. What you'll bring:
- 10+ years building and operating production-grade ML/data platforms with a focus on serving, reliability, and developer experience.
- Evaluate, integrate, and rationalize platform tooling (e.g., MLflow registry, feature stores, serving gateways); lead migrations with clear change management and minimal downtime.
Requirements
- Expertise with online feature store paradigms and underlying storage solutions in ML serving contexts.
- familiarity with GitOps patterns.
- Deep experience with inference solutions: endpoint configuration, containerization, model packaging, autoscaling, serverless vs. real-time trade-offs, MME, A/B and canary releases.
Nice to have
- Solutions d'infrastructure machine et de données pour l'entraînement des modèles.
- Systèmes d'inférence en temps réel pour exploiter et servir des modèles dans un environnement de production en temps réel.
- Capacités de plateforme de fonctionnalités de haute convivialité et précision pour générer, remplir rétrospectivement et stocker des fonctionnalités au niveau de l'utilisateur.
- Couche de service de fonctionnalités à haute précision et faible latence, et solutions de pré-traitement pour prendre en charge le service en ligne des modèles.
- Évaluer, intégrer et rationaliser les outils de plateforme (par exemple, registre MLflow, magasins de fonctionnalités, passerelles de service)
- mener des migrations avec une gestion claire des changements et un temps d'arrêt minimal.
- Ce que vous apporterez
- 10 ans et plus d'expérience dans la construction et l'exploitation de plateformes ML/de données de qualité production, en mettant l'accent sur le service, la fiabilité et l'expérience développeur.
Benefits
- Implement end-to-end observability: data/feature freshness checks, drift/quality gates, model performance/latency SLOs, infra health dashboards, tracing, and alerting-plus incident response and postmortems.
- experience building resilient services, APIs, and automation tooling with high test coverage.
- Working at Mistplay is coupled with a whole array of perks that we've adopted virtually and in-person:
Company info
- We strive to make our work environment as inviting and fun as possible!
- Our culture is deeply rooted in growth and upheld by a team of smart, dynamic, and enthusiastic people.
Apply directly at Mistplay →Create a free account for alerts like thisView Mistplay immigration profile
This listing is sourced directly from Mistplay's careers page and normalized into a canonical job model.