Senior Machine Learning Engineer
Il y a 6 heures
Brussels, Brussels, Belgique
Atos
Temps plein
Gratuit avec email ou Google
Enregistrez cette offre et organisez votre recherche
Créez un compte gratuit pour enregistrer des offres d'emploi, créer des alertes et revenir à cette liste depuis votre tableau de bord.
Gratuit avec email ou Google
Senior MLOps Engineer — Data Platform
Location: Brussels, hybrid
Level: Senior
Languages: English required; French a plus
About the Team We're building a data mesh on Azure and Databricks: domain teams own their data as products, and the central platform team gives them the paved road. We have an
MVP data platform live
with pilot teams onboarded. ML is the next capability: today it's a handful of models running outside the data platform, and we need a production-grade MLOps platform that domains can adopt on their own for these existing and new use-cases. The Role You'll take
ownership of the MLOps capability — its design, its implementation and its roadmap
— and build it on top of the existing platform. You start with understanding the current platform (MLOps setup, landing zones, Unity Catalog, governance, automation), assessing what can be reused, and refining an ML lifecycle that fits it rather than sits beside it. Then you build it, with the domain teams as your first users. Databricks and MLflow are at the core, so deep hands-on expertise there — Unity Catalog model registry, Model Serving, feature engineering, Lakehouse Monitoring, Asset Bundles — is essential rather than one skill among many. It's a hands-on engineering role: you write the pipelines, templates and tooling, and you embed with domain teams to see where they get stuck. We value demonstrated experience over familiarity with the concepts.
What You'll Do Own the MLOps capability end to end
— its design, implementation quality and roadmap — and be the person accountable for where it is going. Assess the existing platform
and refine the MLOps architecture to fit it: how models, features and experiments map onto the landing zone, Unity Catalog and governance model already in place, and what platform gaps need closing before ML can run there. Understand and refine the
target ML lifecycle design
with the platform architect and domain teams, then deliver it incrementally — first a working path for one domain, then a paved road for all. Implement the full ML lifecycle on
Databricks and MLflow
— experiment tracking, Unity Catalog model registry, feature tables, packaging, Model Serving and monitoring, tailored for
regulated environments . Build
controlled promotion across dev, staging and production
with CI/CD (Azure DevOps / GitHub Actions, Databricks Asset Bundles), so model releases are reproducible and auditable. Deliver using
off-the-shelf capabilities where they fit and custom components where they don't , and own that judgement call. Build automated retraining, drift and skew detection with
Lakehouse Monitoring
or equivalent, and the alerting and rollback paths that make them trustworthy. Productionise batch and
near real-time inference . Treat
models as data products
— owners, contracts, SLOs and lineage from source data through features to consumers, with health and cost signals feeding the platform-wide observability and governance views. Give domains
cost visibility for ML workloads
— spend attributed per model and domain, right-sized compute, scale-to-zero serving, and surfacing idle endpoints and abandoned experiments. Manage ML infrastructure as code with
Terraform , following platform standards, and review domain teams' ML deliverables. What Success Looks Like in the First Year Ownership established
— you are recognised by the platform team and domains as the owner of the MLOps capability and its direction. Design agreed
— an ML lifecycle architecture that fits the existing platform, reviewed and backed by the platform architect and domain stakeholders, within the first quarter. First models in production
— at least one domain running monitored, cost-visible models in production via the paved road. Smooth onboarding
— a second domain can take a model from experiment to a monitored production endpoint without platform intervention. Roadmap delivered
— the priority MLOps features [e.g. near real-time inference, automated retraining] shipped and adopted. What We're Looking For 5+ years
in MLOps, ML engineering or platform engineering, with models you've built the delivery path for and supported in production. Someone who can
own the design, implementation and roadmap of an MLOps capability through a shared vision
— assessing an existing platform, designing to fit it, aligning platform and domain teams behind the direction, and delivering incrementally. Deep, hands-on Databricks and MLflow expertise — essential.
MLflow tracking, models and registry in Unity Catalog; Model Serving, feature engineering, Lakehouse Monitoring, Workflows, Asset Bundles, system tables. You should be able to walk through ML platforms you've designed and operated on Da
About the Team We're building a data mesh on Azure and Databricks: domain teams own their data as products, and the central platform team gives them the paved road. We have an
MVP data platform live
with pilot teams onboarded. ML is the next capability: today it's a handful of models running outside the data platform, and we need a production-grade MLOps platform that domains can adopt on their own for these existing and new use-cases. The Role You'll take
ownership of the MLOps capability — its design, its implementation and its roadmap
— and build it on top of the existing platform. You start with understanding the current platform (MLOps setup, landing zones, Unity Catalog, governance, automation), assessing what can be reused, and refining an ML lifecycle that fits it rather than sits beside it. Then you build it, with the domain teams as your first users. Databricks and MLflow are at the core, so deep hands-on expertise there — Unity Catalog model registry, Model Serving, feature engineering, Lakehouse Monitoring, Asset Bundles — is essential rather than one skill among many. It's a hands-on engineering role: you write the pipelines, templates and tooling, and you embed with domain teams to see where they get stuck. We value demonstrated experience over familiarity with the concepts.
What You'll Do Own the MLOps capability end to end
— its design, implementation quality and roadmap — and be the person accountable for where it is going. Assess the existing platform
and refine the MLOps architecture to fit it: how models, features and experiments map onto the landing zone, Unity Catalog and governance model already in place, and what platform gaps need closing before ML can run there. Understand and refine the
target ML lifecycle design
with the platform architect and domain teams, then deliver it incrementally — first a working path for one domain, then a paved road for all. Implement the full ML lifecycle on
Databricks and MLflow
— experiment tracking, Unity Catalog model registry, feature tables, packaging, Model Serving and monitoring, tailored for
regulated environments . Build
controlled promotion across dev, staging and production
with CI/CD (Azure DevOps / GitHub Actions, Databricks Asset Bundles), so model releases are reproducible and auditable. Deliver using
off-the-shelf capabilities where they fit and custom components where they don't , and own that judgement call. Build automated retraining, drift and skew detection with
Lakehouse Monitoring
or equivalent, and the alerting and rollback paths that make them trustworthy. Productionise batch and
near real-time inference . Treat
models as data products
— owners, contracts, SLOs and lineage from source data through features to consumers, with health and cost signals feeding the platform-wide observability and governance views. Give domains
cost visibility for ML workloads
— spend attributed per model and domain, right-sized compute, scale-to-zero serving, and surfacing idle endpoints and abandoned experiments. Manage ML infrastructure as code with
Terraform , following platform standards, and review domain teams' ML deliverables. What Success Looks Like in the First Year Ownership established
— you are recognised by the platform team and domains as the owner of the MLOps capability and its direction. Design agreed
— an ML lifecycle architecture that fits the existing platform, reviewed and backed by the platform architect and domain stakeholders, within the first quarter. First models in production
— at least one domain running monitored, cost-visible models in production via the paved road. Smooth onboarding
— a second domain can take a model from experiment to a monitored production endpoint without platform intervention. Roadmap delivered
— the priority MLOps features [e.g. near real-time inference, automated retraining] shipped and adopted. What We're Looking For 5+ years
in MLOps, ML engineering or platform engineering, with models you've built the delivery path for and supported in production. Someone who can
own the design, implementation and roadmap of an MLOps capability through a shared vision
— assessing an existing platform, designing to fit it, aligning platform and domain teams behind the direction, and delivering incrementally. Deep, hands-on Databricks and MLflow expertise — essential.
MLflow tracking, models and registry in Unity Catalog; Model Serving, feature engineering, Lakehouse Monitoring, Workflows, Asset Bundles, system tables. You should be able to walk through ML platforms you've designed and operated on Da