Databricks Trainer- Dutch Speaker only
Il y a 6 jours
Brussels, Brussels-Capital, Belgique
MEERAB GROUP
Temps plein
Gratuit avec email ou Google
Enregistrez cette offre et organisez votre recherche
Créez un compte gratuit pour enregistrer des offres d'emploi, créer des alertes et revenir à cette liste depuis votre tableau de bord.
Gratuit avec email ou Google
En continuant, vous acceptez nos Conditions d’utilisation & Politique de confidentialité.
Meerab Group is a Belgium-based IT Consulting and Technology Solutions company, supporting organizations across Belgium and the wider European market. We provide technology Solutions for critical projects across Data & AI, Cloud, Cybersecurity, ServiceNow, SAP, Software Development and Project Management
We would like the training to cover the following topics:
• Databricks Functionality o Databricks Unity Catalog (specifically including the Lineage View) o Using Notebooks – %md | %sql | %python o The MS Word-like look and feel of Notebooks o Notebooks Revision History functionality o Notebooks Export/Import functionality o Visualization capabilities in Databricks (both visuals within a cell and separate dashboards) o Using Genie o Defining temporary views to build upon in subsequent commands o Reading Excel files and writing them to a temporary view o Interaction with Power BI – How to load Databricks output datasets into Power BI o Delta Lake storage and partitioning o Performance analysis – identifying the steps the cluster executes and pinpointing where slowdowns occur o Useful PySpark libraries ▪ Installing/importing libraries ▪ PySpark DataFrame and SQL library (pyspark.sql) and scenarios where this is preferred over SQL o Advanced SQL: Nested queries, JOIN clause conditions, CASE WHEN, querying metadata (Information Schema), VERSION AS OF o Using parameters in Databricks o Scheduling workflows in Databricks o Serverless vs. Dedicated Clusters -
• Databricks Functionality o Databricks Unity Catalog (specifically including the Lineage View) o Using Notebooks – %md | %sql | %python o The MS Word-like look and feel of Notebooks o Notebooks Revision History functionality o Notebooks Export/Import functionality o Visualization capabilities in Databricks (both visuals within a cell and separate dashboards) o Using Genie o Defining temporary views to build upon in subsequent commands o Reading Excel files and writing them to a temporary view o Interaction with Power BI – How to load Databricks output datasets into Power BI o Delta Lake storage and partitioning o Performance analysis – identifying the steps the cluster executes and pinpointing where slowdowns occur o Useful PySpark libraries ▪ Installing/importing libraries ▪ PySpark DataFrame and SQL library (pyspark.sql) and scenarios where this is preferred over SQL o Advanced SQL: Nested queries, JOIN clause conditions, CASE WHEN, querying metadata (Information Schema), VERSION AS OF o Using parameters in Databricks o Scheduling workflows in Databricks o Serverless vs. Dedicated Clusters -