Compare the Data Vault approaches in MCHS and X5. What benefits did the new objects (bridge, pit tables) provide?
Data Engineer
How often were dashboards calculated? Was it incremental or full calculation?
Tell us about your work experience
Tell about the architecture of storage: how it was divided into layers, what are the features of each layer?
What incremental loading strategies in dbt have you used?
Which Airflow operators did you use? How did you interact with dbt through Airflow?
How does data get into the storage at the current workplace?
What types of distribution did you use? How did you choose the distribution key?
Analysis of the dbt model with an incremental strategy: what does the model do, are there any issues?
What types of tables have you used in Greenplum? When was heap used, and when was Append-Optimized used?
What are the features of using XCom in Airflow?
Have you interacted with relational databases?
Tell me about indexes in Greenplum and ClickHouse
What data sources have you worked with? How was the interaction with these sources?
Tell about partitioning. What interesting operations have you had to do?
Who were the data consumers and how were the dashboards built?