How can data be transferred when migrating a materialized view from Greenplum to ClickHouse?
Data Engineer
What was your role in a team of 4 people, and how do you see your career in 3-4 years, what goals do you set?
Approximately how many uploads were there — top-level jobs, threads?
How well do you know Python?
When there is high memory consumption on CTE, what happens to the data?
How difficult was it to maintain a large storage of historical data manually in terms of volume and objects?
We input metadata, and it creates a flow based on this metadata — how will this help quickly migrate existing flows from old systems to the planned one?
In terms of requests, what happens to data, including what types of joins — what types of joins can we see there?
What type of read is used in ClickHouse?
How to load data from Kafka into ClickHouse via streaming for real-time reporting?
How did you work with analytical functions (window functions)?
With Kafka, there's nothing difficult, the main thing is to process data — what nuances should be considered?
In a team of 4 people, was there a leader or someone else led it?
How can data from Greenplum and Trino be transferred to ClickHouse?
Was the data warehouse built according to the Data Vault scheme — did you also develop and support it? Did you add hubs and links manually?
Do you already have any job proposals?
You work with Data Science and big data processing — is there a replacement in this library written in Rust, multithreaded, unlike Pandas, very fast. Do you know such?
Is Greenplum more related to OLTP or OLAP?
Is ACID maintained in Greenplum?
How does CTE differ from a subquery?