What is stratified sampling in experiments?
Machine Learning / AI
What is knowledge distillation for edge?
What is a DAG in Airflow?
What are indexes in databases and why are they needed?
Why is a B-tree index most commonly used and what is its algorithmic complexity?
What is Bayesian Optimization and acquisition functions (EI, UCB, PI)?
What is layout analysis in documents?
What problems arise when training logistic regression if the number of features exceeds the number of observations?
What is model quantization?
How many rows will CROSS JOIN return for tables with 100 and 10 records?
What neural network libraries have you used?
What is the difference between DWH and Data Lake?
What are the disadvantages of columnar databases?
What is self-supervised pretraining for GNN?
What solutions would you suggest when system load increases?
Why and when to use graphics cards?