Question Bank
Machine Learning interview questions Naming a model or a metric is only the start of an interview answer. Explain why it fits the task, and connect training, regularization, validation and evaluation to a concrete example. Start with your own explanation, then state its assumptions and limitations.
Hard Machine Learning Мосбиржа
Which ROC-AUC variation is better: 0.1, 0.7, or 0.85?
roc-auc metrics evaluation
Practice in bot Hard Machine Learning Мосбиржа
For heavily imbalanced classes, which to choose: ROC-AUC or PR-AUC?
imbalanced-data metrics pr-auc roc-auc
Practice in bot Medium Machine Learning Lamoda
We have regression task, target from 0 to 100. What will be the output distribution for each algorithm?
regression output-distribution algorithms
Practice in bot Medium Machine Learning Lamoda
What categorical variable encoding methods exist? Pros and cons?
categorical-encoding one-hot target-encoding
Practice in bot Medium Machine Learning Lamoda
What are the disadvantages of target encoding?
target-encoding data-leakage overfitting
Practice in bot Hard Machine Learning Lamoda
How does target encoding work in CatBoost?
catboost target-encoding ordered-statistics
Practice in bot Medium Machine Learning Lamoda
How to build a tree in general, what criteria to use?
decision-trees tree-construction information-gain
Practice in bot Medium Machine Learning Lamoda
What are the advantages and disadvantages of trees?
decision-trees interpretability overfitting
Practice in bot Hard Machine Learning Lamoda
In terms of bias-variance, how does a tree behave?
decision-trees bias-variance overfitting
Practice in bot Medium Machine Learning Lamoda
What kind of trees are built in boosting and random forest?
random-forest gradient-boosting ensemble
Practice in bot Medium Machine Learning Lamoda
Best practice for boosting tuning - what to tune first?
hyperparameter-tuning gradient-boosting optimization
Practice in bot Medium Machine Learning Lamoda
What happens to variance in random forest?
random-forest variance-reduction ensemble
Practice in bot Hard Machine Learning Lamoda
Why does model variance plateau with increasing complexity?
bias-variance model-complexity overfitting
Practice in bot Medium Machine Learning Lamoda
How is gradient used in gradient boosting?
gradient-boosting optimization loss-function
Practice in bot Hard Machine Learning Lamoda
What does the next tree learn in boosting?
boosting gradient-descent residuals
Practice in bot Medium Machine Learning Lamoda
How to evaluate feature importance - feature importance vs SHAP?
feature-importance shap interpretability
Practice in bot Medium Machine Learning
Why doesn't boosting extrapolate beyond training data?
gradient-boosting extrapolation decision-trees
Practice in bot Medium Machine Learning
What do you understand by hyperparameter optimization?
hyperparameter-optimization grid-search bayesian-optimization
Practice in bot Medium Machine Learning
How to prevent data leakage in medical data?
data-leakage medical-data temporal-splitting
Practice in bot Medium Machine Learning
What is model governance and why is it important in production?
model-governance mlops compliance
Practice in bot Prepare for your next interview with Vibe Interview.
Download Vibe Interview