I have a dataset that my instructor provided from a company, and I was asked to prepare it for machine learning.
There are several missing values in the dataset, and I am unsure how they should be handled or imputed. Я также не уверен, каким стандартным практикам или рабочему процессу следует следовать при подготовке данных машинного обучения. Although I conducted my own research, I was unable to find sufficiently clear or satisfactory guidance.
Since this is my first time performing a full data preparation pipeline, I would appreciate guidance on how to approach this process correctly. Мы также будем признательны за любые рекомендации относительно надежных учебных ресурсов или ссылок.
Подробнее здесь: https://stackoverflow.com/questions/798 ... e-learning