Please use this identifier to cite or link to this item: https://hdl.handle.net/1959.11/61878
Title: Application of machine-learning algorithms to predict calving difficulty in Holstein dairy cattle
Contributor(s): Avizheh, Mahdieh (author); Dadpasand, Mohammad (author); Dehnavi, Elena  (author)orcid ; Keshavarzi, Hamideh (author)
Publication Date: 2023-05-15
Open Access: Yes
DOI: 10.1071/AN22461
Handle Link: https://hdl.handle.net/1959.11/61878
Abstract: 

Context. An ability to predict calving difficulty could help farmers make better farm-management decisions, thereby improving dairy farm profitability and welfare. Aims. This study aimed to predict calving difficulty in Iranian dairy herds using machine-learning (ML) algorithms and to evaluate sampling methods to deal with imbalanced datasets. Methods. For this purpose, the history records of cows that calved between 2011 and 2021 on two commercial dairy farms were used. Using WEKA software, four commonly used ML algorithms, namely naïve Bayes, random forest, decision trees, and logistic regression, were applied to the dataset. The calving difficulty was considered as a binary trait with 0, normal or unassisted calving, and 1, difficult calving, i.e. receiving any help during parturition from farm personnel involvement to surgical intervention. The average rate of difficult calving was 18.7%, representing an imbalanced dataset. Therefore, down-sampling and cost-sensitive techniques were implemented to tackle this problem. Different models were evaluated on the basis of F-measure and the area under the curve. Key results. The results showed that sampling techniques improved the predictive model (P = 0.07, and P = 0.03, for down-sampling and cost-sensitive techniques respectively). F-measure ranged from 0.387 (decision tree) to 0.426 (logistic regression) with the balanced dataset. However, when applied to the original imbalanced dataset, naïve Bayes had the best performance of up to 0.388 in terms of F-measure. Conclusions. Overall, sampling techniques improved the prediction model compared with original imbalanced dataset. Although prediction models performed worse than expected (due to an imbalanced dataset, and missing values), the implementation of ML algorithms can still lead to an effective method of predicting calving difficulty. Implications. This research indicated the capability of ML algorithms to predict the incidence of calving difficulty within a balanced dataset, but that more explanatory variables (e.g. genetic information) are required to improve the prediction based on an unbalanced original dataset.

Publication Type: Journal Article
Source of Publication: Animal Production Science, 63(10-11), p. 1095-1104
Publisher: CSIRO Publishing
Place of Publication: Australia
ISSN: 1836-5787
1836-0939
Fields of Research (FoR) 2020: 300305 Animal reproduction and breeding
Socio-Economic Objective (SEO) 2020: 100401 Beef cattle
Peer Reviewed: Yes
HERDC Category Description: C1 Refereed Article in a Scholarly Journal
Appears in Collections:Animal Genetics and Breeding Unit (AGBU)
Journal Article

Files in This Item:
2 files
File Description SizeFormat 
openpublished/ApplicationDehnavi2023JournalArticle.pdfPublished Version1.31 MBAdobe PDF
Download Adobe
View/Open
Show full item record

SCOPUSTM   
Citations

2
checked on Sep 28, 2024
Google Media

Google ScholarTM

Check

Altmetric


This item is licensed under a Creative Commons License Creative Commons