Bridging the Gap between Field Experiments and Machine Learning: The EC H2020 B-GOOD Project as a Case Study towards Automated Predictive Health Monitoring of Honey Bee Colonies.
Coby Van DooremalenZeynep N UlgezenRaffaele Dall'OlioUgoline GodeauXiaodong DuanJosé Paulo SousaMarc O SchäferAlexis L BeaurepairePim van GennipMarten SchoonmanClaude J FlenerSeverine MatthijsDavid Claeys BouuaertWim VerbekeDana FreshleyDirk-Jan ValkenburgTrudy van den BoschFamke SchaafsmaJeroen PetersMang XuYves Le ConteCedric AlauxAnne DalmonRobert J PaxtonAnja TehelTabea StreicherDaniel Severus DezmireanAlexandru-Ioan GiurgiuChristopher J ToppingJames Henty WilliamsNuno CapelaSara LopesFátima AlvesJoana AlvesJoão BicaSandra SimõesAntónio Alves da SilvaSílvia CastroJoão LoureiroEva HorčičkováMartin BencsikAdam McVeighTarun KumarArrigo MoroApril van DeldenElżbieta ZiółkowskaMichał FilipiakŁukasz MikołajczykKirsten LeufgenLina De SmetDirk C de GraafPublished in: Insects (2024)
Honey bee colonies have great societal and economic importance. The main challenge that beekeepers face is keeping bee colonies healthy under ever-changing environmental conditions. In the past two decades, beekeepers that manage colonies of Western honey bees ( Apis mellifera ) have become increasingly concerned by the presence of parasites and pathogens affecting the bees, the reduction in pollen and nectar availability, and the colonies' exposure to pesticides, among others. Hence, beekeepers need to know the health condition of their colonies and how to keep them alive and thriving, which creates a need for a new holistic data collection method to harmonize the flow of information from various sources that can be linked at the colony level for different health determinants, such as bee colony, environmental, socioeconomic, and genetic statuses. For this purpose, we have developed and implemented the B-GOOD (Giving Beekeeping Guidance by computational-assisted Decision Making) project as a case study to categorize the colony's health condition and find a Health Status Index (HSI). Using a 3-tier setup guided by work plans and standardized protocols, we have collected data from inside the colonies (amount of brood, disease load, honey harvest, etc.) and from their environment (floral resource availability). Most of the project's data was automatically collected by the BEEP Base Sensor System. This continuous stream of data served as the basis to determine and validate an algorithm to calculate the HSI using machine learning. In this article, we share our insights on this holistic methodology and also highlight the importance of using a standardized data language to increase the compatibility between different current and future studies. We argue that the combined management of big data will be an essential building block in the development of targeted guidance for beekeepers and for the future of sustainable beekeeping.
Keyphrases
- big data
- machine learning
- public health
- healthcare
- artificial intelligence
- electronic health record
- health information
- mental health
- quality improvement
- decision making
- human health
- deep learning
- health promotion
- cancer therapy
- south africa
- life cycle
- data analysis
- gene expression
- simultaneous determination
- copy number