Browse Articles

Using data science along with machine learning to determine the ARIMA model’s ability to adjust to irregularities in the dataset

Choudhary et al. | Jul 26, 2021

Using data science along with machine learning to determine the ARIMA model’s ability to adjust to irregularities in the dataset

Auto-Regressive Integrated Moving Average (ARIMA) models are known for their influence and application on time series data. This statistical analysis model uses time series data to depict future trends or values: a key contributor to crime mapping algorithms. However, the models may not function to their true potential when analyzing data with many different patterns. In order to determine the potential of ARIMA models, our research will test the model on irregularities in the data. Our team hypothesizes that the ARIMA model will be able to adapt to the different irregularities in the data that do not correspond to a certain trend or pattern. Using crime theft data and an ARIMA model, we determined the results of the ARIMA model’s forecast and how the accuracy differed on different days with irregularities in crime.

Read More...

Quantitative analysis and development of alopecia areata classification frameworks

Dubey et al. | Jun 03, 2024

Quantitative analysis and development of alopecia areata classification frameworks

This article discusses Alopecia areata, an autoimmune disorder causing sudden hair loss due to the immune system mistakenly attacking hair follicles. The article introduces the use of deep learning (DL) techniques, particularly convolutional neural networks (CNN), for classifying images of healthy and alopecia-affected hair. The study presents a comparative analysis of newly optimized CNN models with existing ones, trained on datasets containing images of healthy and alopecia-affected hair. The Inception-Resnet-v2 model emerged as the most effective for classifying Alopecia Areata.

Read More...

Comparing the Dietary Preference of Caenorhabditis elegans for Bacterial Probiotics vs. Escherichia coli.

Lulla et al. | Dec 18, 2020

Comparing the Dietary Preference of <i>Caenorhabditis elegans</i> for Bacterial Probiotics vs. <i>Escherichia coli</i>.

In this experiment, the authors used C. elegans as a simple model organism to observe the impact of probiotics on the human digestive system. The results of the experiments showed that the C. elegans were, on average, most present in Chobani cultures over other tested yogurts. While not statistically significant, these results still demonstrated that C. elegans might prefer Chobani cultures over other probiotic yogurts, which may also indicate greater gut benefits from Chobani over the other yogurt brands tested.

Read More...

Substance Abuse Transmission-Impact of Parental Exposure to Nicotine/Alcohol on Regenerated Planaria Offspring

Bennet et al. | Jul 02, 2024

Substance Abuse Transmission-Impact of Parental Exposure to Nicotine/Alcohol on Regenerated Planaria Offspring

The global mental health crisis has led to increased substance abuse among youth. Prescription drug abuse causes approximately 115 American deaths daily. Understanding intergenerational transmission of substance abuse is complex due to lengthy human studies and socioeconomic variables. Recent FDA guidelines mandate abuse liability testing for neuro-active drugs but overlook intergenerational transfer. Brown planaria, due to their nervous system development similarities with mammals, offer a novel model.

Read More...

Optimizing data augmentation to improve machine learning accuracy on endemic frog calls

Anand et al. | Mar 09, 2025

Optimizing data augmentation to improve machine learning accuracy on endemic frog calls
Image credit: Anand and Sampath 2025

The mountain chain of the Western Ghats on the Indian peninsula, a UNESCO World Heritage site, is home to about 200 frog species, 89 of which are endemic. Distinctive to each frog species, their vocalizations can be used for species recognition. Manually surveying frogs at night during the rain in elephant and big cat forests is difficult, so being able to autonomously record ambient soundscapes and identify species is essential. An effective machine learning (ML) species classifier requires substantial training data from this area. The goal of this study was to assess data augmentation techniques on a dataset of frog vocalizations from this region, which has a minimal number of audio recordings per species. Consequently, enhancing an ML model’s performance with limited data is necessary. We analyzed the effects of four data augmentation techniques (Time Shifting, Noise Injection, Spectral Augmentation, and Test-Time Augmentation) individually and their combined effect on the frog vocalization data and the public environmental sounds dataset (ESC-50). The effect of combined data augmentation techniques improved the model's relative accuracy as the size of the dataset decreased. The combination of all four techniques improved the ML model’s classification accuracy on the frog calls dataset by 94%. This study established a data augmentation approach to maximize the classification accuracy with sparse data of frog call recordings, thereby creating a possibility to build a real-world automated field frog species identifier system. Such a system can significantly help in the conservation of frog species in this vital biodiversity hotspot.

Read More...

Using machine learning to develop a global coral bleaching predictor

Madireddy et al. | Feb 21, 2023

Using machine learning to develop a global coral bleaching predictor
Image credit: Madireddy, Bosch, and McCalla

Coral bleaching is a fatal process that reduces coral diversity, leads to habitat loss for marine organisms, and is a symptom of climate change. This process occurs when corals expel their symbiotic dinoflagellates, algae that photosynthesize within coral tissue providing corals with glucose. Restoration efforts have attempted to repair damaged reefs; however, there are over 360,000 square miles of coral reefs worldwide, making it challenging to target conservation efforts. Thus, predicting the likelihood of bleaching in a certain region would make it easier to allocate resources for conservation efforts. We developed a machine learning model to predict global locations at risk for coral bleaching. Data obtained from the Biological and Chemical Oceanography Data Management Office consisted of various coral bleaching events and the parameters under which the bleaching occurred. Sea surface temperature, sea surface temperature anomalies, longitude, latitude, and coral depth below the surface were the features found to be most correlated to coral bleaching. Thirty-nine machine learning models were tested to determine which one most accurately used the parameters of interest to predict the percentage of corals that would be bleached. A random forest regressor model with an R-squared value of 0.25 and a root mean squared error value of 7.91 was determined to be the best model for predicting coral bleaching. In the end, the random model had a 96% accuracy in predicting the percentage of corals that would be bleached. This prediction system can make it easier for researchers and conservationists to identify coral bleaching hotspots and properly allocate resources to prevent or mitigate bleaching events.

Read More...

Large Language Models are Good Translators

Zeng et al. | Oct 16, 2024

Large Language Models are Good Translators

Machine translation remains a challenging area in artificial intelligence, with neural machine translation (NMT) making significant strides over the past decade but still facing hurdles, particularly in translation quality due to the reliance on expensive bilingual training data. This study explores whether large language models (LLMs), like GPT-4, can be effectively adapted for translation tasks and outperform traditional NMT systems.

Read More...

Modeling the heart’s reaction to narrow blood vessels

Athulathmudali et al. | May 22, 2023

Modeling the heart’s reaction to narrow blood vessels

Cardiovascular diseases are the largest cause of death globally, making it a critical area of focus. The circulatory system is required to make the heart function. One component of this system is blood vessels, which is the focus of our study. Our work aims to demonstrate the numeric relationship between a blood vessel's diameter and the number of pumps needed to transport blood.

Read More...

Analyzing breath sounds by using deep learning in diagnosing bronchial blockages with artificial lung

Bae et al. | Jan 22, 2024

Analyzing breath sounds by using deep learning in diagnosing bronchial blockages with artificial lung

Many common respiratory illnesses like bronchitis, asthma, and chronic obstructive pulmonary disease (COPD) lead to bronchial inflammation and, subsequently, a blockage. However, there are many difficulties in measuring the severity of the blockage. A numeric metric to determine the degree of the blockage severity is necessary. To tackle this demand, we aimed to develop a novel human respiratory model and design a deep-learning program that can constantly monitor and report bronchial blockage by recording breath sounds in a non-intrusive way.

Read More...