Statistical learning theory serves as the foundational bedrock of Machine learning (ML), which in turn represents the backbone of artificial intelligence, ushering in innovative solutions for real-world challenges. Its origins can be linked to the point where statistics and the field of computing meet, evolving into a distinct scientific discipline. Machine learning can be distinguished by its fundamental branches, encompassing supervised learning, unsupervised learning, semi-supervised learning, and reinforcement learning. Within this tapestry, supervised learning takes center stage, divided in two fundamental forms: classification and regression. Regression is tailored for continuous outcomes, while classification specializes in categorical outcomes, with the overarching goal of supervised learning being to enhance models capable of predicting class labels based on input features. This review endeavors to furnish a concise, yet insightful reference manual on machine learning, intertwined with the tapestry of statistical learning theory (SLT), elucidating their symbiotic relationship. It demystifies the foundational concepts of classification, shedding light on the overarching principles that govern it. This panoramic view aims to offer a holistic perspective on classification, serving as a valuable resource for researchers, practitioners, and enthusiasts entering the domains of machine learning, artificial intelligence and statistics, by introducing concepts, methods and differences that lead to enhancing their understanding of classification methods.
This study aims to document and describe the speech sounds and sound inventory that are present in Sorani which is dialect of Kurdish and compare the results with their English counterparts. The research concentrates on the voicing system and the quality of Sorani sounds which are measured by using the voice onset time (VOT) of the stop consonants, and the first three formants of the vowel sounds; the closure duration of voiceless stop consonants in medial position is measured as well.
Ten native speakers of the Sorani dialect (5 males and 5 females) participated in this experiment. All speakers are between 20 and 50 years of age, were born in Sulaimanyiah, migrated to the US, and remain in the US at the time of recording.
... Show MoreAccurate land use and land cover (LU/LC) classification is essential for various geospatial applications. This research applied a Spectral Angle Mapper (SAM) classifier on the Landsat 7 (ETM+ 2010) & 8 (OLI 2020) satellite scenes to identify the land cover materials of the Shatt al-Arab region which is located in the east of Basra province during ten years with an estimate of the spectral signature using ENVI 5.6 software of each cover with the proportion of its area to the area of the study region and produce maps of the classified region. The bands of these datasets were analyzed using the Optimum Index Factor (OIF) statistic. The highest OIF represents the best and most appropr
This paper proposes a new approach, of Clustering Ultrasound images using the Hybrid Filter (CUHF) to determine the gender of the fetus in the early stages. The possible advantage of CUHF, a better result can be achieved when fuzzy c-mean FCM returns incorrect clusters. The proposed approach is conducted in two steps. Firstly, a preprocessing step to decrease the noise presented in ultrasound images by applying the filters: Local Binary Pattern (LBP), median, median and discrete wavelet (DWT), (median, DWT & LBP) and (median & Laplacian) ML. Secondly, implementing Fuzzy C-Mean (FCM) for clustering the resulted images from the first step. Amongst those filters, Median & Lap
This paper proposes a new approach, of Clustering Ultrasound images using the Hybrid Filter (CUHF) to determine the gender of the fetus in the early stages. The possible advantage of CUHF, a better result can be achieved when fuzzy c-mean FCM returns incorrect clusters. The proposed approach is conducted in two steps. Firstly, a preprocessing step to decrease the noise presented in ultrasound images by applying the filters: Local Binary Pattern (LBP), median, median and discrete wavelet (DWT),(median, DWT & LBP) and (median & Laplacian) ML. Secondly, implementing Fuzzy C-Mean (FCM) for clustering the resulted images from the first step. Amongst those filters, Median & Laplace has recorded a better accuracy. Our experimental evaluation on re
... Show MorePrecise and interpretable classification of autism-related behaviors is important for initial diagnosis, personalized intervention, and support arrangements. This study proposes an interpretable machine learning (ML) model using Light Gradient Boosting Machine (LightGBM) and Categorical Boosting (CatBoost) to classify behavioral patterns into four categories (normal, mild, moderate, and severe) associated with Autism Spectrum Disorder (ASD) based on a custom 377-instance survey dataset from Iraqi parents and teachers of children aged 6-12. The model observes 16 key features across communication and social interaction, repetitive behaviors, language, and adaptive skills, preprocessed via interquartile range (IQR) outlier removal, me
... Show MoreEarth’s climate changes rapidly due to the increases in human demands and rapid economic growth. These changes will affect the entire biosphere, mostly in negative ways. Predicting future changes will put us in a better position to minimize their catastrophic effects and to understand how humans can cope with the new changes beforehand. In this research, previous global climate data set observations from 1961-1990 have been used to predict the future climate change scenario for 2010-2039. The data were processed with Idrisi Andes software and the final Köppen-Geiger map was created with ArcGIS software. Based on Köppen climate classification, it was found that areas of Equator, Arid Steppes, and Snow will decrease by 3.9 %, 2.96%, an
... Show MoreAlthough its wide utilization in microbial cultures, the one factor-at-a-time method, failed to find the true optimum, this is due to the interaction between optimized parameters which is not taken into account. Therefore, in order to find the true optimum conditions, it is necessary to repeat the one factor-at-a-time method in many sequential experimental runs, which is extremely time-consuming and expensive for many variables. This work is an attempt to enhance bioactive yellow pigment production by Streptomyces thinghirensis based on a statistical design. The yellow pigment demonstrated inhibitory effects against Escherichia coli and Staphylococcus aureus and was characterized by UV-vis spectroscopy which showed lambda maximum of
... Show More