Dergiler / Bahçe / 2019 / Cilt: 48 - Sayı: 2
BİTKİ ISLAHINDA GENOTİP VERİM DEĞERİNİN REGRESYON YÖNTEMLERİ İLE TAHMİNİ
- Dergi
- Bahçe
- Sayfa
- 79–85
- DOI
- —
Özet
Bilimsel araştırmaların amacı yapılan çalışmaların gözlem ve denemelerinden genel sonuçlara ulaşmaktır.Gelişen teknolojilerle beraber bu sonuçlar dijital olarak kayıt altına alınmakta ve bu kayıtlar büyük veri(big data) yığınlarını meydana getirmektedir. Bu yığınların işlenmesi yani anlamlı bilgiye dönüştürülmesi1950’li yıllarda başlamış ve veri madenciliği kavramı ortaya çıkmıştır. Tahmin ya da karar vermesüreçlerinde kullanılan veri madenciliği, günümüzde tarımsal faaliyetlerin tahmin çalışmalarında dakendine yer bulmaktadır. Bitki ıslah çalışmalarının temeli, istenilen fenotip ve genotip özelliklerinin verimve çevre şartlarına göre karşılaştırılması esasına dayanmaktadır. Bu sonuçların değerlendirilmesinde çeşitliistatistik paket programları kullanılmaktadır. Kullanılan bu programlar bir ıslahçının verim ile yapılacakgenotip seçimleri için gerekli analiz ve raporlama kabiliyetlerini tam olarak karşılamamakladır. Buçalışmada, 12 lokasyondan, 24 genotipe ait 4 tekerrürlü toplam 1153 adet verim değerine göre genotipe aitverim tahmini yapılmıştır. Tahminlemede bitki ıslahında kullanılan doğrusal regresyonun yanında makineöğrenmesi metotlarından Sıralı Minimal Optimizasyon (SMO), En Yakın k–Komşu (k–EYK), RastgeleOrman (RO) metotları seçilmiştir. Seçilen metotların başarıları Ortalama Karesel Hatanın Karekökü veOrtalama Mutlak Hata metriklerine göre karşılaştırılmıştır. RO, diğer üç yönteme göre daha yüksekperformans göstermiş ve bitki ıslah programlarında kullanılan doğrusal regresyon yöntemi ile beraberkullanılması önerilmiştir.
Abstract
The aim of scientific research is to reach general results from the observations and experiments of the studies. Together with the developing technologies, these results are recorded digitally, and these records form big data stacks. The process of processing these masses into meaningful information began in the 1950s and the concept of data mining emerged. Data mining, which is used in forecasting or decision making processes, now finds its place in forecasting agricultural activities. The basis of the plant breeding studies is based on the comparison of the desired phenotype and genotype properties according to the efficiency and environmental conditions. Various statistical package programs are used in the evaluation of these results. These programs do not fully meet the analysis and reporting capabilities required for a breeder’s genotype selection. In this study, the yield of the genotype was estimated from 12 locations with a total of 1153 yields of 24 replicates of 24 genotypes. In addition to the linear regression used in plant breeding, sequential Minimal Optimization (SMO), Nearest k–Neighbor (k–EYK), Random Forest (RO) methods were selected from the methods of machine learning. The success of selected methods was compared according to the mean square root of the mean square error and the mean absolute error metrics. RO has shown higher performance than the other three methods and it has been proposed to use with the linear regression method used in plant breeding programs.