Madde düzeyinde boyutluluk modellerinin bilgisayar ortamında bireyselleştirilmiş test yöntemleri üzerindeki etkisinin incelenmesi

Özdemir, Burhanettin

dc.contributor.advisor	Gelbal, Selahattin
dc.contributor.author	Özdemir, Burhanettin
dc.date.accessioned	2020-12-29T13:53:21Z
dc.date.available	2020-12-29T13:53:21Z
dc.date.submitted	2015
dc.date.issued	2018-08-06
dc.identifier.uri	https://acikbilim.yok.gov.tr/handle/20.500.12812/435257
dc.description.abstract	Bu çalışmanın amacı, farklı yetenek kestirimi yöntemleri, madde seçim yöntemleri ve test sonlandırma kurallarını dikkate alarak bireylerin yabancı dil yeteneklerinin telafi-edici modellere dayalı Çok Boyutlu Bilgisayar Ortamında Bireyselleştirilmiş (BOB) Testi yöntemleri ile ölçülmesi ve madde-içi ve maddeler-arası boyutluluğun çok-boyutlu BOB testi yöntemlerinin performansları üzerindeki etkisinin incelenmesidir. Bu amaç doğrultusunda, Hacettepe Üniversitesi tarafından uygulanan dinleme, okuduğunu anlama ve dilbilgisi olmak üzere üç boyuttan oluşan İngilizce Yeterlik Sınavlarına (İYS) ilişkin gerçek veri seti kullanılarak gerçek verilere dayalı simülasyon (post-hoc simulation) yapılmıştır.Bu çalışmada, 2009-2013 eğitim-öğretim yıllarında uygulanan 10 İngilizce yeterli sınavına ait veri seti kullanılmış ve her bir testte yer alan maddelere ait madde parametreleri telafi-edici (compensatory) çok boyutlu 2 parametreli lojistik model (CM-2PLM) kullanılarak kestirilmiştir. Madde-içi boyutluluk modeline ait madde havuzu 565 maddeden oluşurken, maddeler-arası boyutluluk modeline ait madde havuzu ise 559 maddeden oluşmaktadır. Bu çalışmada en uygun çok-boyutlu BOB testine karar vermek için iki farklı yetenek kestirim yöntemi (Fisher'in puanlama ve Bayesyen MAP yöntemi), üç farklı madde seçim yöntemi (A-optimality, D-optimality, Seçkisiz madde seçim yöntemi) ve iki farklı test sonlandırma kuralı (sabit madde sayısı ve hata varyansı durdurma kuralı) kullanılmıştır. Toplamda 72 koşul analiz edilmiş ve her bir koşula ilişkin analiz sonuçları güvenirlik katsayıları, ölçmenin standart hatası, ortalama madde sayısı, gerçek ve kestirilen yetenek parametreleri arasındaki korelasyon ve RMSD değerleri açısından karşılaştırılmıştır.Madde düzeyinde boyutluluk modellerine dayalı çok boyutlu BOB testi analiz sonuçlarına bakıldığında, farklı madde seçim ve yetenek kestirim yöntemlerinin kullanımının standart hata, testin uzunluğu, gerçek ve kestirilen yetenek parametreleri arasındaki korelasyon ve RMSD değerlerini etkilediği bulgusuna ulaşılmıştır. D-optimality madde seçim yöntemi yerine A-optimality madde seçim yöntemi kullanıldığında her bir boyutluluk modeli için hem test uzunluğunun ve RMSD değerlerinin azaldığı hem de her bir boyuta ilişkin testin güvenirliğinin arttığı bulgusuna ulaşılmıştır. Diğer taraftan, madde seçim yöntemlerinden D-optimality ve yetenek kestirim yöntemlerinden MLE'ye dayalı Fisher'in puanlama yönteminin madde düzeyinde boyutluluk modellerinden etkilendiği görülmektedir. Gerçek verilere dayalı (post-hoc) simülasyon analizi bulgularına göre kağıt-kalem testleri ile karşılaştırıldığında çok boyutlu BOB testlerinin daha az madde ile daha yüksek güvenirlikte ölçümler yaptığı görülmektedir. Sonuç olarak, A-optimality madde seçim ve Bayesyen MAP yetenek kestirim yöntemlerinin kullanıldığı madde-içi boyutluluk modeline dayalı çok boyutlu BOB testlerinin diğer çok boyutlu BOB testlerine göre daha güvenilir ve tutarlı sonuç verdiği söylenebilir. Bu çalışmanın sonuçları İYS sınavının gerçek çok-boyutlu BOB testi yöntemleri ile uygulanmasında önemli bir katkı sağlayabilir.
dc.description.abstract	The purpose of this study is to measure students' language abilities with Compensatory Multidimensional Computerized Adaptive Testing (MCAT) designs using different ability estimation, item selection methods and stopping rules; and to examine the effect of item-level dimensionality models on MCAT. For this purpose, real data set from English Proficiency Test (EPT) administered by Hacettepe University was used to conduct post-hoc simulation, in which each test consist of three dimensions listening, reading and grammar, respectively. In this study, 10 EPT data sets administered between 2009 and 2013, were used to conduct analysis. Item parameters were estimated with compensatory multidimensional 2 parameter logistic model (CM-2PLM) and item pool for with-in item dimensionality model consisted of 565 items, while item pool for between item dimensionality consisted of 559 items. In order to determine the best MCAT algorithm for EPT, two different theta estimation (Fisher scoring and Bayesian MAP) methods, three different fisher information based item selection methods (A-optimality, D-optimality and Random) and two different termination methods (fixed number of item, precision based) were used. In total, 72 different conditions were taken into consideration, and results of these conditions were compared with respect to, reliability index, SEM, averaged number of items administered and RMSD values between full bank theta and estimated MCAT theta. MCAT Results indicated that using different theta estimation and item selection methods affected SEM, averaged number of administered items, correlation between true and estimated theta and RMSD values. Using A-optimality rather than D-optimality to select items both decreased average number of items administered, RMSD values and increased test reliability for both dimensionality models. On the other hand, both D-optimality item selection and MLE-based Fisher's scoring methods were affected from item-level dimensionality methods. Results also indicated that post-hoc MCAT simulation for EPT provided ability estimations with higher reliability and fewer items compared to paper and pencil format. Overall, MCAT designs based on within-item models with A-optimality and Bayesian theta estimation method outperformed other MCAT designs. Results of this study would also provide an important guideline for live MCAT application of EPT.	en_US
dc.language	Turkish
dc.language.iso	tr
dc.rights	info:eu-repo/semantics/openAccess
dc.rights	Attribution 4.0 United States	tr_TR
dc.rights.uri	https://creativecommons.org/licenses/by/4.0/
dc.subject	Eğitim ve Öğretim	tr_TR
dc.subject	Education and Training	en_US
dc.subject	İstatistik	tr_TR
dc.subject	Statistics	en_US
dc.title	Madde düzeyinde boyutluluk modellerinin bilgisayar ortamında bireyselleştirilmiş test yöntemleri üzerindeki etkisinin incelenmesi
dc.title.alternative	Examining the effects of item level dimensionality models on multidimensional computerized adaptive testing methods
dc.type	doctoralThesis
dc.date.updated	2018-08-06
dc.contributor.department	Eğitim Bilimleri Anabilim Dalı
dc.subject.ytm	Matter
dc.subject.ytm	Item statistics
dc.subject.ytm	Ability tests
dc.subject.ytm	Ability estimates
dc.subject.ytm	Computerized adaptive testing
dc.subject.ytm	Test methods
dc.subject.ytm	Measurement and evaluation
dc.subject.ytm	Individual tests
dc.subject.ytm	Tests
dc.identifier.yokid	10084914
dc.publisher.institute	Eğitim Bilimleri Enstitüsü
dc.publisher.university	HACETTEPE ÜNİVERSİTESİ
dc.identifier.thesisid	418187
dc.description.pages	152
dc.publisher.discipline	Eğitimde Ölçme ve Değerlendirme Bilim Dalı

Files in this item

Name:: yokAcikBilim_10084914.pdf
Size:: 3.551Mb
Format:: PDF
Description:: File_10084914

View/Open

This item appears in the following Collection(s)

TEZLER

Show simple item record

Except where otherwise noted, this item's license is described as info:eu-repo/semantics/openAccess