İçeriğe geç
akaturk Akademik ölçüm

Makale detayı · 2025

Comparative analysis of AI and expert evaluations in engineering design pedagogy

Dergi

PLOS ONE

ISSN 1932-6203

YÖKSİS OpenAlex Açık erişim · gold SJR Q1 JCR Q2 Atıf 2 Yüzdelik 72.7% FWCI 0.59
Yıl
2025
Tür
article

Veri kaynağı ayrımı

  • YÖKSİS YÖKSİS makale kaydı
  • YÖKSİS dergi adı PLOS One
  • Katalog eşleşmesi (ISSN) PLOS ONE
  • OpenAlex OpenAlex zenginleştirmesi (özet, atıf, konular)

Özet

OpenAlex · İngilizce

BACKGROUND: Integrating engineering design processes into science education has become a significant priority in STEM instruction. However, many science teachers face difficulties incorporating these processes due to limited pedagogical expertise. Generative artificial intelligence (GAI) tools such as ChatGPT offer potential support mechanisms by evaluating lesson plans and providing formative feedback. This study investigates the reliability and validity of GAI evaluations compared to expert assessments. METHODS: This mixed-methods study involved 43 science teachers who received professional development over four months to integrate engineering design into their lesson plans. A total of 52 lesson plans were evaluated using structured and unstructured prompts via ChatGPT 4.5, alongside evaluations by expert mentors. Quantitative data were analyzed using the Intraclass Correlation Coefficient (ICC) and Bland-Altman methods to assess inter-rater consistency. Qualitative data was analyzed through open and deductive coding to interpret differences in evaluation rationale. RESULTS: Findings revealed high consistency between structured prompt AI evaluations and expert assessments (ICC = 0.708), while unstructured prompts showed low and non-significant agreement (ICC = 0.076). Qualitative analysis indicated that AI evaluations, particularly those using structured prompts, tend to be more positive and holistic, whereas experts offered more detailed and critical feedback. Differences were also observed in evaluating dcomponents like problem definition, testability, and interdisciplinary integration. CONCLUSION: Structured AI prompts offer reliable and valid evaluation results comparable to expert assessments and could serve as scalable tools in teacher support systems. However, unstructured prompts produce inconsistent outcomes and require refinement. The study highlights both the potential and limitations of using GAI tools for pedagogical evaluation in STEM education.

Konular

Atıflar

OpenAlex cited_by_count. WoS veya Scopus atıf sayısı değildir; o kaynaklar için ayrı kolon yoktur.

2 atıf

OpenAlex cited_by_count (önbellek / veritabanı)

Yerel katalogda bu makaleye atıf yapan 1 yayın (OpenAlex referans eşleşmesi; tam dünya listesi değildir).

  1. Correction: Comparative analysis of AI and expert evaluations in engineering design pedagogy 2025 Atıf 0 · OpenAlex

Yazarlar

  1. TUĞRA KARADEMİR COŞKUN SİNOP ÜNİVERSİTESİ
  2. ESRA BOZKURT ALTAN