PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 21, 2026Informatics in Education0 citationsOpen Access

Large Language Models for Educational Task Authoring: A Bebras Challenge Case Study

LBLeonard Busuttil

Key Points

  • This research aims to explore how large language models can create educational tasks for the Bebras Challenge.
  • Single-case study approach utilizing large language models
  • Exemplar-based prompting with authentic Bebras tasks
  • Evaluation by international expert reviewers based on established quality criteria
  • Collaborative revision by international editors to address gaps in accessibility and technical accuracy
  • Generated task accepted for inclusion in the 2025 Bebras challenge
  • Alignment with pedagogical requirements confirmed by expert reviewers
  • Identification of gaps in accessibility compliance and technical framing
  • Collaborative workflow revealed to enhance LLM-generated content

Abstract

This study explores the application of large language models (LLMs) to create computational thinking tasks for the Bebras International Challenge through a single-case study approach. Using exemplar-based prompting with seven authentic Bebras tasks from the 2024 cycle as contextual input, a task was developed that was subsequently accepted for inclusion in the 2025 international Bebras challenge. Comparison with the exemplar tasks confirmed that the generated content drew from multiple sources rather than replicating any single task, combining grid-based constraint satisfaction, rule-based filtering, and logical deduction into a novel navigation puzzle with engaging narrative context. International expert reviewers evaluated the task using established Bebras quality criteria, confirming successful alignment with core pedagogical requirements including age-appropriateness, clarity, and cultural neutrality. However, two significant gaps emerged in the broader authoring workflow: accessibility compliance in the researcher-authored visual components and technical inaccuracies in the LLM-generated informatics framing. Following collaborative revision by international editors that addressed these concerns while preserving the LLM’s creative contributions, the task achieved acceptance for international use. The findings reveal a collaborative pipeline comprising contextual preparation, LLM-guided generation, human technical implementation, expert community review, and collaborative revision. Results from this case suggest that LLMs can efficiently generate educationally sound creative foundations while requiring integrated human expertise to meet specialised standards and ensure inclusive design, with the task’s acceptance providing encouraging evidence for the viability of this collaborative approach.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Leonard Busuttil (2026) studied this question.

synapsesocial.com/papers/69be37866e48c4981c6773bbhttps://doi.org/10.15388/infedu.2509.022
Ask AI
Helpful
Bookmark
Share
View Full Paper