Structured output constraints affect LLM performance differently. Gemini 2.5 Pro benefits from structured prompting, while GPT-5-Thinking declines. Open-weight models show minimal impact from output constraining. Proprietary models outperform open models in radiology protocol selection.
Bahaaeldin et al. (Fri,) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: