No takes yet. Share an insight, caveat, or question.
Systematic evaluation reveals the importance of prompt templates in LLM judge performance on alignment tasks.
Wei et al. (2024) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: