Authors
Loading...
Computational analysis reveals that routine text-preprocessing choices severely distort zero-shot language model outputs in automated pipelines, highlighting systemic risks for official statistics.
Komendantova et al. (2026) studied this question.