← all papers · overview

Humans Vs Large Language Models: Judgmental Forecasting In An Era Of Advanced AI

Abstract

This study investigates the forecasting accuracy of human experts versus Large Language Models (LLMs) in the retail sector, particularly during standard and promotional sales periods. Utilizing a controlled experimental setup with 123 human forecasters and five LLMs, including ChatGPT4, ChatGPT3.5, Bard, Bing, and Llama2, we evaluated forecasting precision through Mean Absolute Percentage Error. O

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).