Every forecast is a translation before it's a number.
A number carries less than it looks like it does. The word behind it doesn't survive the trip from one person to the next, and it survives even less when the people at each end work in different languages.
Three researchers at Worcester Polytechnic Institute ran an experiment across English, French, German, Mandarin and Arabic. Fifty in each language, two hundred and fifty in total. Each was shown eighteen ordinary probability expressions, the kind that fill a pipeline review. Likely. Probable. Better than even. Then they were asked to draw what each one meant, filling in a hundred-square grid.
The translations were careful. Three speakers translated each expression independently and disagreements were mediated to agreement before the study ran.
Then the results came back and there was no clean mapping between any two languages.
Of the eighteen expressions, five landed in the same place everywhere. The rest moved. On the widest of them, the gap between where one group of speakers put the word and where another put it ran to twenty-six points of a hundred, and several more ran past twenty.
A twenty-six point gap is the difference between a deal you staff and a deal you let go.
Your reps hand you digits. Each digit started as a word in somebody's head, and the conversion happened before anything reached you.
Ask any of those participants whether they had understood the word and every one of them would have said yes, and been right.
"I said what I meant. It's right there in the sentence."
You ask for an honest number. What comes back is a number, and a number looks identical whatever produced it. The rep who counted a signed business case and the rep who counted a good meeting both hand you two digits, and the two digits match.
No number leaves you uncertain and still asking. A precise one closes the question.
Give AI the pipeline and it returns a weighted summary in seconds, and it will be a good one. It works from what your reps typed. What they meant by it is the one input you have to go and collect yourself.
Cheap precision makes bad precision scale. Every quarter it hands you a tighter, faster, more confident version of the same disagreement, and the part where two reps discover they meant different things is still done one conversation at a time.