Live data from Hacker News

Why we should care about the Nate Silver vs. Nassim Taleb Twitter war

towardsdatascience.com

21–30 of 212 posts

Re: Why we should care about the Nate Silver vs. Nassim Taleb Twitter war

#21
The article seems like an oversimplification of the underlying dispute.

- It’s not necessary or even useful to define an arbitrary “decision boundary” unless you actually have to make a decision. From a Bayesian perspective, a probability stands for itself: 100% means the event is certain to occur, 50% means you have no idea, and numbers in between convey varying levels of certainty. In reality, 538’s predictions are not true Bayesian probabilities because they don’t take epistemic uncertainty into account, but that’s a totally different issue.

- “Wild fluctuations in a prediction” from new information are absolutely a “normal part of forecasting” - sometimes. If I’m planning to flip two coins, the probability of getting two heads is 25%; but once I flip the first coin, the probability changes to either 50% (if I get heads) or 0% (if I get tails). In the case of Comey reopening the investigation, even if the model had included a probability of that happening, it would have been low and thus wouldn’t affect the overall forecast much. But once that low probability became a certainty, you would expect a sudden swing. The real question is whether 538’s predictions are more swingy than they logically should be (particularly earlier on), but again, that’s a different issue.

Re: Why we should care about the Nate Silver vs. Nassim Taleb Twitter war

#22
post #6

Earlier quoted context omitted.

Taleb actually makes a great point here and published a pretty cool paper about it. The point is that if a probability of the event changes too much, you can arbitrage it. E.g. assume these are payoff odds and you can sell your position before the event. He then found some nice no-arbitrage conditions that "real" probabilities must meet and showed that Nate's predictive timeline allowed arbitrage. Unfortunately Nate…

But 538 was giving predictions for different events each day in the moving "now cast" so there was no arbitrage.

The forecast still fluctuated way too much to be correct, as can be seen in Taleb's paper, which cites the "forecast" and not the "now cast."

Re: Why we should care about the Nate Silver vs. Nassim Taleb Twitter war

#23
> Practically what this means is that they do not predict a winner or looser but instead report a likelihood

Yes, this is good. It means that you can evaluate their accuracy to a greater degree than % of times correct.

> Further complicating the issue, these predictions are reported as point estimates

They've learned from this, and now show their distribution of results very clearly. However, there's nothing mathematically fraught about reporting a prediction as a point estimate.

> The problem is that models are not perfect replicas of the real world and are, as a matter of fact, always wrong in some way.

...yes? And so what? This is a fully general counterargument to all of science.

> Predictions have two types of uncertainty; aleatory and epistemic.

Don't say things like this with 100% certainty when your source itself says "the validity of this categorization is open to debate", and also when your source is Wikipedia.

> However, as you can see, there is still a noticable variation of 2–5% of actual proportion to predictions. This is a signal of un-addressed epistemic uncertainty. It also means you cannot take one of these forecast probabilities at face value.

Models aren't perfect. Also, 538 is very careful to address their epistemic uncertainty! They also try very hard to not change their models significantly after they publish them, in order to avoid letting their personal biases tinker with the results.

This post goes out of it's way a to avoid mentioning that there are pretty well established ways of measuring the accuracy of predictions, instead using things like "look, it's not quite a line" and "look, the line went up and down before the election".

If you're interested in actually reading about this from a mathematical viewpoint, read Taleb's paper, which has some actually interesting thoughts on why 538's algorithms are too eager, or read Madeka's paper, which includes comparisons of 538 to other people trying to make predictions, and finds that actually, they were better than most other news sources.

https://arxiv.org/pdf/1703.06351.pdf

https://arxiv.org/pdf/1704.02664.pdf

Re: Why we should care about the Nate Silver vs. Nassim Taleb Twitter war

#24
Author does informative critique of 538. I'm not sure at all that it has anything to do with Silver vs Taleb twitter war though. It's more like premise for expressing author's own thoughts.

However I totally dig what 538 does and appreciate their analysis. I understand that there is a complex underlying model with many parameters taken subjectively. The results of Monte Carlo runs of this model are very interesting to me. The fact that they describe distribution of outcomes and not just a single number is important property of a Monte Carlo model and I would have an issue if that wouldn't provide that info. If I would like to analyze complex event that has lots of moving parts and uncertainties I would also build a model to see what kind of distribution I would get. The fact that somebody published results of their model is quite useful.

The analogy would be European and American weather models - no one say that their results are exact and everybody understand that uncertainty in result comes from inherent uncertainty of initial state as well as shortcuts and approximations each of the model takes. No one says that the results of these models are useless because of that. And everybody finds it valuable if weather forecast predicts rain tomorrow with 40% chance. So I don't see how political prediction is only valuable (by the words of the author) if it comes without probabilities attached to it.

Re: Why we should care about the Nate Silver vs. Nassim Taleb Twitter war

#25
This is an intellectually confused post. It is so unclear and muddled that only a charitable reader would suppose the author understands the mathematical and modeling issues here. I’m surprised the author published it in this form - you don’t want to jump into a controversy like this with such an unclear discussion.

As just one example, the whole digression on a “decision boundary” is conceptually mistaken. Once you report a posterior probability, it’s up to the user to establish a simple real number threshold for placing a bet (if we’re talking about means to put skin in the game). The posterior you just learned holds all information, your only action is to threshold it. That’s the Neyman-Pearson lemma.

If you allow generalizations to interval-valued probability, which might be sensible under these conditions, the situation gets more complicated. But the writer of the post did not mention this.

Another place this came out is his initial discussion of aleatory vs. epistemic uncertainty — this can be done clearly, but here we read:

> Aleatory uncertainty is concerned with the fundamental system (probability of rolling a six on a standard die). Epistemic uncertainty is concerned with the uncertainty of the system (how many sides does a die have? And what is the probability of rolling a six?).

This is just not helpful. He has said “probability of rolling a six” twice.

Source: do UQ in day job.

Re: Why we should care about the Nate Silver vs. Nassim Taleb Twitter war

#26
post #21

The article seems like an oversimplification of the underlying dispute. - It’s not necessary or even useful to define an arbitrary “decision boundary” unless you actually have to make a decision. From a Bayesian perspective, a probability stands for itself: 100% means the event is certain to occur, 50% means you have no idea, and numbers in between convey varying levels of certainty. In reality, 538’s predictions are…

Seems like in addition to whether they're more swingy than they should be, are they being presented in a way that is interpreted as more certain or authoritative than they can possibly claim?

Re: Why we should care about the Nate Silver vs. Nassim Taleb Twitter war

#27
Nate Silver built an aleatory model whereby low probability high effect events (Comey reopening the investigation on eve of the election) are discounted because if you add enough of these potential disruptive events the model will be 50%-50% and that has no value to anybody. Who wants to read an article that said "election forecast 50-50" every day.

If a 538 model says one outcome is 75% probable, it really should come with a giant caveat saying "at current trends assuming nothing out of the ordinary occurs". Taleb's beef is that if that is the case, the 75% does not mean in reality this result is 75% likely to happen.

What is a really interesting takeaway is that if in an event where the outcome you desire based appears unlikely, your job is introduce as much previously undefined or discounted uncertainty. Ideally you engineer a black swan event or at least do what you can to make it happen. This would be fun to model from a game-theory approach instead.

Re: Why we should care about the Nate Silver vs. Nassim Taleb Twitter war

#30
The author doesn't seem to understand how the 538 model works. The 538 model does take epistemic uncertainty into account; it just doesn't call it that. The model calls it "likelihood of polls moving by X% between now and election day". In 2016 they offered a "now-cast" which didn't include that, and answer the question "based on where the polls are right now, how likely is any given outcome if the election were held today?" while the "classic" model answered the question "based on where the polls are right now and the fact that unanticipated things will happen between now and the election, how likely is any given outcome?". For 2018, they dropped the "now-cast", because people didn't understand what it was saying.

True, the epistemic uncertainty is limited -- it's based on past elections, because hey, 538 works with actual data. One could argue that rather than saying "71.4% Clinton 28.6% Trump", a better model would have said "71.3% Clinton 28.4% Trump 0.2% Election is cancelled due to nuclear war / natural disaster / etc" but I don't think anyone sensible is interpreting the model as excluding such extreme outcomes.

Post reply on HN