Stackoverflow's data is cc-by-sa 4.0 licensed and can be found here: https://archive.org/details/stackexchange So nothing is stopping anyone to spin-up multiple instances.
It's much harder to work with the data if you don't have the code. SO is not open source.
Stackoverflow's data is cc-by-sa 4.0 licensed and can be found here: https://archive.org/details/stackexchange So nothing is stopping anyone to spin-up multiple instances.
It's much harder to work with the data if you don't have the code. SO is not open source.
IMO it would be more interesting to not use their code and build some kind of specialized search service. One that learns to rank results from user metrics and feedback?
They wouldn't find a solution because all the good answers were moderated away into oblivion :D
I’ll undelete any good answer you link to here.
The biggest tragedy is good answers to bad questions. The answer might teach you something you never knew before, but if the question is judged harshly it might get deleted or hidden in search results.
So what is their stack, or where can I find more about it? A totally custom webserver? Someone commented to this effect the yesterday [1], but I couldn't find much in the way of a good writeup, except this answer from 2009 [2]. 1 https://news.ycombinator.com/item?id=32321726 2 https://stackoverflow.com/questions/676326/how-does-stackove...
And this is at least the 4th or 5th time within 10 days that either the website is down or performance is badly degraded. So bad for a website that is basically the go-to resource where IT professionals get their answers. So bad that the whole platform is closed-source, so we can't even spin multiple instances. The main resource used by developers and sysadmins around the world is locked and centralized, and when tha…
> We need crawlers and scrapers to download EVERYTHING out of the SE platforms, and they need to do so on a daily basis. All content on SE must be mirrored across the world, and if they don't want to do it then we'll do the scraping for them. I think that's already the case. Whenever I look up a technical issue, the first result is usually from SO; the next results are usually from websites that copied that first SO…
You're lucky, for me recently its the other way around, spam copies first.
For anyone looking on this from the future, the error page has Content-Type set to text/html but literally returns the text: "The service is unavailable." and a newline character.
And this is at least the 4th or 5th time within 10 days that either the website is down or performance is badly degraded. So bad for a website that is basically the go-to resource where IT professionals get their answers. So bad that the whole platform is closed-source, so we can't even spin multiple instances. The main resource used by developers and sysadmins around the world is locked and centralized, and when tha…
I’m not trying to be funny but it’s interesting how much panic one can read in the parent comment.
We all joke about the dependency on SO. It seems to actually be more true than perhaps we’d like to admit.
So what is their stack, or where can I find more about it? A totally custom webserver? Someone commented to this effect the yesterday [1], but I couldn't find much in the way of a good writeup, except this answer from 2009 [2]. 1 https://news.ycombinator.com/item?id=32321726 2 https://stackoverflow.com/questions/676326/how-does-stackove...