Live data from Hacker News

We might have been slower to abandon StackOverflow if it wasn't a toxic hellhole

pcloadletter.dev

141–150 of 225 posts

Re: We might have been slower to abandon StackOverflow if it wasn't a toxic hellhole

#141
post #32

StackOverflow was successful explicitly because of the people/question it excluded . The "toxicity" was the point. It was trying very very hard not to become Yahoo Questions. If you want to hear more about this you should watch Joel Spolsky's talk "The Cultural Anthropology of Stack Exchange": https://www.youtube.com/watch?v=LpGA2fmAHvM The point of StackOverflow was explicitly not to help the question-askers, but to…

Well put. It's maddening that people still don't get it 17 and a half years later. (It's interesting that Spolsky is giving the talk, since to my understanding Atwood was considerably more hard-line about it, and was always the one blogging about the high concept.)

I mostly chalk it up to UI affordances. The most obvious one: the site constantly presents an "Ask Question" button; it gives you a form to type in a question; people come to the site because they have a question, and it goes live to a general audience[1] as soon as it's posted. No amount of emphasis on search is ever going to override that.

Less obvious but much more important is that the community can't actually put information about community norms in front of new users, except by scolding them for mistakes. No matter how polite you are about giving people links to the tour[2] or to policies[3],

Then of course, they wanted the site to actually grow at the start, so we got that terribly conceived reputation system best described as Goodhart's law incarnate[4]. And it was far too successful at that early growth, such that if anyone actually understood the idea properly at the start, they were overwhelmed by new users (including the experts answering questions) and had no chance to instill a site culture. It took until 2012 or so until a significant chunk of the experts were getting frustrated with... all the same things they were historically frustrated with on actual forums; then we got the "What Stack Overflow is Not" incident[5]. A lot of the frustration was misdirected except for a general annoyance at certain stereotypes of typical users. It took until at least 2014, from my assessment of the old meta posts, for a real consensus to start emerging about what makes a good question, and even then there was a lot of confusion[6].

Newer sites like Codidact[7] have a chance to learn from this mess, establishing ideas about what good questions look like, and about site scope, from the start[8].

1. Notwithstanding more recent efforts, like the Staging Ground and now a new "question type" feature (https://meta.stackoverflow.com/questions/435293) which seems to have been recently rolled back in preparation for something bigger (https://meta.stackoverflow.com/questions/437856), and the various attempts to force AI into the process, etc.

2. https://stackoverflow.com/tour

3. Especially things like 'Under what circumstances may I add "urgent" or other similar phrases to my question, in order to obtain faster answers?' (https://meta.stackoverflow.com/questions/326569) and 'Why is "Can someone help me?" not a useful question?' (https://meta.stackoverflow.com/questions/284236). See also https://news.ycombinator.com/item?id=46485817 .

4. https://meta.stackexchange.com/questions/387356/the-stack-ex... ; the anchor is for my own answer but please scroll around and read other points of view.

5. See https://meta.stackexchange.com/questions/137795. Back then I was actively using the site but not active on meta; I pretty well gave up in 2015 for largely unrelated (personal) reasons, then came back in mid 2019, coincidentally shortly before the Monica situation[9].

6. In particular, see 'How much research effort is expected of Stack Overflow users?' (https://meta.stackoverflow.com/questions/261592), originally authored 2013, and especially compare the original answers to newer ones. Notably there were also quite a few deleted answers on this one, for those of you with the reputation to view them. Also see 'How do I ask and answer homework questions?' (https://meta.stackoverflow.com/questions/334822) which largely misses the point: it's not so much about the ethics of someone cheating on homework, but about the question fitting the site model.

7. https://codidact.com , with subdomains for various topics. Notably, "programming" as a topic is not privileged; unlike how the Stack Exchange network started with Stack Overflow which still dominates everything else put together, software.codidact.com is just another section of the site. Full disclosure: I am a moderator for that section.

8. See for example https://software.codidact.com/posts/285035/289176#answer-289... ; https://software.codidact.com/posts/291064 ; https://software.codidact.com/posts/284979 ; https://software.codidact.com/posts/292960 ; https://software.codidact.com/posts/294610 ; https://meta.codidact.com/posts/289910 ; https://meta.codidact.com/posts/290028 ; https://meta.codidact.com/posts/291121/291156#answer-291156 ; https://meta.codidact.com/posts/289687 ; https://meta.codidact.com/posts/289951 ; https://meta.codidact.com/posts/284169. Yes, this is a carefully hand-picked list. I have a fairly clear mental image of one more but was somehow unable to search for it.

9. See https://meta.stackexchange.com/questions/333965 and many others. It's a deep rabbit hole. It's also the triggering incident leading to the creation of Codidact[7].

Re: We might have been slower to abandon StackOverflow if it wasn't a toxic hellhole

#142

There should be a lesson for other communities but, unfortunately, there is none. StackOverflow used to be a fun and welcoming place in 2009. It became a toxic hellhole overtime There are no obvious flows in the original design, and there were no endemic wrongdoings in the governance either. It just rotted slowly like any other community does. And nobody in the world knows how to keep communities from becoming toxic.…

And nobody in the world knows how to keep communities from becoming toxic. Mostly true, but there are exceptions... HN is about as good as I've seen for a publicly accessible forum, but has very active moderation (both from Dang and team and a pretty good vote & flag mechanism). The other good forums I've seen are all private and/or extremely niche and really only findable via word of mouth. And have very active mode…

> HN is about as good as I've seen for a publicly accessible forum, but has very active moderation

I wonder how much of it is because of good moderation versus having a site that deliberately doesn't appeal to the masses. The layout is purely text (Aside from the small "Y" logo in the top). No embedded images or videos. Comment scores (aside from your own) are hidden, usernames aren't emphasized with bold, and there are no profile pictures, so karma farming is even more pointless (no pun intended) than on reddit. There's no visible score to give the dopamine from the knowledge of others seeing your high score.

In other words, a smaller community is easier to keep clean, and HN's design naturally keeps the community small.

Re: We might have been slower to abandon StackOverflow if it wasn't a toxic hellhole

#143
post #55

This may be a controversial take based on the last few years of SO discourse, but the strictness in SO was part of what made it a good resource. For a question asker, it could be really toxic. I've had toxic responses as well. The problem is, there are _a lot_ of bad questions that would pollute that site otherwise. For a case study into what it would look like if it invited all questions, look at many subreddits. I'…

StackOverflow was great because it's not like a support forum or a mailing list. It's more like a repository of knowledge. It's been extremely helpful to me when arriving from Google, and I've gotten a couple useful responses to my own questions either. Awesome resource. Where SO started failing in my opinion is when the "no duplicate questions" rule started to be interpreted as "it's a duplicate if the same or very…

The duplicate flagging is definitely the most common bad moderation I saw on SO, for sure. Good point.

Re: We might have been slower to abandon StackOverflow if it wasn't a toxic hellhole

#144

Earlier quoted context omitted.

This is a prime example of why I prefer HN’s moderation. On SO, even though you’re objectively wrong — this is a link to a completely different site and discussing it from a different angle — your active participation on the site would probably have given you the ability to close this conversation and ruin the discussion for those interested. I’m glad that HN doesn’t allow that, so we can roll our eyes at the obvious…

There is nothing in the submitted article that isn't part of or significantly different than the large discussion already ongoing/passed. The discussion about toxic community etc. The short blog post is basically just another comment on that greater discussion. The discussion which is already over there. No need to split it up or repeat it again days later. It's a duplicate discussion .

I disagree, and that’s what I value here: you can add your tags, and we can opt to follow or ignore them as we see fit. That’s far better than the SO model of a site participant having the ability to shut down the conversation for everyone.

But also, a note: conversations are never finished. People talked about this broad topic yesterday, but I didn’t see it yesterday to be able to weigh in. I’m here today, saw this topic, and started talking about it with the other people who stumbled across it just now. I would be highly annoyed with a friend if I brought up an interesting subject and they replied that they’d already discussed it with someone else over dinner last night so there’s no need to talk about it again. I wasn’t there last night. Even if there was a recording of it, that would be a stale artifact I could interact with, other than to contact last night’s debaters and try to continue on with a subject they’d already finished with.

Re: We might have been slower to abandon StackOverflow if it wasn't a toxic hellhole

#145

Isn't this a sign that LLMs are self-cannibalizing? What data will LLMs be trained on when sites like StackOverflow die?

You are now often debugging problems directly in an LLM context window, providing back console output, error logs, etc. there is an absolute mountain of data to mine from developer chat logs.

Re: We might have been slower to abandon StackOverflow if it wasn't a toxic hellhole

#146
The average answer writer on SO is not actually that toxic: really it's the voters.

The top answer will almost always be an explanation of why the asker is wrong to want to do the thing they want to do. But that's not most answers, or most answerers, just the top ones.

The bottom answer will almost always be an honest attempt to write code that answers the question. Sometimes the code doesn't work, or needs explanation: these problems will be solved a couple answers up from the bottom.

My technique for years has been to click the SO link in search results, then hit End on my keyboard to jump to the bottom of the page. This a little slower than reading a cached LLM answer, but faster than waiting for the LLMs to generate something.

Re: We might have been slower to abandon StackOverflow if it wasn't a toxic hellhole

#147

Earlier quoted context omitted.

So much is removed that it gives a false appearance of consensus and harmony.

AFAIK, not a lot in HN gets outright removed. A decent amount of stuff will get flagged (and thus becomes invisible) especially when it's anywhere near politics. But even in those spaces, few things end up actually being flagged even when the flames are burning hot.

This is not my experience. So many comments are flagged and removed. It’s just popularity.

I think it’s fine they are hidden by default. But unt we can see all removed comments we can’t understand the debate.

Re: We might have been slower to abandon StackOverflow if it wasn't a toxic hellhole

#149
post #79

Earlier quoted context omitted.

> I remember that half the time the top result was someone asking the question I had, only for it to be duped to a different question that didn’t answer the original. This is an entirely different problem than toxicity is it not? Like, if the moderators are bad at their job that seems uniquely different than the moderators were mean to me while doing their job.

It's not an entirely different problem because the main method through which moderators are mean is in closing new questions as dupes. A more positive q&a community might "steel-man" the question and try to find what's different about it, but SO's culture leaned heavily towards essentially telling people "go away, you don't have anything new and worthwhile for us".

1. They overwhelmingly are not moderators, and they are not doing moderation by closing questions. This is curation, and duplicate closures overwhelmingly are done by subject-matter experts: users with a gold badge in one or more of the tags originally applied to the question. The requirement for such a badge is based on answering questions:

> Earn at least 1000 total score for at least 200 non-community wiki answers in the $TAG tag. These users can single-handedly mark $TAG questions as duplicates and reopen them as needed.

So these are definitely not people averse to the idea of answering questions.

2. I can guarantee you that the overwhelming majority of these cases are not people trying to be "mean". Users are actively incentivized against closing duplicates, which has historically led to nowhere near enough duplicate questions being recognized and closed (although there have been many proposals to fix this). Dupe-hammering questions "to be mean" is considered abusive, and suspicion of it is grounds to go to the meta site and discuss the matter.

No, people close these questions because they genuinely believe the question is a duplicate, and genuinely believe they improve the site with this closure. It's important to understand that: a) people who ask a question are not entitled to a personalized answer; b) leaving duplicate questions open actively harms the site by allowing answers to get spread around, making it harder for the next person to find all the good ones; c) the Stack Overflow conception of duplication is not based on just what the OP understands or finds useful, but on what everyone else afterward will find useful.

For example, there are over a thousand duplicate links to https://stackoverflow.com/questions/45621722 , most of which is from my own effort — spending many consecutive days closing dozens of questions a day (and/or redirecting duplicate closures so that everything could point at a "canonical"). Yes, that's a question about how to indent Python code properly. I identified candidates for this from a search query and carefully reviewed each one, verifying the issue and sending other duplicates to more specific canonicals in many cases (such as https://stackoverflow.com/questions/10239668). And put considerable effort into improvements to questions and existing answers, writing my own answer, and adding links and guidelines for other curators so that they can choose more appropriate duplicate targets in some cases. I also looked at a wider search that probably had a fairly high false positive rate, but implies that there could be thousands more that I missed.

3. When your question is closed as a duplicate, you immediately get a link to an answer. You don't even need to wait for someone to write it! It's someone saying "here, I was able to find it for you, thanks perhaps to my familiarity with other people asking it".

4. Stack Overflow users really do "try to find what's different about" the question. It just... doesn't actually matter in a large majority of cases. "I need to do X with a tuple, not a list" — well, you do it the same way. "I need to Y the Xs" — well, it seems like you understand how to Y an X and the real problem is with finding the Xs; here's the existing Q&A about finding Xs; you shouldn't need someone else to explain how to feed that into your Y-the-things loop, or if you do, we can probably find a separate duplicate for that. Things like that happen constantly.

Sometimes a question shows up with multiple duplicates. This almost always falls into two patterns: the user is really asking multiple separate things (due to failing to try to break up a problem into logical steps) and each one is a duplicate; or the question is constantly asked but nobody knows a good version of the question, and gives multiple links to previous attempts out of frustration with constantly seeing it. (The latter is bad; someone is supposed to write the good version and send everything else there. But that typically requires behind-the-scenes coordination. Better would be if the first bad attempt got fixed, but you know.)

5. Closing a question is emphatically not about telling people to go away. The intended message (unless the question is off topic or the OP just made a typo or had a brainfart) is "please stay and fix this". However, it's perfectly reasonable that an explicit attempt to catalog and organize useful information treats redundant indices by pointing them at the same target rather than copies of the target. And questions are indices in the Q&A model.

Re: We might have been slower to abandon StackOverflow if it wasn't a toxic hellhole

#150
post #42
post #32

StackOverflow was successful explicitly because of the people/question it excluded . The "toxicity" was the point. It was trying very very hard not to become Yahoo Questions. If you want to hear more about this you should watch Joel Spolsky's talk "The Cultural Anthropology of Stack Exchange": https://www.youtube.com/watch?v=LpGA2fmAHvM The point of StackOverflow was explicitly not to help the question-askers, but to…

As someone who did predominately use stack overflow through Google search… I remember that half the time the top result was someone asking the question I had, only for it to be duped to a different question that didn’t answer the original. So they failed there.

Yes, sometimes you search and find someone else's attempt to ask something that looks very much like your question, but it's duped to a different question. There are a few common failure modes:

* The originally asked question was very low quality; for example, it might have basically been a code dump and a "what's wrong?" where many things were wrong, one of which is what you were both asking about. Someone else may have decided that something else was the more proximate issue.

* The OP was confused, and didn't really have your question. Or the question title was misleading or clickbaity. These should get deleted, but they tend to get forgotten about for a variety of reasons.

* Sometimes two very different problems are described with all the same keywords, and it takes special effort to disentangle them. Even when the questions are properly separated, and even if every dupe is sent to the correct one of the two options, search engines can get confused. On the flip side, sometimes there are very different valid ways to phrase fundamentally the same question.

My favourite example of the latter: "How can I sort a list, according to where its elements appear in another list?" (https://stackoverflow.com/questions/18016827) is a very different question from "Given parallel lists, how can I sort one while permuting (rearranging) the other in the same way?" (https://stackoverflow.com/questions/9764298). But the latter is fundamentally the same problem as in "Sorting list according to corresponding values from a parallel list" (https://stackoverflow.com/questions/6618515). It's very easy to imagine how someone with one of these problems could find the wrong Q&A with a search engine. And there were a lot of other duplicate questions I found that were directed to the wrong one, and if the site were as active as it was in 2020, I'm sure it would still be happening.

And that's after the effort I (and others) put in to improve the prose and especially the titles, and add reference sections. The original titles for these questions were, respectively: "python sort list based on key sorted list"; "Is it possible to sort two lists(which reference each other) in the exact same way?"; "Sorting list based on values from another list?". No wonder people didn't get what they wanted.

Post reply on HN