Live data from Hacker News

What AI did to stackoverflow in a graph

data.stackexchange.com

551–560 of 615 posts

Re: What AI did to stackoverflow in a graph

#551

Earlier quoted context omitted.

> Sometimes a new question was in fact a duplicate and should be closed as such. But in the quest to close duplicates I pretty frequently had to argue with the reviewer that "No, this isn't a duplicate just because these two questions related to the same library". Please show examples. People arguing that XYZ is not a duplicate made up a large fraction of volume on the meta site, and in the overwhelming majority of c…

Christ, this is actually the SO problem in a nutshell. I'm not going to go dig through SO to refind the examples of improperly "closed as duplicate" questions I stumbled on years ago while looking up a problem. It's just not that important to me for a dead site. I get it, that means "just trust me bro" is in play. Feel free to completely ignore my comment in that case. You win. SO was filled with this sort of "techni…

> It's just not that important to me for a dead site. I get it, that means "just trust me bro" is in play.

See, this is the part where I'm actually getting insulted by the invocation of the "bro" language. No; what's happening here is that my lived experience as a curator of the site (which was quite stressful) is being denied, and it is supposed that evidence to counter me is not necessary.

This isn't about "winning"; this is about me wanting you to understand something that you apparently don't understand. Again, it's hurtful to have that motivation ascribed to me.

> SO was filled with this sort of "technically this is a duplicate and you are just a nasty rule breaker" style comments with litigation that ultimately goes nowhere.… Others experienced and are reporting here and elsewhere experiencing the harsh moderation of SO.

First off, I want to make this extremely clear: posting a question that gets closed is not, and was never, considered "breaking the rules". When people close a question, they are not moderating the site, but curating it; which is exactly why the overwhelming majority of such actions are taken by non-moderators. People are wrong and ignorant to call them mods, and it's hurtful that they don't take the time to understand this. I understand perfectly well that people felt that these actions were harsh, or even cruel. Many of them were not executed properly. But nobody ever takes the time to understand my perspective; they assume many false and hurtful things about me (and others) and refuse to be corrected on those assumptions. Worse, the site software, and the company, actively interfered with the intended curation of the site in many different ways over the years (changing according to whatever the company was currently trying in order to be profitable).

There is nothing "technical" about things being marked as duplicates. There is a clear standard for it which was reasoned out according to a very specific purpose: people who search the site should find the best possible version of the Q&A, so that they can directly get the answer to their question.

The only time that "litigation goes nowhere" is when people come to the meta site with the assumption that the rules do and ought to work the way they expected them to, and refuse to accept that they actually work differently. Stack Overflow was (and still is, to the extent that anyone remains) a community, and communities are supposed to have a right to self-determination. It's not the fault of the curators that the site presented a misleading interface, to the extent that that happened; they had no power to modify the interface beyond making more Q&A posts on the meta site.

Re: What AI did to stackoverflow in a graph

#552

Earlier quoted context omitted.

> As someone who came into the industry in college, the problem with SO was simply that it was too hard to ask a question. That's because you were intended to use the site like GP describes, and not by asking simply because you want to know something. > They were up your ass about minutia that really didn't matter. What you consider "minutia" were critically important, because the entire point was to optimize for GP'…

and how did that work out? the site is dead now because of it

The site was always going to die. That is completely okay. Things are allowed to exist without chasing a profit motive. Communities are allowed to form on for-profit platforms, and they should not be required to have the platform's commercial interest in mind.

Re: What AI did to stackoverflow in a graph

#553
post #536

Earlier quoted context omitted.

Would you be willing to show questions you attempted to ask? I'd be happy to help explain how policy works/worked there.

I love the implication that thousands of people, from a broad spectrum, who all had the same negative experiences with what was obviously a serious endemic problem with the site, all just needed to understand the policy better. That was clearly the issue.

If thousands of Americans travelled to a tiny country where cars drive on the left, and all tried to drive on the right as they do at home, would that make the country in the wrong? Would it be a "serious endemic problem" with the country?

Re: What AI did to stackoverflow in a graph

#554

Earlier quoted context omitted.

You could not unilaterally edit anonymously; your edit would have been put in a review queue. It is good that leaving comments is hard. First off, because it was learned repeatedly, the hard way, that removing that barrier leads to ungodly amounts of literal spam. Second because even insightful comments detract from the main intended flow of using the site, which is: you find a question from a search engine, read the…

See, what you're saying makes some logical sense in a vacuum but it's so utterly unlike how any other website works that unless you take a huge amount of effort to explain this to people, gently , then it's not at all surprising that people just bounce off of it.

> but it's so utterly unlike how any other website works that unless you take a huge amount of effort to explain this to people, gently, then it's not at all surprising that people just bounce off of it.

It is, and that's fine.

It's incompatible with a site becoming large; and that, too, is fine.

A large site was not required to accomplish any of Atwood or Spolsky's goals aside from perhaps making a profit off of ads. Many other choices were also counterproductive to the goals aside from profit.

Everyone involved would have loved to "explain this to people gently". There was no space provided in which to do it. We had no access to modify the help pages or do anything about user onboarding flows. Downvotes were the tool provided to de-emphasize things and take them out of the way, and to rate quality; so they were used appropriately. (Despite this, and despite public perception, voting was historically very biased in the positive direction.)

Because new questions were open by default, they had to be closed explicitly and in public in order for them to be dealt with. People walked away with the completely incorrect impression that closing a question is inherently a rejection; the explicit purpose and sole effect of closing a question is to prevent answers unless and until it is fixed. It is done with an explicit expectation that fixable questions will get fixed; but the nature of some problems (in particular: missing information) is such that only the OP can do this. There was a period of time in which "closed" was rebranded as "on hold" in the interface, to try to impress upon people that no, they aren't being turned away, they are being asked to take further action. But this was reverted because it was somehow apparently making things even worse.

But also, yes, after telling countless people the same things daily, sometimes humans lose their ability to be gentle. But nobody who's dealt with roughly ever gets to see what the other side had to deal with earlier. The way that people talk about "toxicity" on Stack Overflow comes across like they think it was constantly full of insults and slurs. The evidence shows that the large majority of that sort of content (which explicitly violated the Code of Conduct, and there was a full system in place to deal with this, that pundits often ignore or even lie about) was directed at the curators. One of the most common questions on the meta site is "why aren't people required to explain downvotes?", with the main retort being "why don't you think they should be required to explain upvotes?". But the main answer to that question is because people who stuck out their necks to try to explain the site's standards politely, would frequently get sworn at and even threatened. (I was once openly called sexist for applying these standards, with no evidence beyond the fact that the OP was a woman, which I didn't know at the time).

Re: What AI did to stackoverflow in a graph

#555

Earlier quoted context omitted.

It is taken very seriously in Stack Overflow editing policy that edits are not supposed to go against the author's apparent intent. But it does involve removing a lot of things that you might want the answer to say, but don't meet the guidelines for what answers are supposed to contain. When you join the site, you agree to Terms and Conditions that, among other things, grant a Creative Commons license to the communit…

That's ridiculous and you're wrong. I grant them a CC license to my content to fold, spindle and mutilate it for their own purposes, but not to leave my name attached to the folded, spindled, mutilated version. My answer has my name on it . It's right there saying "kstrauser said these words". If I didn't say them, I don't want the site lying and saying that I did. I don't mind if someone fixes an obvious typo, or up…

> That's ridiculous and you're wrong. I grant them a CC license to my content to fold, spindle and mutilate it for their own purposes, but not to leave my name attached to the folded, spindled, mutilated version.

The choice of license is made clear at https://stackoverflow.com/legal/terms-of-service/public#lice..., and the consequences elaborated at https://meta.stackexchange.com/questions/347758/creative-com... . The license itself is at https://creativecommons.org/licenses/by-sa/4.0/ . It is one of the BY licenses, so the site is legally required to leave your name on the work. The interface explicitly marks both the original author and the most recent editor on the main facing page; and it offers the ability to click in and see the entire editing history.

> But I've had people add extra sentences or paragraphs, and oh hell no.

This is overwhelmingly done in good faith, and it's explicitly policy not to change your meaning. In fact, there are many posts on meta where people complained about attempting a good-faith edit and having been judged as falling afoul of that for no reason. See e.g. https://meta.stackoverflow.com/questions/349950 .) But it's also explicitly policy that they can "edit your words, attributed to you" (see e.g. https://meta.stackexchange.com/questions/403176 ). The people who make these changes, generally, sincerely believe that they are either saying things you intended to say, or at least adding information that you would have added if you'd come back to keep things up to date. There has been a ton of discussion trying to clarify and/or figure out where the line is: see e.g. https://meta.stackoverflow.com/search?q=%5Bediting%5D+intent... .

If you saw someone change something that fell afoul of this, there was a meta site where you could raise the objection. But we were constantly inundated with people who thought it was wrong that we took "thank you in advance, my_username" off the end of the their questions or similar (see e.g. https://meta.stackoverflow.com/questions/260776 ), as if they thought they had a right to that.

Re: What AI did to stackoverflow in a graph

#556
post #511

Earlier quoted context omitted.

It is taken very seriously in Stack Overflow editing policy that edits are not supposed to go against the author's apparent intent. But it does involve removing a lot of things that you might want the answer to say, but don't meet the guidelines for what answers are supposed to contain. When you join the site, you agree to Terms and Conditions that, among other things, grant a Creative Commons license to the communit…

Found the SO mod…

No; editing a post to meet the site's style expectations is not moderation, and I was not a moderator, just like 99.9% of everyone described anonymously and incorrectly as "a mod" in discussions of Stack Overflow on the Internet.

It's really annoying to see people insist that they should be the ones who get to decide how Stack Overflow either did, or should have, worked, when they can't even be bothered to understand who the mods were.

Re: What AI did to stackoverflow in a graph

#557

Earlier quoted context omitted.

> Their rules, (I believe unintentionally) give iron-fisted fiefdom rulers a toolbox of justifications to control and alienate under the guise of protecting the quality of the site data. That isn't what happens. I know many, many people believe it to be what happens, but I know from years of seeing the process on the inside that it's absolutely not what happens in the overwhelming majority of cases. The "quality of t…

> That isn't what happens. I know many, many people believe it to be what happens, but I know from years of seeing the process on the inside that it's absolutely not what happens in the overwhelming majority of cases. If quite literally every person I interact with professionally has an anecdote about this happening, your anecdote about it not happening is not very convincing. Have you/SO staff/SO mods considered why…

Your claim is that "quite literally every person you interact with professionally has an anecdote about" the internal mental state of other people.

Even if everyone else on the Internet thinks a person has a certain belief or motivation or ideology, that doesn't mean that person actually has that belief or motivation or ideology.

I'm speaking to my own mental state, as well as to my general impression of what other curators told me, and the frustrations they espoused. There were chat rooms where people could have easily spent all day talking about how cool it is to power-trip, if they'd actually been power-trippers. Instead, they talked about how helpless* they felt as new people kept coming in and making the same innocent mistakes; about how they had no tools to deal with this except the ones that left offsite users with a bad impression; about how annoying it was to be constantly aware of that offsite bad impression when they were trying to do things necessary to the site's objective; etc.

What you also seem to be missing is that since the Prosus acquisition, "SO mods" (and the rest of the community, who are much greater in number and take the large majority of the actions incorrectly deemed "moderation" by outsiders) and "SO staff" were in active conflict most of the time. The LLM-content ban policy was against the staff's (really the company, who employ a handful of staff as liaisons) wishes; the fighting over this got so bad that there was an actual moderation strike and an extended period where the site got the amount of spam you'd expect for a site that size instead of the almost miraculously small amount it would normally get. (The spam filtering also involved community-operated automated systems that were suspended in solidarity.) Then we got hit with proposal after proposal for AI-powered features that were obviously terrible half-baked nonsense with no real value ad; we constantly tore into them on the meta site; the company showed zero sign whatsoever of caring in the slightest, while Prashanth Chandrasekhar (the new CEO) would give TedX talks and such demonstrating a complete ignorance of how the community felt about AI in general.

Re: What AI did to stackoverflow in a graph

#558

Earlier quoted context omitted.

> Their rules, (I believe unintentionally) give iron-fisted fiefdom rulers a toolbox of justifications to control and alienate under the guise of protecting the quality of the site data. That isn't what happens. I know many, many people believe it to be what happens, but I know from years of seeing the process on the inside that it's absolutely not what happens in the overwhelming majority of cases. The "quality of t…

Mod vs non-mod-user-with-edit-privileges is a semantic difference from a user perspective. Nobody gives a shit what the internal labeling system looks like or hierarchy among the people with edit privileges. I’m not going to litigate my case in front of a clique of other officious hall monitors just for the privilege of making that site better. I don’t care if it was against the rules for me to be annoyed by their ob…

> Mod vs non-mod-user-with-edit-privileges is a semantic difference from a user perspective.

If your lack of interest in how the existing community works is that profound, then you could hardly be said to have a good-faith interest in joining the community in the first place. Which means you're actually upset that the site doesn't provide the service you expected it to.

Sites are not required to provide services simply because the people who arrive there believe that's the service the site provides.

If you're sticking around to keep answering questions, and it's because you think the place has a cool concept behind it and you want to become part of something, you owe it to yourself to make sure you understand what the "something" is. And you owe it to the other people there to accept that you are only one voice out of N in deciding that.

> and they failed to provide a reasonable forum to do that despite being its sole purpose

That is emphatically not what the purpose was. See for example https://meta.stackexchange.com/questions/92107 .

The point system was a broken mess. But it was not intended to represent what you appear to mean by "expertise"; it was intended to represent trust and engagement.

By answering questions, you obtained reputation which granted you privileges, which you were expected to use to maintain the site according to the site's objectives. This includes, notably, filtering questions according to what helps build the site as a useful reference for users of search engines, which is the real motivating force behind the standard question-closure reasons (even if it took a while to achieve consensus on that, and to decide how to phrase it all). And to refrain from answering questions which should be closed, because that acts counter to the site's objective.

The design was specifically intended to ensure that you get scrutinized by peers or superiors — in the domain of using the site properly and working towards its goal, not in the domain of the subject matter where you have expertise. I agree that the reputation system utterly failed in effecting this.

Re: What AI did to stackoverflow in a graph

#559

Earlier quoted context omitted.

> StackExchange had ridiculously high barriers to participation Only in terms of asking new questions, because this was never intended to be the primary use of the site. The entire point of the model is that one person asking a good question could save many others the effort of asking — and answering. But everyone who made an account had permission right off the bat to propose edits to questions and answers, for exam…

I disagree. Stack Overflow was made to succeed, not fail ... originally. Then the founders left, and meta basically chronicles the echo chamber of mistakes made. It's not the intent: that was to help coders code. Meta was dominated by people who lost the intent, and wanted to build a library.

Atwood and Spolsky were quite explicit, on multiple occasions, about what they wanted the site to be. That vision was not compatible with having a site with millions of users and millions of questions, because the masses simply aren't interested in the kind of interaction that the vision required. They may not have understood this at the time. But it's clear that the early community attempted a lot of gatekeeping, much of it misguided in the details (you want lots of questions that address beginner-level problems, because most people in a field are beginners at any given time for any given field, so that's what's most useful) but directionally correct (most of the time, you don't want the question that a beginner asks, because the beginner is not adept at asking useful questions and lacks the perspective needed for a proper Q that allows for presenting a useful A).

The intent was very much to build a library. It's literally on the tour page at the top:

> Stack Overflow is a question and answer site for professional and enthusiast programmers. It's built and run by you as part of the Stack Exchange network of Q&A sites. With your help, we're working together to build a library of detailed, high-quality answers to every question about programming.

That is not something inserted by people on the meta site. We literally did not have access to make that kind of change. Believe me, there are so many other things we would have fixed if it had been made possible.

Most of the important policy was established around 2013 or 2014 and those are the dates of the most important posts on meta that get brought up again and again (and used as duplicate targets, and cited in arguments). That was also right around when site activity was peaking. Atwood and Spolsky didn't leave until the acquisition in 2019.

Re: What AI did to stackoverflow in a graph

#560
post #549

Earlier quoted context omitted.

> overt deletionism. It wasn't enough to just close the questions, they actually felt the need to delete the content people put effort into writing as well. The site has more than three times as many publicly-visible questions as Wikipedia has articles . And that's with the scope restricted to just programming.

You dare to compare a Wikipedia article (even a short one) to an SE "article," which is, according to SE's own rules, narrowly and up-to-the-point concentrated on an equally narrow formulated question?

Yes, because there's no reason not to. A Wikipedia article (even a long one) is "narrowly and up-to-the-point concentrated" on a single concept. There maybe a lot to say about dogs, but https://en.wikipedia.org/wiki/Dog only covers things common to all of them, and individual dog breeds get their own pages (as does the concept of a "dog breed").

Similarly, SO has good, valid questions about general concepts, which are narrowly scoped about the concept despite the broad applicability of that concept. For example, https://stackoverflow.com/questions/26337003 .

Post reply on HN