Live data from Hacker News

Rides of Glory – Uber Blog (2012)

web.archive.org

21–30 of 116 posts

Re: Rides of Glory – Uber Blog (2012)

#21
post #18

Off-topic: Because of situations like these, I'm surprised that part of the checklist when launching a PR blog is not: "Block googlearchive/archive.org robots" There have been very, very few times when a company's webpage was down and I needed to go to google-archive or archive.org to refer to some innocuous information. However, the times that I've used those sites to gather evidence of possible whitewashing? Many,…

What is funny is that using robots.txt would have been enough.

Re: Rides of Glory – Uber Blog (2012)

#23

In another deleted post [0] the author talks about using a name-to-gender API to look at ride locations by gender, which implies that these analyses were not done using anonymized data. [0] https://web.archive.org/web/20140827195715/http://blog.uber....

Personally, I think they do it with raw access to the database. https://www.uber.com/legal/usa/privacy under the heading "How do we use the information collected" says nothing about anonymizing rideshare data.

Re: Rides of Glory – Uber Blog (2012)

#24
post #20
post #16

Stuff like this is one of the many reasons I love archive.org. I think i's really important to capture historical artifacts for future analysis. The service they provide doesn't allow the "Ministry of Truth"[1] to doctor historical documents to meet their present day narrative. [1] https://en.wikipedia.org/wiki/Ministry_of_Truth

Sadly, it does. archive.org respect the robots.txt of the current website owner. This can mean that they have the data but choose not to give you access to them. I have seen cases in the past where a website I once frequented became defunct, then the domain expired, then someone parked a holding page on that domain including a robots.txt that keeps archive.org from displaying the old data (which do not even belong to…

Thanks for the info, I didn't know that.

Re: Rides of Glory – Uber Blog (2012)

#25
post #16

Stuff like this is one of the many reasons I love archive.org. I think i's really important to capture historical artifacts for future analysis. The service they provide doesn't allow the "Ministry of Truth"[1] to doctor historical documents to meet their present day narrative. [1] https://en.wikipedia.org/wiki/Ministry_of_Truth

This is a good example of using the power of technology to enable the weak against the powerful.

And yes I don't care what you think, but a company with a billion(ish) of funding is more powerful than YOU.

Re: Rides of Glory – Uber Blog (2012)

#26
I gotta say, I'm not really seeing the creepy / cringey / evil / whatever-else here...

Anyone (especially the HN crowd) should know they have the data, and if you think they're not carefully analyzing it behind the scenes (like every other tech company who has your data), I've got things to sell you. I personally think a tiny peek like this into the data, much like the usage posts that OKCupid, YouPorn, and others give, is neat.

Re: Rides of Glory – Uber Blog (2012)

#27
post #20
post #16

Stuff like this is one of the many reasons I love archive.org. I think i's really important to capture historical artifacts for future analysis. The service they provide doesn't allow the "Ministry of Truth"[1] to doctor historical documents to meet their present day narrative. [1] https://en.wikipedia.org/wiki/Ministry_of_Truth

Sadly, it does. archive.org respect the robots.txt of the current website owner. This can mean that they have the data but choose not to give you access to them. I have seen cases in the past where a website I once frequented became defunct, then the domain expired, then someone parked a holding page on that domain including a robots.txt that keeps archive.org from displaying the old data (which do not even belong to…

The Internet Archive should implement some sort of digital signature system to allow website owners with foresight to prevent this.

Re: Rides of Glory – Uber Blog (2012)

#28

Is it sad that in the realm of Uber blunders I find this relatively tame?

Yeah, frankly this doesn't seem all that bad compared to some of the words coming out of executives' mouths.

One could do an analysis like this while still working with anonymized data. Still a bit creepy, but not that different from reports and blog posts you see from other startups and tech companies.

Re: Rides of Glory – Uber Blog (2012)

#29

Uber employees are the kind of people who kept a telescope in their bedroom window to peep on girls down the street. I'll never understand why they're still in business.

They provide a convenient, affordable service that people need. I doubt any significant number of their users care or even know about the company's shady and unethical doings.

Re: Rides of Glory – Uber Blog (2012)

#30

In another deleted post [0] the author talks about using a name-to-gender API to look at ride locations by gender, which implies that these analyses were not done using anonymized data. [0] https://web.archive.org/web/20140827195715/http://blog.uber....

You have to start with the original data, which is obviously de-anonymized. Full data -> [gender, time, origin neighborhood, destination neighborhood] leaves you with a pretty anonymous dataset, and is all that would be required for this analysis.

Internal metrics teams nearly always have access to complete data. The issue is sharing non-anonymized data externally.

Post reply on HN