Live data from Hacker News

Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

m.wikimediafoundation.org

41–50 of 192 posts

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#42
post #35

The title is misleading. It's 250k, not 2.5m and the goal is a knowledge engine, not a search engine.

http://i.imgur.com/w89dQ4i.png sounds like a search engine to me...

It is a search engine called the knowledge engine. That is what I read of it.

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#43
post #39
post #36

Earlier quoted context omitted.

You should really consider using multiple server blocks instead of relying on the Host field.

I do. Nginx does the matching using the Host field http://nginx.org/en/docs/http/ngx_http_core_module.html#serv...

Hmm, I could've sworn that server_name doesn't rely on the Host header. My mistake.

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#44

This is a very, very, very small amount of money if you want to build a search engine, let alone one "to rival Google" (source?). Looks like the goals are realistic, though - look how wikipedia search could be extended beyond results from wikipedia.org, build some test sets. And get a better idea what it really is that is supposed to be built.

It is a small amount of money and the database for it has to be huge to index a lot of websites. So the server will have a large hard drive and be able to serve up enough users to not suffer outages.

I imagine they will build a proof of concept and get more money for it later. Have volunteers work on building it to save money. Open source the project and have others look into fixing the issues with it.

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#45
I support it 100%, but still think they should use search advertising to cover costs and further development instead of asking for donations every year.

especially if they can make something that actually does rival google... other companies have spent billions and not gotten very close.

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#46
post #37

The biggest problem is the lack of data about what people are searching. It's a catch-22 that's very hard to break in the face of Google's search dominance and ubiquity. By Google being the best, it only becomes better, and introduces a huge barrier to entry to competitors. It used to be possible to know what people were searching for to end up in a given Wikipedia article, but the process is now only asynchronous (a…

I think facebook should be able to build a search engine. Don't know why they don't have one yet

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#47
post #46
post #37

The biggest problem is the lack of data about what people are searching. It's a catch-22 that's very hard to break in the face of Google's search dominance and ubiquity. By Google being the best, it only becomes better, and introduces a huge barrier to entry to competitors. It used to be possible to know what people were searching for to end up in a given Wikipedia article, but the process is now only asynchronous (a…

I think facebook should be able to build a search engine. Don't know why they don't have one yet

Facebook wants a single platform to provide the Web through. If they built a search engine, it would only search Facebook.

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#48
post #37

The biggest problem is the lack of data about what people are searching. It's a catch-22 that's very hard to break in the face of Google's search dominance and ubiquity. By Google being the best, it only becomes better, and introduces a huge barrier to entry to competitors. It used to be possible to know what people were searching for to end up in a given Wikipedia article, but the process is now only asynchronous (a…

Keyword Planner tool tells you what people search for.

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#49
post #35

The title is misleading. It's 250k, not 2.5m and the goal is a knowledge engine, not a search engine.

Not only is it a search engine, but it is a grant application that has had WMF staff leaving in droves, and has greatly upset many, many others - who will quite likely also leave.

It's very, very sad. And it's also a shameful moment for the WMF.

edit: and don't just think it's me saying it. The WMF has had a mass exodus of staff in the last week or so. If you speak to any WMF non-executive staff members directly, you'll quickly find out that morale is at an all time low, and confidence in the WMF Board is sitting at something like 12%.

Post reply on HN