Live data from Hacker News

Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

m.wikimediafoundation.org

21–30 of 192 posts

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#21

Given the terrible state of advertising, I would welcome a search engine that penalizes pages with popovers, animated ads, auto playing audio, and so on. Google would never build this, given its business model. I hope Wikipedia brings some innovation to search, untethered from advertising revenue.

Google does actually penalize this. It's also not that easy to catch given the layers involved.

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#23
post #12

This is exciting. Recently I started to work on an answer engine/search engine. It still sucks but it's a good project to work on when bored http://kairos.xyz/ In a few weeks I'll publish the source code and do a Show HN. I wish a lot of luck to https://www.mojeek.com/ and http://www.lexxe.com/ too. DuckDuckGo also started to crawl the web with its own bot (right now they're using Yandex's api). We need more competit…

(Default nginx page showing up on your website)

It works for me. Care to show me a screenshot?

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#24
post #23

Earlier quoted context omitted.

(Default nginx page showing up on your website)

It works for me. Care to show me a screenshot?

"Welcome to nginx on Debian!

If you see this page, the nginx web server is successfully installed and working on Debian. Further configuration is required.

For online documentation and support please refer to nginx.org

Please use the reportbug tool to report bugs in the nginx package with Debian. However, check existing bug reports before reporting a new bug.

Thank you for using debian and nginx."

I meant here: http://kairos.xyz/

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#25
post #23

Earlier quoted context omitted.

(Default nginx page showing up on your website)

It works for me. Care to show me a screenshot?

When I make a request from another IP, it gives me the right page.

" (...) Try...

Who is Richard Stallman? define bravado 10 USD to EUR RFC 2460 generate username help - about"

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#26
post #11

This is a very, very, very small amount of money if you want to build a search engine, let alone one "to rival Google" (source?). Looks like the goals are realistic, though - look how wikipedia search could be extended beyond results from wikipedia.org, build some test sets. And get a better idea what it really is that is supposed to be built.

I run a knowledge engine project that actively mines facts from web-based sources and third-party data dumps. It was featured on the front page of HN a while ago, and has a total of $10 in funding (from a single donation; not a typo). I have, however, put a ton of time into it, and it's something I'm very passionate about. I'm fairly confident Wikipedia can have success making initial headway on their grant objective…

Over 15 years ago, for my undergraduate thesis I set up a "Hypermedia Textbook" on the history of my field. I had to manually collate all the info, manually scan in every photo, and type in every last bit of text and html. The end result was a couple hundred pages that looked very, very similar to what your Einstein page looks like! At the time, I knew a better way would emerge, but didn't know how or when. It's moments like this that I (a) feel old :( and (b) am amazed by the times we live in and the speed at which things are happening! :) Thank you for providing such a wonderful, if unintentional, moment of self-reflection!

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#27
post #23

Earlier quoted context omitted.

(Default nginx page showing up on your website)

It works for me. Care to show me a screenshot?

I also get the Nginx default page when I click on your link from my iPhone. I'd take a screenshot but not sure how to link that from here.

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#28
post #11

Earlier quoted context omitted.

I run a knowledge engine project that actively mines facts from web-based sources and third-party data dumps. It was featured on the front page of HN a while ago, and has a total of $10 in funding (from a single donation; not a typo). I have, however, put a ton of time into it, and it's something I'm very passionate about. I'm fairly confident Wikipedia can have success making initial headway on their grant objective…

Over 15 years ago, for my undergraduate thesis I set up a "Hypermedia Textbook" on the history of my field. I had to manually collate all the info, manually scan in every photo, and type in every last bit of text and html. The end result was a couple hundred pages that looked very, very similar to what your Einstein page looks like! At the time, I knew a better way would emerge, but didn't know how or when. It's mome…

Thank you for checking it out! And if you have any ideas for how things can be improved, I'd love to hear them :)

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#29
post #23

Earlier quoted context omitted.

It works for me. Care to show me a screenshot?

When I make a request from another IP, it gives me the right page. " (...) Try... Who is Richard Stallman? define bravado 10 USD to EUR RFC 2460 generate username help - about"

This is so strange. If I fetch the page from http://archive.is I get the same default Nginx page

http://archive.is/AxHFV

the only thing I can think off right now is that the http "Host" header field is not sent. I have several sites on the same server and Nginx is used as a reverse proxy and uses the Host field to redirect traffic to different ports.

Re: Wikipedia starts work on $2.5M internet search engine project to rival Google [pdf]

#30

This is a very, very, very small amount of money if you want to build a search engine, let alone one "to rival Google" (source?). Looks like the goals are realistic, though - look how wikipedia search could be extended beyond results from wikipedia.org, build some test sets. And get a better idea what it really is that is supposed to be built.

I think "this is a very, very, very small amount of money if you want to build an encyclopedia" was probably said by a bunch of people in the early Wikipedia days.
Post reply on HN