Live data from Hacker News

Fast Nearest Neighbor Queries in Haskell

izbicki.me

1–10 of 20 posts

Re: Fast Nearest Neighbor Queries in Haskell

#2
"I’d guess that it took about 10 times as much work to create my Haskell-based cover tree than it would have taken to create a similar C-based implementation. (I’m roughly as proficient in each language.)"

Interesting since I would imagine implementing something in C to take much longer than doing the same in something like Python or Java. On the other hand, I suppose getting something to work is not the same as getting it to run fast.

Re: Fast Nearest Neighbor Queries in Haskell

#3
post #2

"I’d guess that it took about 10 times as much work to create my Haskell-based cover tree than it would have taken to create a similar C-based implementation. (I’m roughly as proficient in each language.)" Interesting since I would imagine implementing something in C to take much longer than doing the same in something like Python or Java. On the other hand, I suppose getting something to work is not the same as gett…

He says "Because Haskell code is so high level, it requires aggressive compiler optimizations to perform well."

I think you are basically right. It is easier to make something correct in Haskell, it is easier to make something reasonably fast in Haskell too (generally) but in specific circumstances when you want max performance and care about the assembly or intermediate code output then it becomes easier in a low level language like C to do the performance optimization.

Having said that the 10* effort is an investment if you get a reusable library and will pay off because the rest of your code is in Haskell not C!

Re: Fast Nearest Neighbor Queries in Haskell

#4
post #3
post #2

"I’d guess that it took about 10 times as much work to create my Haskell-based cover tree than it would have taken to create a similar C-based implementation. (I’m roughly as proficient in each language.)" Interesting since I would imagine implementing something in C to take much longer than doing the same in something like Python or Java. On the other hand, I suppose getting something to work is not the same as gett…

He says "Because Haskell code is so high level, it requires aggressive compiler optimizations to perform well." I think you are basically right. It is easier to make something correct in Haskell, it is easier to make something reasonably fast in Haskell too (generally) but in specific circumstances when you want max performance and care about the assembly or intermediate code output then it becomes easier in a low le…

Great point. His concluding statement aligns with your observation:

"So then why did I use Haskell? To make cover trees 10 times easier for programmers to use."

Default Haskell is certainly not geared toward numeric performance even though it certainly can be with the right libraries.

Re: Fast Nearest Neighbor Queries in Haskell

#5
Not that i'm the proshit haskell expert, but I don't think that GHC does the most amazing job in applying optimizations and I definitely think its performance paradigm is really hard to reason about.

Still, it was a nice breakdown of a real non-toy problem in haskell!

Re: Fast Nearest Neighbor Queries in Haskell

#6
Since you're using the auto-tuned index for FLANN it's going to build a large number of indexes to try and find the best performing one, which can take a long time. Properly tuned I would expect FLANN to be much faster (though that's expected since it doesn't do exact searches). Also benchmarking the index build time together with query times doesn't make sense to me since depending on the application you may not care how long it takes to build the index.

Re: Fast Nearest Neighbor Queries in Haskell

#7
post #4
post #3

Earlier quoted context omitted.

He says "Because Haskell code is so high level, it requires aggressive compiler optimizations to perform well." I think you are basically right. It is easier to make something correct in Haskell, it is easier to make something reasonably fast in Haskell too (generally) but in specific circumstances when you want max performance and care about the assembly or intermediate code output then it becomes easier in a low le…

Great point. His concluding statement aligns with your observation: "So then why did I use Haskell? To make cover trees 10 times easier for programmers to use." Default Haskell is certainly not geared toward numeric performance even though it certainly can be with the right libraries.

You might be interested in this if you are interested in high performance numerical Haskell:

https://github.com/wellposed/numerical

Carter claims he can perform optimizations that aren't possible in other languages, though I can't personally speak to them since I'm not well versed in numerical code.

Re: Fast Nearest Neighbor Queries in Haskell

#8
post #5

Not that i'm the proshit haskell expert, but I don't think that GHC does the most amazing job in applying optimizations and I definitely think its performance paradigm is really hard to reason about. Still, it was a nice breakdown of a real non-toy problem in haskell!

I've heard quite the opposite. I've heard that GHC is state of the art in the depth and breadth of optimizations it can perform. I'm struggling to remember if the context was "of everything" or just within pure functional programming.

Re: Fast Nearest Neighbor Queries in Haskell

#10
post #8
post #5

Not that i'm the proshit haskell expert, but I don't think that GHC does the most amazing job in applying optimizations and I definitely think its performance paradigm is really hard to reason about. Still, it was a nice breakdown of a real non-toy problem in haskell!

I've heard quite the opposite. I've heard that GHC is state of the art in the depth and breadth of optimizations it can perform. I'm struggling to remember if the context was "of everything" or just within pure functional programming.

GHC has great optimizations within the scope of modules (and seems to balance for the right level of polymorphism/modularity/linking), and can inline some things across module boundaries. But at the level of entire programs more advanced optimizations are possible; I gather there are tradeoffs too.

http://mlton.org/

https://www.reddit.com/r/haskell/comments/2tpmbo/what_on_ear...

http://stackoverflow.com/questions/4720499/possible-optimiza...

Post reply on HN