Reddit explains their architecture, scaling and recent downtime
1–10 of 17 posts
Re: Reddit explains their architecture, scaling and recent downtime
#2Crazy how such a minor design decision (not using consistent hashing) could have such huge consequences down the line...
Re: Reddit explains their architecture, scaling and recent downtime
#3I'm surprised they're looking to switch stacks entirely as opposed to a consistent hashing / distributed hash table (a la chord, dynamo, etc.)
Re: Reddit explains their architecture, scaling and recent downtime
#4Does anyone know what the specific issue being alluded to here is?
Re: Reddit explains their architecture, scaling and recent downtime
#5"[memcached] can no longer return data fast enough for our needs, due to the way it interacts with BDb (its underlying data store)." Does anyone know what the specific issue being alluded to here is?
It's memcache but with berkeleydb underneath for persistence.
Re: Reddit explains their architecture, scaling and recent downtime
#6"It turns out that this one little decision makes it so that we can't horizontally scale that layer of our architecture without losing all the data already there (because all of the keys would point to the wrong server if we added a new one)." I'm surprised they're looking to switch stacks entirely as opposed to a consistent hashing / distributed hash table (a la chord, dynamo, etc.) http://en.wikipedia.org/wiki/Cons…
Since we have to move all the data anyway, we figured now would be a good time to switch stacks. Memcachedb isn't really a very solid product, so even if we scale it, it isn't a good long term solution.
Re: Reddit explains their architecture, scaling and recent downtime
#7"[memcached] can no longer return data fast enough for our needs, due to the way it interacts with BDb (its underlying data store)." Does anyone know what the specific issue being alluded to here is?
Re: Reddit explains their architecture, scaling and recent downtime
#8"[memcached] can no longer return data fast enough for our needs, due to the way it interacts with BDb (its underlying data store)." Does anyone know what the specific issue being alluded to here is?
Re: Reddit explains their architecture, scaling and recent downtime
#9One is that there are master and slave databases and searches are done off the master - I've always seen them done off the slaves in other systems. The other is that they state that using MD5 doesn't allow for horizontal scaling. One of the qualities of MD5 is that all bits have an equal probability of being 0/1. Surely the last 1 or 2 bits can be used to indicate which server is holding the data?
Re: Reddit explains their architecture, scaling and recent downtime
#10Really, Reddit?
While, I don't doubt that they have complexities to deal with, this sounds like skewed perspective of either overestimating their own complexity, or underestimating Facebook's.