Live data from Hacker News

FastSpark: A New Fast Native Implementation of Spark from Scratch

medium.com

1–10 of 93 posts

Re: FastSpark: A New Fast Native Implementation of Spark from Scratch

#10

Nice, but I can't find any reason to choose Spark over modern Distributed SQL databases (CockroachDB, CitusDB, TiDB etc. or cloud vendor-specific SQL DBs)

depends what you’re doing. For querying large datasets? 100% with you.

For data cleaning, processing, analytics, ML on decently large datasets? Spark wins out

Post reply on HN