Keywords: Spark Streaming, Cassandra data modeling.
Well-versed in digging through data to find key insights and curating a compelling story from complex analyses, passionate about delving into data from different systems, at different timescales, and in complex formats to uncover hidden relationships.
Machine Learning with Spark: Linear / Logistic Regression, Decision Trees, NaiveBayes, Alternating Least Squares (Recommender Systems), TF-IDF, Frequent Pattern Mining
Professional Background (formerly): ETL Developer / Traditional DWHs / Kimball's Methodology
Computer Science Skills: Data Structures, Algorithms, Functional Programming Paradigm, Relational Databases
Big Data / Core Skill: Spark
Big Data / Core Skill: Apache Cassandra => Data Modeling
Big Data / Other: Apache Kafka => Spark Streaming from Kafka topics
Programming Languages: Scala, Python
Keen interest in experimenting with open-source Big Data technologies.
E-mail address in the profile.