Cognifest NYC 2017: Yuri Bogomolov, Distributed Time Series Analysis: Before and after Spark SQL
ai.bythebay.io Nov 2025, Oakland, full-stack AI conference Scale By the Bay 2019 is held on November 13-15 in sunny Oakland, California, on the shores of Lake Merritt: https://scale.bythebay.io. Join us! ----- Distributed Time Series Analysis: Before and after Spark SQL The first part of the talk will be focused on Flint: an open source time series analysis library on Spark. We will discuss the pros and cons of creating your own RDD types, and dive into performance aspects. Then we will talk about the new feature of Spark 2.2: Spark session extensions and it’s value for developers of third-party libraries on Spark. Bio: Yuri Bogomolov is a Software Engineer at Two Sigma. He is one of key contributors to Flint, and focuses on building scalable and high performance tools with Apache Spark. Prior to joining Two Sigma, Yuri did big data analysis at Amazon.com (http://amazon.com/) and Yandex. Yuri is specialized in algorithms and distributed systems, also he has won several awards in international programming competitions (ACM ICPC, IEEEXtreme).