Data by the Bay
Data by the Bay is a By the Bay technology conference.
Talks & recordings
100 connected sessions
A Real Time Analytics Framework
We'll talk through the elements of a framework which was built to easily construct multiple real time analytics applications to address the needs of multiple teams at Yahoo's Publishing Products group. In this talk you'll learn about the architecture of what makes up a real time analytics stack and its use cases. We'll cover the supporting services and libraries that had to be built to support the various use cases. We'll also cover the optimizations we made to address resource and network I/O considerations. Our focus will mainly be on the data processing not the analytics.
Akka Streams for Large Scale Data Proces...
With over 50 million members, Credit Karma is the most utilized and trusted personal finance platform in the U.S. To handle tens of millions of Americans’ credit information, we use Akka Streams for high throughput data transfer. We will discuss how we quickly built a solution using Akka Actors to help us parallelize, parse and send data to our ingestion service. We then used Akka Streams to pull data through our actors based on demand, allowing us to easily control our memory buffers and prevent Out of Memory issues. Akka Streams also allowed us to apply parsing and simple business logic. In this panel, Credit Karma shares best practices on how to implement Akka Streams at scale.
Apache Flink: A Very Quick Guide
Apache Flink is the exciting newcomer in the Big Data space, which integrates batch and stream processing in a novel way. Join us in exploring Apache Flink and its architecture, internals, and its data and programming models. We point out the unique features and differences in comparison to Apache Spark. You will see Flink’s Scala APIs in action as we code and run a selected set of examples that nicely illustrate its features.
Connections
Top 40 of 150 ranked relationships