Scale by the Bay
Scale by the Bay is a By the Bay technology conference.
Talks & recordings
19 connected sessions
Aging: Evolving Software, Tools, and Ourselves
If it ain't broke, should you fix it? If it is broke, shouldn't you just rewrite it instead? Evolving software over years (and decades) isn't glamorous or exciting, but it's the necessary work that enables our field to create real economic value. How should we think about – and plan for – the ongoing evolution of our software? How can we know when to evolve and when to push back? When should we or shouldn't we throw it away and go green-field?
Apache Iceberg: enabling an open lakehouse architecture for large-scale analytics
Data Lakes have been built with a desire to democratize data - to allow more and more people, tools, and applications to make use of data. A key capability needed to achieve it is hiding the complexity of underlying data structures and physical data storage from users. The de-facto standard has been the Hive table format, released by Facebook, which addresses some of these problems, but falls short on data, user, and application scale. Apache Iceberg is a foundational technology for implementing an open data lakehouse, an architecture that addresses the limitations of traditional data architecture patterns. These limitations include having to ETL the data into each tool creating data drift and data silos, high costs making it cost prohibitive to make warehouse features available to all of your data and lack of flexibility forcing you to adjust your workflow to the tool your data is locked in. Apache Iceberg provides the capabilities, performance, scalability and savings that fulfil
Connections
19 relationships