• Building the Backend: Data Solutions that Power Leading Organizations

  • De: Travis Lawrence
  • Podcast

Building the Backend: Data Solutions that Power Leading Organizations

De: Travis Lawrence
  • Resumen

  • Welcome to the Building the Backend Podcast! We’re a data podcast focused on uncovering the data technologies, processes, and patterns that are driving today’s most successful companies. You will hear from data leaders sharing their knowledge and insights with what’s working and what’s not working for them. Our goal is to bring you valuable insights that will save you and your team time when building a modern data architecture in the cloud. Topics will span from big data, AI, ML, governance, visualizations, and best practices for enabling your organization to be data-driven. If you are a chief data officer, data architect, data engineer, data analyst, and those building the backend data solutions then HIT SUBSCRIBE!
    © 2023 Building the Backend: Data Solutions that Power Leading Organizations
    Más Menos
Episodios
  • The Analytics Engine for All Your Data with Justin Borgman @ Starburst
    Mar 15 2022

    In this episode we speak with Justin Borgman, Chairman & CEO at Starburst, which is based on open source Trino (formerly PrestoSQL) and was recently valued at $3.35 billion after securing their series D funding.  In this episode we discuss convergence of DW’s / DL's, why data lakes fail and much much more. 

    Top 3 takeaways

    • The data mesh architecture is gaining adoption more quickly in Europe due to GDPR.
    • There were two main limitations of data lakes when comparing to DW’s, performance and CRUD operations. Performance has been resolved with query engines like Starburst and tools like Apache Iceberg, Apache Hudi and Delta Lake are starting to close the gap with CRUD operations. 
    • The principle of a single source of truth / storing everything in a single DL or DW is not always feasible or possible depending on regulations. Starburst is bridging that gap and enabling data mesh and data fabric architectures. 
    Más Menos
    36 m
  • Transform Your Object Storage Into a Git-like Repository With Paul Singman @ LakeFS
    Mar 1 2022

    In this episode we speak with Paul Singman Developer Advocate at Treeverse / LakeFS. LakeFS is an open source project  that allows you to transform your object storage into a Git-like repository. 

    Top 3 takeaways

    • LakeFS enables use cases like debugging to quickly view historical versions of your data at a specific point in time and running ML experiments over the same set of data with branching..
    • The current data landscape is very fragmented with many tools available.. Over the coming years there will most likely be consolidation of tools that are more open and integrated. 
    • Data quality and observability continue to be key components of successful data lakes and having visibility into job runs. 
    Más Menos
    27 m
  • Enable Faster Data Processing and Access with Apache Arrow with Matt Topol @ Factset
    Feb 1 2022

    In this episode we speak with Matt Topol, Vice President, Principal Software Architect @ FactSet and dive deep into how they are taking advantage of Apache Arrow for faster processing and data access. 

    Below are the top 3 value bombs:

    • Apache Arrow is an open-source in-memory columnar format that creates a standard way to share and process data structures.
    • Apache Arrow Flight eliminates serialization and deserialization which enables faster access to query results compared to traditional JDBC and ODBC interfaces.
    • Don’t put all your eggs in one basket, whether you're using commercial products or open source, make sure you design a modular architecture that does not tie you down to any one piece of technology.
    Más Menos
    49 m

Lo que los oyentes dicen sobre Building the Backend: Data Solutions that Power Leading Organizations

Calificaciones medias de los clientes

Reseñas - Selecciona las pestañas a continuación para cambiar el origen de las reseñas.