home.social

#deltalake — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #deltalake, aggregated by home.social.

  1. Azure costs rising? It may not be compute. Learn how fixing small files, snapshots, and streaming writes cut data costs by 40%. hackernoon.com/how-i-cut-our-c #deltalake

  2. Azure costs rising? It may not be compute. Learn how fixing small files, snapshots, and streaming writes cut data costs by 40%. hackernoon.com/how-i-cut-our-c #deltalake

  3. The next is on 19 March 2026 at Elisa, Ratavartijankatu 5.

    On the menu: optimisation problems in electricity markets, a deep dive into read performance without a cluster, two lightning talks, and an intro to Data & AI at our sponsor Elisa.

    And of course our already traditional quiz, where the main challenge is understanding how answers are scored.

    meetup.com/pydatahelsinki/even

  4. The next #PyData #Helsinki #meetup is on 19 March 2026 at Elisa, Ratavartijankatu 5.

    On the menu: optimisation problems in electricity markets, a deep dive into #DeltaLake read performance without a cluster, two lightning talks, and an intro to Data & AI at our sponsor Elisa.

    And of course our already traditional quiz, where the main challenge is understanding how answers are scored.

    meetup.com/pydatahelsinki/even

    #DataScience #DataEngineering #Python

  5. My most frequently asked question in 2025 was: how do you design for
    high-throughput ingestion with #deltalake

    Over the last week I took the time to write down exactly how I design for throughput in this blog post: buoyantdata.com/blog/2026-01-0

  6. My most frequently asked question in 2025 was: how do you design for
    high-throughput ingestion with #deltalake

    Over the last week I took the time to write down exactly how I design for throughput in this blog post: buoyantdata.com/blog/2026-01-0

  7. Spark SQL for Data Engineering 1 : I am going to start spark sql sessions as series. #sparksql

    Spark SQL Part 1 : I am going to start spark sql sessions as series. #sparksql #deltalake #pyspark ' Databricks Notebooks code for ... source

    quadexcel.com/wp/spark-sql-for

  8. Spark SQL for Data Engineering 1 : I am going to start spark sql sessions as series. #sparksql

    Spark SQL Part 1 : I am going to start spark sql sessions as series. #sparksql #deltalake #pyspark ' Databricks Notebooks code for ... source

    quadexcel.com/wp/spark-sql-for

  9. kafka-delta-ingest was the project that spawned the development of #deltalake for #rustlang, also known as delta-rs.

    Last week I decommissioned the last of those processes. I have since made our
    ingestion even cheaper but kafka-delta-ingest will always hold a spot in our history

    brokenco.de/2025/10/30/kafka-d

  10. Tomorrow at 7am PT (14:00 UTC) I'll be doing some #deltalake hacking with some other 🦀 folks on #deltalive!

    twitch.tv/agentdero/schedule

  11. @Schneems I don't know about best practices, but once upon a while ago I wrote why we re-export some symbols in #deltalake brokenco.de/2023/07/26/rust-re

  12. Did you know you can purchase #deltalake the definitive guide from reputable non-monopolistic book sellers?

    bookshop.org/p/books/delta-lak

    You can also pester some of the other authors and I next week in Mountain View, CA: lu.ma/okxq0bt1

  13. #ITByte: #DeltaLake is an open-source #Storage framework that enables building a #Lakehouse #Architecture with different compute engines.

    Delta Lake preserves the integrity of the original data without sacrificing the performance and agility required for real-time analytics, artificial intelligence (AI), and machine learning (ML) applications.

    knowledgezone.co.in/trends/exp

  14. #ITByte: #DeltaLake is an open-source #Storage framework that enables building a #Lakehouse #Architecture with different compute engines.

    Delta Lake preserves the integrity of the original data without sacrificing the performance and agility required for real-time analytics, artificial intelligence (AI), and machine learning (ML) applications.

    knowledgezone.co.in/trends/exp

  15. Just caught up with the recent Delta Lake webinar,

    > Revolutionizing Delta Lake workflows on AWS Lambda with Polars, DuckDB, Daft & Rust

    Some interesting hints there regarding lightweight processing of big-ish data. Easy to relate to any other framework instead of Lambda, e.g. #ApacheAirflow tasks

    youtu.be/BR9oFD0QMAs

    #dataengineering #datascience #duckdb #daft #polars #pandas #python #spark #deltalake #databricks #airflow #bigdata #smalldata

  16. Just caught up with the recent Delta Lake webinar,

    > Revolutionizing Delta Lake workflows on AWS Lambda with Polars, DuckDB, Daft & Rust

    Some interesting hints there regarding lightweight processing of big-ish data. Easy to relate to any other framework instead of Lambda, e.g. #ApacheAirflow tasks

    youtu.be/BR9oFD0QMAs

    #dataengineering #datascience #duckdb #daft #polars #pandas #python #spark #deltalake #databricks #airflow #bigdata #smalldata

  17. Revolutionizing Data Management: Mooncake-Labs Unveils pg_mooncake v0.1.0

    In a significant leap for data processing, Mooncake-Labs has launched pg_mooncake v0.1.0, introducing robust features that enhance PostgreSQL's capabilities. This release empowers users to seamlessly ...

    news.lavx.hu/article/revolutio

    #news #tech #PostgreSQL #DataManagement #DeltaLake

  18. I hope you'll be able to join the #deltalake AMA today at 9am PST with the authors of Delta Lake The Definitive Guide 🤗

    linkedin.com/events/asktheauth

  19. I am so so excited to have a physical copy of #deltalake "The Definitive Guide"

  20. I had some talks earlier this year about #deltalake and #rustlang

    I also did so with one hand...because I broke my wrist at the beginning of the conference 🤦

    brokenco.de/2024/10/17/data-ai

  21. If you use #deltalake from Python or Rust, I sure would appreciate it if your company sponsored my work!

    github.com/sponsors/rtyler

  22. The #deltalake Python and Rust bindings have change data capture (CDC) read and write support in the works right now.

    *Very* exciting stuff happening between now and Data and AI Summit in June!

  23. Finally hitting a groove with this #deltalake code at ... almost 23:00 local time 🤢

  24. #ITByte: #DeltaLake is an open-source #Storage framework that enables building a #Lakehouse #Architecture with different compute engines.

    Delta Lake preserves the integrity of the original data without sacrificing the performance and agility required for real-time analytics, artificial intelligence (AI), and machine learning (ML) applications.

    knowledgezone.co.in/trends/exp

  25. #ITByte: #DeltaLake is an open-source #Storage framework that enables building a #Lakehouse #Architecture with different compute engines.

    Delta Lake preserves the integrity of the original data without sacrificing the performance and agility required for real-time analytics, artificial intelligence (AI), and machine learning (ML) applications.

    knowledgezone.co.in/trends/exp

  26. I am thrilled that I am able to hire another full time Senior Data Engineer in North America for my organization at Scribd.

    At a high level my team:

    * created the #deltalake Rust bindings, kafka-delta-ingest, and a number of other Rust based data tools.
    * has presented multiple times at Data and AI Summit
    * works at the leading edge of the Databricks platform
    * is remote-first, spanning 3-5 continents depending on the time of year 😆 (we write a lot of stuff down!)

    jobs.lever.co/scribd/80fcabda-