home.social

#polars — Public Fediverse posts

Live and recent posts from across the Fediverse tagged #polars, aggregated by home.social.

  1. Vous cherchez des idées de lecture pour l'été? Jetez un œil au 24 Heures du jour (avec ma pomme et en exergue "Sophocle, c'est du thriller") :beetjoy: #polars #thriller #samedilecture #suisseromande

  2. Vous cherchez des idées de lecture pour l'été? Jetez un œil au 24 Heures du jour (avec ma pomme et en exergue "Sophocle, c'est du thriller") :beetjoy: #polars #thriller #samedilecture #suisseromande

  3. Vous cherchez des idées de lecture pour l'été? Jetez un œil au 24 Heures du jour (avec ma pomme et en exergue "Sophocle, c'est du thriller") :beetjoy: #polars #thriller #samedilecture #suisseromande

  4. #NVIDIA published a blog post where they present GQE, a GPU-based query engine. Querying data from databases with GPU accelleration is beyond cool, and will certainly optimize the storage requirements for #bigdata due to enabling for more efficient compression algorithms. Here is the blog post:

    Designing GPU-Accelerated Query Engines with NVIDIA GQE

    #bigdata #databases #polars #pandas #datascience

  5. #NVIDIA published a blog post where they present GQE, a GPU-based query engine. Querying data from databases with GPU accelleration is beyond cool, and will certainly optimize the storage requirements for #bigdata due to enabling for more efficient compression algorithms. Here is the blog post:

    Designing GPU-Accelerated Query Engines with NVIDIA GQE

    #bigdata #databases #polars #pandas #datascience

  6. #NVIDIA published a blog post where they present GQE, a GPU-based query engine. Querying data from databases with GPU accelleration is beyond cool, and will certainly optimize the storage requirements for #bigdata due to enabling for more efficient compression algorithms. Here is the blog post:

    Designing GPU-Accelerated Query Engines with NVIDIA GQE

    #bigdata #databases #polars #pandas #datascience

  7. #NVIDIA published a blog post where they present GQE, a GPU-based query engine. Querying data from databases with GPU accelleration is beyond cool, and will certainly optimize the storage requirements for #bigdata due to enabling for more efficient compression algorithms. Here is the blog post:

    Designing GPU-Accelerated Query Engines with NVIDIA GQE

    #bigdata #databases #polars #pandas #datascience

  8. #NVIDIA published a blog post where they present GQE, a GPU-based query engine. Querying data from databases with GPU accelleration is beyond cool, and will certainly optimize the storage requirements for #bigdata due to enabling for more efficient compression algorithms. Here is the blog post:

    Designing GPU-Accelerated Query Engines with NVIDIA GQE

    #bigdata #databases #polars #pandas #datascience

  9. Polars is a lightning fast DataFrame library/in-memory query engine with parallel execution and cache efficiency. And now you can use is with the tidyverse syntax: tidypolars.etiennebacher.com/ #rstats #polars #optimisation

  10. Polars is a lightning fast DataFrame library/in-memory query engine with parallel execution and cache efficiency. And now you can use is with the tidyverse syntax: tidypolars.etiennebacher.com/ #rstats #polars #optimisation

  11. Polars is a lightning fast DataFrame library/in-memory query engine with parallel execution and cache efficiency. And now you can use is with the tidyverse syntax: tidypolars.etiennebacher.com/ #rstats #polars #optimisation

  12. Polars is a lightning fast DataFrame library/in-memory query engine with parallel execution and cache efficiency. And now you can use is with the tidyverse syntax: tidypolars.etiennebacher.com/ #rstats #polars #optimisation

  13. Polars is a lightning fast DataFrame library/in-memory query engine with parallel execution and cache efficiency. And now you can use is with the tidyverse syntax: tidypolars.etiennebacher.com/ #rstats #polars #optimisation

  14. As a big fan of #Polars when doing #ETL pipelines, and processing #data, I am happy to see them having their distributed engine available for #Kubernetes deployments.

    Read the blog post here:
    pola.rs/posts/polars-distribut

    #datascience #python #rust

  15. As a big fan of #Polars when doing #ETL pipelines, and processing #data, I am happy to see them having their distributed engine available for #Kubernetes deployments.

    Read the blog post here:
    pola.rs/posts/polars-distribut

    #datascience #python #rust

  16. As a big fan of #Polars when doing #ETL pipelines, and processing #data, I am happy to see them having their distributed engine available for #Kubernetes deployments.

    Read the blog post here:
    pola.rs/posts/polars-distribut

    #datascience #python #rust

  17. As a big fan of #Polars when doing #ETL pipelines, and processing #data, I am happy to see them having their distributed engine available for #Kubernetes deployments.

    Read the blog post here:
    pola.rs/posts/polars-distribut

    #datascience #python #rust

  18. As a big fan of #Polars when doing #ETL pipelines, and processing #data, I am happy to see them having their distributed engine available for #Kubernetes deployments.

    Read the blog post here:
    pola.rs/posts/polars-distribut

    #datascience #python #rust

  19. So I moved into industry 1.5 months ago, which has meant a proper switch from R :rstats: to Python :python: (I love both). Here are a few observations for statistics-related stuff in this switch (mainly GLMs, statistical inference, contrasts)

    - #polars is really great, I love LazyFrames & streaming millions of rows of parquet files, categorical data, missing data.
    - I don't really like using #statsmodels, the interface is clunky and the formula API is unfinished

    1/n

    #DataScience #Statistics

  20. So I moved into industry 1.5 months ago, which has meant a proper switch from R :rstats: to Python :python: (I love both). Here are a few observations for statistics-related stuff in this switch (mainly GLMs, statistical inference, contrasts)

    - is really great, I love LazyFrames & streaming millions of rows of parquet files, categorical data, missing data.
    - I don't really like using , the interface is clunky and the formula API is unfinished

    1/n

  21. So I moved into industry 1.5 months ago, which has meant a proper switch from R :rstats: to Python :python: (I love both). Here are a few observations for statistics-related stuff in this switch (mainly GLMs, statistical inference, contrasts)

    - #polars is really great, I love LazyFrames & streaming millions of rows of parquet files, categorical data, missing data.
    - I don't really like using #statsmodels, the interface is clunky and the formula API is unfinished

    1/n

    #DataScience #Statistics

  22. So I moved into industry 1.5 months ago, which has meant a proper switch from R :rstats: to Python :python: (I love both). Here are a few observations for statistics-related stuff in this switch (mainly GLMs, statistical inference, contrasts)

    - #polars is really great, I love LazyFrames & streaming millions of rows of parquet files, categorical data, missing data.
    - I don't really like using #statsmodels, the interface is clunky and the formula API is unfinished

    1/n

    #DataScience #Statistics

  23. So I moved into industry 1.5 months ago, which has meant a proper switch from R :rstats: to Python :python: (I love both). Here are a few observations for statistics-related stuff in this switch (mainly GLMs, statistical inference, contrasts)

    - #polars is really great, I love LazyFrames & streaming millions of rows of parquet files, categorical data, missing data.
    - I don't really like using #statsmodels, the interface is clunky and the formula API is unfinished

    1/n

    #DataScience #Statistics

  24. Alltså... Polars LazyFrames är ju min nya bästa kompis. Har haft "lära sig polars" på att-göra-listan i tusen år... får igen gräma sig över att det tagit så länge att sätta tänderna i det.

    #python #polars #datascience

  25. Alltså... Polars LazyFrames är ju min nya bästa kompis. Har haft "lära sig polars" på att-göra-listan i tusen år... får igen gräma sig över att det tagit så länge att sätta tänderna i det.

    #python #polars #datascience

  26. Alltså... Polars LazyFrames är ju min nya bästa kompis. Har haft "lära sig polars" på att-göra-listan i tusen år... får igen gräma sig över att det tagit så länge att sätta tänderna i det.

    #python #polars #datascience

  27. Alltså... Polars LazyFrames är ju min nya bästa kompis. Har haft "lära sig polars" på att-göra-listan i tusen år... får igen gräma sig över att det tagit så länge att sätta tänderna i det.

    #python #polars #datascience

  28. Alltså... Polars LazyFrames är ju min nya bästa kompis. Har haft "lära sig polars" på att-göra-listan i tusen år... får igen gräma sig över att det tagit så länge att sätta tänderna i det.

    #python #polars #datascience

  29. @thealexmerced thanks! Added to wish list in manning. Better 2buy there vs Amazon to get the ai features?

    I guess Manning got rid of old option to buy coins 2 read individual pages? was a cool feature 2 bad.

    Thanks for reminder about #datafusion i guess it & #polars have excellent #iceberg support & can be used from #rust

    I was thinking about replacing a #pyspark glue job with a rust #lambda on #aws

    Just found your excellent medium account. Best of luck at your upcoming talk!

  30. @thealexmerced thanks! Added to wish list in manning. Better 2buy there vs Amazon to get the ai features?

    I guess Manning got rid of old option to buy coins 2 read individual pages? was a cool feature 2 bad.

    Thanks for reminder about #datafusion i guess it & #polars have excellent #iceberg support & can be used from #rust

    I was thinking about replacing a #pyspark glue job with a rust #lambda on #aws

    Just found your excellent medium account. Best of luck at your upcoming talk!

  31. @thealexmerced congrats on new #iceberg book! Looks good & good timing with all the new features coming out lately. Do you cover much re: cloud architecture or is that too specific (& which cloud would you choose anyway)?

    saw you wrote #polars book as well. What do you think about using polars from #rust ? should work even better than from #python i would think but wasn’t sure. For career security in #ai age i was thinking about switching my #dataengineering focus to rust — thoughts?

  32. @thealexmerced congrats on new #iceberg book! Looks good & good timing with all the new features coming out lately. Do you cover much re: cloud architecture or is that too specific (& which cloud would you choose anyway)?

    saw you wrote #polars book as well. What do you think about using polars from #rust ? should work even better than from #python i would think but wasn’t sure. For career security in #ai age i was thinking about switching my #dataengineering focus to rust — thoughts?

  33. CW: eBooks

    Oui, les livres sont chers en Suisse. Et oui, l'accès au livre de poche est difficile. Pour celles et ceux qui veulent lire Immaculée connexion à prix doux, il y a le livre électronique [allergiques à Google, passez votre chemin], par exemple sur ce site: play.google.com/store/books/de
    #lectures #romans #polars

  34. CW: eBooks

    Oui, les livres sont chers en Suisse. Et oui, l'accès au livre de poche est difficile. Pour celles et ceux qui veulent lire Immaculée connexion à prix doux, il y a le livre électronique [allergiques à Google, passez votre chemin], par exemple sur ce site: play.google.com/store/books/de
    #lectures #romans #polars

  35. CW: eBooks

    Oui, les livres sont chers en Suisse. Et oui, l'accès au livre de poche est difficile. Pour celles et ceux qui veulent lire Immaculée connexion à prix doux, il y a le livre électronique [allergiques à Google, passez votre chemin], par exemple sur ce site: play.google.com/store/books/de
    #lectures #romans #polars

  36. CW: eBooks

    Oui, les livres sont chers en Suisse. Et oui, l'accès au livre de poche est difficile. Pour celles et ceux qui veulent lire Immaculée connexion à prix doux, il y a le livre électronique [allergiques à Google, passez votre chemin], par exemple sur ce site: play.google.com/store/books/de
    #lectures #romans #polars

  37. CW: eBooks

    Oui, les livres sont chers en Suisse. Et oui, l'accès au livre de poche est difficile. Pour celles et ceux qui veulent lire Immaculée connexion à prix doux, il y a le livre électronique [allergiques à Google, passez votre chemin], par exemple sur ce site: play.google.com/store/books/de
    #lectures #romans #polars

  38. Just updated my small :python: package polarsgrid for tidyverse-style expand_grid functionality in python polars 🐻‍❄️

    pypi.org/project/polarsgrid/

    Version 0.4.0 has unit tests, more efficient row-index computation, better CI.

    Major new user-facing improvement is that you can now enter any iterable as the input, not just lists. So create your grid with range(10) instead of list(range(10)).

    See this preprint for use-case
    arxiv.org/abs/2509.11741

    #python #polars #rstats

  39. Just updated my small :python: package polarsgrid for tidyverse-style expand_grid functionality in python polars 🐻‍❄️

    pypi.org/project/polarsgrid/

    Version 0.4.0 has unit tests, more efficient row-index computation, better CI.

    Major new user-facing improvement is that you can now enter any iterable as the input, not just lists. So create your grid with range(10) instead of list(range(10)).

    See this preprint for use-case
    arxiv.org/abs/2509.11741

  40. Just updated my small :python: package polarsgrid for tidyverse-style expand_grid functionality in python polars 🐻‍❄️

    pypi.org/project/polarsgrid/

    Version 0.4.0 has unit tests, more efficient row-index computation, better CI.

    Major new user-facing improvement is that you can now enter any iterable as the input, not just lists. So create your grid with range(10) instead of list(range(10)).

    See this preprint for use-case
    arxiv.org/abs/2509.11741

    #python #polars #rstats

  41. Just updated my small :python: package polarsgrid for tidyverse-style expand_grid functionality in python polars 🐻‍❄️

    pypi.org/project/polarsgrid/

    Version 0.4.0 has unit tests, more efficient row-index computation, better CI.

    Major new user-facing improvement is that you can now enter any iterable as the input, not just lists. So create your grid with range(10) instead of list(range(10)).

    See this preprint for use-case
    arxiv.org/abs/2509.11741

    #python #polars #rstats

  42. Just updated my small :python: package polarsgrid for tidyverse-style expand_grid functionality in python polars 🐻‍❄️

    pypi.org/project/polarsgrid/

    Version 0.4.0 has unit tests, more efficient row-index computation, better CI.

    Major new user-facing improvement is that you can now enter any iterable as the input, not just lists. So create your grid with range(10) instead of list(range(10)).

    See this preprint for use-case
    arxiv.org/abs/2509.11741

    #python #polars #rstats

  43. RE: mastodon.energy/@catalystcoop/

    It's easy to query the bulk Parquet outputs with #polars or @duckdb or @pandas_dev

    So far we've only applied basic transforms -- real dtypes & NULL values, standardized categorical values, etc. But the upgrade to a modern cloud-native format is a huge improvement over nested zipfiles.

    And a big thanks to the folks at the AWS Open Data Registry for giving us a TB of free public storage so we can publish big data like this conveniently.

    Let us know what you think, and what additional kinds of data cleaning, entity classification, record linkage, or additional aggregated outputs would be useful.

    #OpenData #Energy #FOSS

  44. RE: mastodon.energy/@catalystcoop/

    It's easy to query the bulk Parquet outputs with #polars or @duckdb or @pandas_dev

    So far we've only applied basic transforms -- real dtypes & NULL values, standardized categorical values, etc. But the upgrade to a modern cloud-native format is a huge improvement over nested zipfiles.

    And a big thanks to the folks at the AWS Open Data Registry for giving us a TB of free public storage so we can publish big data like this conveniently.

    Let us know what you think, and what additional kinds of data cleaning, entity classification, record linkage, or additional aggregated outputs would be useful.

    #OpenData #Energy #FOSS

  45. RE: mastodon.energy/@catalystcoop/

    It's easy to query the bulk Parquet outputs with #polars or @duckdb or @pandas_dev

    So far we've only applied basic transforms -- real dtypes & NULL values, standardized categorical values, etc. But the upgrade to a modern cloud-native format is a huge improvement over nested zipfiles.

    And a big thanks to the folks at the AWS Open Data Registry for giving us a TB of free public storage so we can publish big data like this conveniently.

    Let us know what you think, and what additional kinds of data cleaning, entity classification, record linkage, or additional aggregated outputs would be useful.

    #OpenData #Energy #FOSS

  46. RE: mastodon.energy/@catalystcoop/

    It's easy to query the bulk Parquet outputs with #polars or @duckdb or @pandas_dev

    So far we've only applied basic transforms -- real dtypes & NULL values, standardized categorical values, etc. But the upgrade to a modern cloud-native format is a huge improvement over nested zipfiles.

    And a big thanks to the folks at the AWS Open Data Registry for giving us a TB of free public storage so we can publish big data like this conveniently.

    Let us know what you think, and what additional kinds of data cleaning, entity classification, record linkage, or additional aggregated outputs would be useful.

    #OpenData #Energy #FOSS

  47. RE: mastodon.energy/@catalystcoop/

    It's easy to query the bulk Parquet outputs with #polars or @duckdb or @pandas_dev

    So far we've only applied basic transforms -- real dtypes & NULL values, standardized categorical values, etc. But the upgrade to a modern cloud-native format is a huge improvement over nested zipfiles.

    And a big thanks to the folks at the AWS Open Data Registry for giving us a TB of free public storage so we can publish big data like this conveniently.

    Let us know what you think, and what additional kinds of data cleaning, entity classification, record linkage, or additional aggregated outputs would be useful.

    #OpenData #Energy #FOSS

  48. Here is how you can explore massive datasets from the terminal 💯

    🌀 **datui** — A high-performance TUI for analyzing tabular data

    🔥 Query with SQL, render charts, transform data & stream huge files with Polars

    🦀 Written in Rust & built with @ratatui_rs

    ⭐ GitHub: github.com/derekwisong/datui

    #rustlang #ratatui #tui #data #analytics #polars #cli #devtools