I can do anything I want 😁
I can do anything I want 😁
"Accelerating Data Reads in Iceberg: Caching and Optimization Strategies"
eyedle.ai/accelerating...
#python #iceberg #dataengineering #pyiceberg
"Accelerating Data Reads in Iceberg: Caching and Optimization Strategies"
eyedle.ai/accelerating...
#python #iceberg #dataengineering #pyiceberg
Betting $100 it is AI slop because the commands return results from the Java API.
Betting $100 it is AI slop because the commands return results from the Java API.
pypi.org/project/pydu... (github: github.com/jghoman/pydu...)
pypi.org/project/pydu... (github: github.com/jghoman/pydu...)
Iceberg tables automatically show up in the iceberg_tables view, which means pyiceberg can connect to it as a (sql) catalog.
www.crunchydata.com/blog/crunchy...
github.com/Eventual-Inc...
Iceberg tables automatically show up in the iceberg_tables view, which means pyiceberg can connect to it as a (sql) catalog.
www.crunchydata.com/blog/crunchy...
#ADLS #opentableformat #PyIceberg.
#ADLS #opentableformat #PyIceberg.
buff.ly/H2SovBh
buff.ly/H2SovBh
load with pyiceberg (local file cache) took: 6.6 sec
*plan files: 3.5sec, project_table 3.1sec
load with polars (local files) took: 270ms
Same partitioning
load with pyiceberg (local file cache) took: 6.6 sec
*plan files: 3.5sec, project_table 3.1sec
load with polars (local files) took: 270ms
Same partitioning
At Re:invent AWS announced Amazon S3 Tables - which are managed Apache Iceberg tables stored/accessed in S3. They are optimized for analytics usages and promise to be faster for these cases. Seeing examples using these helps and here is one using PyIceberg. (1/3)
🧵
At Re:invent AWS announced Amazon S3 Tables - which are managed Apache Iceberg tables stored/accessed in S3. They are optimized for analytics usages and promise to be faster for these cases. Seeing examples using these helps and here is one using PyIceberg. (1/3)
🧵
Dynamic Routing Lightweight ETL with AWS Lambda, DuckDB, and PyIceberg
#aws #dataengineering #duckdb #icebereg
Dynamic Routing Lightweight ETL with AWS Lambda, DuckDB, and PyIceberg
#aws #dataengineering #duckdb #icebereg
PyIceberg on AWS Lambda: Comparing GlueCatalog and REST Catalog Access Methods
#aws #iceberg #dataengineering
PyIceberg on AWS Lambda: Comparing GlueCatalog and REST Catalog Access Methods
#aws #iceberg #dataengineering
1) Manage Iceberg tables and metadata with ease
2) Work seamlessly with existing tools like PyIceberg, Snowflake, and Spark
3) Query data from any cloud or region with zero egress fees
1) Manage Iceberg tables and metadata with ease
2) Work seamlessly with existing tools like PyIceberg, Snowflake, and Spark
3) Query data from any cloud or region with zero egress fees
Accelerate lightweight analytics using PyIceberg with AWS Lambda and an AWS Glue Iceberg REST endpoint
#AWS #BigData
Rounding out our speaker spotlights for #icebergSummit, we have Fokko Driesprong highlighting what you can expect from his #PyIceberg session. Don't miss your chance to dive into this #apacheIceberg implementation tomorrow! 🐍 🧊
Rounding out our speaker spotlights for #icebergSummit, we have Fokko Driesprong highlighting what you can expect from his #PyIceberg session. Don't miss your chance to dive into this #apacheIceberg implementation tomorrow! 🐍 🧊
Join us to explore Daft, a next-gen data engine redefining catalogs in Python!
📅 Apr 7 | 10 AM PT
🔗 Register: lu.ma/BeyondJVMs
#deltalake #oss
Join us to explore Daft, a next-gen data engine redefining catalogs in Python!
📅 Apr 7 | 10 AM PT
🔗 Register: lu.ma/BeyondJVMs
#deltalake #oss