Tianji Li
10/04/2024, 4:29 AMjay
10/04/2024, 5:24 AMTianji Li
10/04/2024, 3:53 PMPhil Chen
10/04/2024, 5:12 PMKevin Wang
10/04/2024, 5:27 PMTianji Li
10/04/2024, 5:40 PMKevin Wang
10/04/2024, 5:41 PMjay
10/04/2024, 6:24 PMjay
10/04/2024, 6:25 PMTianji Li
10/04/2024, 6:28 PMKevin Wang
10/04/2024, 6:35 PMjay
10/06/2024, 2:49 AMdaft.read_parquet(..., hive_folder_partitioning=True)
This will then attempt to auto-detect the partitioning info on the folder paths.Phil Chen
10/06/2024, 1:17 PMjay
10/06/2024, 4:15 PMdaft.register_catalog(“my_glue_catalog”, type=“aws_glue”)
df = daft.read_table(“my_glue_catalog.x.y.z”)
We should be able to use the glue API to find all the tables I think, and perhaps we can start with supporting just Hive tables and Iceberg tables.
We’re still iterating on the possible API spec on our table/catalog support, would love to get your inputs once we have some proposals here @Phil Chen !Phil Chen
10/06/2024, 5:50 PM