Hi @jay, the primary use-case is primarily ETL. Read the jsons, transform them and append to an Iceberg table.
Daft can be either run on single node or multiple nodes using Ray, to process the files. In both the cases, if there is any crash of node or cluster, how to resume processing from the point of failure ?