acoustic-van-30942
12/08/2023, 5:22 PMforeach on Metaflow? A user wants to process ~2M files, and they're already doing some batching and then each batch will take advantage of parallel_map. So wondering what would be a good balance between the horizontal fan out and the batch size?victorious-lawyer-58417
12/09/2023, 1:11 AMvictorious-lawyer-58417
12/09/2023, 1:13 AMvictorious-lawyer-58417
12/09/2023, 1:18 AM--max-workers is likely to be much lower than 6000, so there isn't much downside of doing larger tasks probablyacoustic-van-30942
12/09/2023, 2:18 AMvictorious-lawyer-58417
12/09/2023, 2:19 AMacoustic-van-30942
12/09/2023, 2:20 AMacoustic-van-30942
12/15/2023, 10:20 PMforeach children? I think Thomas can circumvent this issue by decrease the batch size, and he's switched to high memory instance types, but we probably still want to have a reasonable batch size to process the 10M files faster. Any suggestions or recommendations on things to consider?acoustic-van-30942
12/15/2023, 10:22 PM@batch decorator. But besides this,
Would any of the following params in the @batch decorator: sharedMemorySize, maxSwap and swappiness
help to alleviate OOMs as well?