Anywhere from 1M to 4M will yield the best performance, especially if you’re not doing “in-place modification”, which very few non-DB software actually does.
Going higher than 8M can reduce performance gains from losing parallel processing of multiple blocks at a time.
A 100KB file in a dataset with a recordsize of 4M will consume 100K.