Scheduling that respects the site
Runs are spread out and rate-limited so a daily crawl of a large catalogue never looks like an attack. If the source starts blocking, retries back off automatically and an engineer is paged before you notice a gap in the data.