I have data partitioned by day stored in S3, i.e. customer/year=2020/month=04/day=05, and I have a crawler cataloging that data. Data arrives daily. Is there an option in Glue to update the customer table in that example? For instance, let's say that new customers are discovered on day=06, then, it got added to the table, but let's say that existing customers have updated fields, then, is there an option to only update the table? Or is it a new record to the table?
Currently, when configuring a crawler to discover partitioned data, the partition fields get added to the record. I guess what I'd like to know is if it's possible to constantly have a table representing the latest state of the data?
Thanks in advance. K