My Hadoop MapReduce custom InputFormat for splitting the input performs some additional work that I want to know about when the job is finished. Essentially I need to know some metrics for the number of certain operations my InputFormat implementation performed.
What's the best way to pass additional information out of InputFormat back to the MapReduce job? If InputFormat were passed a Job instance, I could just update counters; unfortunately the JobContext Hadoop passes (I'm using v2.10.x) doesn't provide access to counters.
Should I store the information in the configuration, which I can access via JobContext, and allow the job to access it later? That seems like a bit of a kludge.