Impact of modifying value of batch_size_fail_threshold_in_kb in cassandra.yaml

Viewed 1126
1 Answers

I would not change that setting without a very good reason. The value of batch_size_fail_threshold_in_kb exists to protect the coordinator node from crashing, in the event that an extremely large batch is sent. Batch statements utilized in a RDBMS fashion (sending thousands of batched writes to the same table) are known to be problematic, and this setting helps to protect against them.

I have had application teams approach me about increasing this value (when using batches correctly), because their payload columns exceed this threshold. I take those on a case-by-case basis.

I cannot speak to how DataStax decided on their (much higher) default. My guess, is that it might have something to do with their Solr or Graph integrations.

tl;dr;

batch_size_fail_threshold_in_kb is one setting where the default should almost never need to be adjusted.

Related