Batch Size and LLM Inference Efficiency
Having an optimal Batch size can decrease your models cost per token at the time of inference.
This will be an explanation on how Batch size affects the cost at the inference, We will be going deep an
swaritshukla.hashnode.dev4 min read