You can see a few interesting things, mostly that time per byte actually decreases dramatically with larger payloads which is a good sign (see screenshot below)
You can also see that at the very end of the largest benchmark we get weird outliers which might be alloc related? more to investigate