Appending to a tensor

torch.cat is super efficient, and basically bandwidth bound. it’s not time consuming.