I would add that the task is relevant too. I feel there’s not yet a model that is consistently better at everything. I still revert to plain old GPT-4 for direct translation of text into English that requires creative editing to fit a specific style. Of all the Claudes and GPTs, it’s the one that gives me the best output (to my taste). On the other hand, for categorisation tasks, depending on the subject and the desired output, GPT-4o and Claude 3.5 might perform better than the other interchangeably. The same applies to coding tasks. With complex prompts, however, it does seem that Claude 3.5 is better at getting more details right.
In certain airports/flights, where you have to take a bus from the gate to the plane after boarding at the gate counter, 'speedy boarding' can actually work against its intended purpose. This is because you'll end up boarding the bus first, and if you choose to take a seat rather than standing close to the doors, everyone else will leave the bus before you to board the plane. I've seen this happening countless times.