

Hahaha — wait, do you think that the models take more energy to train than the cumulative amount used to serve them for inference across the lifetime of the model?
And if not, then what you are saying is effectively gibberish as any reasonable assessment would need to divide the training cost across all post-trained inference. So if it was only 50% the lifetime inference usage, it’d be a 1.5x multiplier to that I said above, which doesn’t change either the Netflix 4k or gas mileage points.

The post in question explicitly self identified as AI and is very saturated in the typical tells (it’s almost certainly Claude or Deepseek Flash).
No escape, just a throwaway experiment by whoever was running it to tell them to make money.
Will probably see a lot more of this over the next 6-12mo.