A short observation on inference pricing that I think is underappreciated: per-token pricing makes the customer bear the model's verbosity. If two providers serve the same quality at the same per-token rate, but one produces answers 40% longer, the second is meaningfully cheaper and no price sheet will tell you that. The comparable unit is cost per completed task, not cost per token, and almost nobody publishes it because it depends on your workload. Worth measuring yourself on a representative sample before committing. The ranking often inverts.

BitFan
Public Service Atlas for Bittensor