Hacker News·5 min read·hard
$0.09 and $290.12 are both the price of 1M output tokens
O
OsamaJaber✦AI Summary
An analysis of AI inference costs reveals a massive price disparity between different providers, driven largely by hardware choices rather than the models themselves. The author argues that the accelerator market is inefficiently priced and that self-hosting on specific hardware like AMD's MI355X can be significantly cheaper than using hosted APIs.
Nine cents. Two hundred and ninety dollars and twelve cents.
technologybusiness
✦
Get the full story
Sign up for Headlinne to unlock AI insights, political bias analysis, and your personalized news feed.
Create free accountAlready have an account? Sign in