techwiki
LinkedIn·Tuesday, 25 August 2026·1d ago

We cut a client's AI bill from $100k to $60k/mo Without losing any quality: When you use AI you pay on the number of tokens in your…

Assem Chammah
CEO @ Nexus | AI transformation for enterprises | Clients inc. Orange Group & Lambda
We cut a client's AI bill from $100k to $60k/mo Without losing any quality: When you use AI you pay on the number of tokens in your prompts, and the number of output tokens in the response So Opus 4.8 costs $5 per million input tokens, $25 per million output tokens But, benchmark performance on Opus 4.8 is similar to Sonnet 5. And Sonnet currently costs $2 per million input tokens, $10 per million output tokens And for the workflows this client is running, they can get the same level of intelligence required for a much cheaper cost by switching model Two quick caveats: - Sonnet 5 consumes slightly more tokens than 4.8 for the same job - Sonnet 5 is on introductory pricing until August 31. From September it's $3 in, $15 out Pretty much any organization can run this exercise and dramatically cut their AI bill
3
View on LinkedIn

Cross-referenced

Related on the wire