news.ycombinator.com

Untitled

minimaxir · 0 points · 0 comments · wczoraj

> Cached input costs just $0.10 per million tokens—95% less than standard input pricing and 50% less than GPT‑6 Sol’s cached input pricing This is the actual big announcement. 50% cheaper cache than GPT-6 Sol will get you far more mileage on Codex.

Comments

5 preview comments · loading full thread
joshstrangewczoraj

> 50% cheaper cache than GPT-6 Sol will get you far more mileage on Codex. Cache doesn't help you much when you are compacting every 5 minutes... I was shocked at how quickly I ran out my $100/mo subscription with a single agent (sol medium).

TuxSHwczoraj

Exactly half as expensive as Opus 5.5 in every API pricing metric

pvab36 godzin temu

gpt 6 Sol was already supposedly better and 50% cheaper than 5.6 Sol right? I didn't understand why they were keeping 5.6 Sol

vcryan7 godzin temu

People's volume and approach varies. I'm a happy customer and I use my entire double max subscription on planning and analysis and have other models doing all my implementation work because I would burn through my subscription in a day or less. It's difficult to calculate, but I'm something like 10-20 billion token per week consumer and I can't use a US-based model to do this volume of implementation work. Also, a lot of this work is verification to ensure that AI generated code does what is intended and is safe to merge and deploy. That verification work is critical and uses a lot of tokens.

verdvermwczoraj

cache is typically 10%, is this OAI setting a new level at half, 5%?