news.ycombinator.com
Untitled
> Cached input costs just $0.10 per million tokens—95% less than standard input pricing and 50% less than GPT‑6 Sol’s cached input pricing This is the actual big announcement. 50% cheaper cache than GPT-6 Sol will get you far more mileage on Codex.
评论
5 条预览评论 · 正在加载完整讨论请先登录 h4cker 账号,然后连接 Hacker News 后发表评论。
> 50% cheaper cache than GPT-6 Sol will get you far more mileage on Codex. Cache doesn't help you much when you are compacting every 5 minutes... I was shocked at how quickly I ran out my $100/mo subscription with a single agent (sol medium).
Exactly half as expensive as Opus 5.5 in every API pricing metric
gpt 6 Sol was already supposedly better and 50% cheaper than 5.6 Sol right? I didn't understand why they were keeping 5.6 Sol
People's volume and approach varies. I'm a happy customer and I use my entire double max subscription on planning and analysis and have other models doing all my implementation work because I would burn through my subscription in a day or less. It's difficult to calculate, but I'm something like 10-20 billion token per week consumer and I can't use a US-based model to do this volume of implementation work. Also, a lot of this work is verification to ensure that AI generated code does what is intended and is safe to merge and deploy. That verification work is critical and uses a lot of tokens.
cache is typically 10%, is this OAI setting a new level at half, 5%?