mirror of
https://github.com/HeyPuter/puter.git
synced 2026-10-03 10:27:48 +00:00
* feat(ai): sync GPT-6 Sol, GPT-6 Luna, and Claude Opus 5.5 * chore: remove documentation changes from model sync * fix(ai): bill OpenAI cache writes and long-context pricing GPT-5.6 and later bill prompt-cache writes at 1.25x input and report them in `cache_write_tokens`, inside the input total. Both OpenAI calculators now split them out of the prompt count and meter them under their own key, and the six GPT-5.6+ models carry the rate. GPT-6, GPT-5.6, GPT-5.5 and GPT-5.4 (incl. Pro) bill a request with more than 272K input tokens at 2x input (cached reads and cache writes included) and 1.5x output for the whole request. Models declare this as `long_context_pricing`, and the ledger overrides, the reported `usd_cents` and the credit gate all apply the multipliers. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> * fix(ai): bill GPT-5.6 Sol at OpenAI's promotional pricing The catalog carried GPT-5.5's $5/$0.50/$30, but OpenAI bills GPT-5.6 Sol at $4/$0.40/$20 per million input/cached/output tokens, promotional through at least 2026-11-21. PUT-1943 tracks re-checking before then. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com> --------- Co-authored-by: Claude Opus 5.5 (1M context) <noreply@anthropic.com>