OpenAI updates prompt caching for GPT-6 API

Shared prompt prefixes reused within a 30-minute window now qualify for input token discounts of up to 90%.

Developers building multi-turn agents on GPT-6 can lower API costs and latency while dynamically modifying session configurations.

OpenAI updates prompt caching for GPT-6 API

Sources

Read this as text

Back to the AI news