GPT-6 Gets Smarter at Remembering—and Cheaper to Use
OpenAI has improved how GPT-6 handles repeated information, making it faster and less expensive. The upgrade means fewer delays when you ask questions about the same documents or data.
GPT-6 Gets Smarter at Remembering—and Cheaper to Use
Imagine telling a friend the same long story twice. It would be nice if they remembered it the second time without you repeating every detail. OpenAI's GPT-6 now does something similar with prompt caching—a technique that lets the AI remember parts of conversations or documents so it doesn't have to reprocess them every single time.
When you use an AI like GPT-6, it "reads" everything you send it each time you ask a question. If you're asking questions about the same long document, contract, or email thread over and over, that's wasteful—like rereading the entire book to find a new answer. Prompt caching lets the AI store those repeated parts in a kind of temporary memory. The improvements announced mean this memory system now works better: it remembers more often (higher "cache hit rates"), shows you what it's doing (new diagnostics), and gives you more control over exactly what gets saved. The result? Faster answers and lower costs for people and companies using GPT-6.
For everyday users, this means fewer seconds waiting for responses when asking follow-up questions. For businesses paying per query, it means smaller bills. Think of it like a library that's gotten much better at quickly finding the book you've already asked about before.
Original source: OpenAI Blog
