We’ve improved prompt caching in the API for GPT-6, helping agents run faster and cost…
This is a AI post classified by Jev as AI infra & evals (a launch), kept by the AI Radar because it carries real work, not commentary.
We’ve improved prompt caching in the API for GPT-6, helping agents run faster and cost less. Higher cache-hit rates by default mean more input tokens benefit from cached-input discounts of up to 90%.
Posted by OpenAI Developers (431.7k followers) 59 min ago · 154 likes · 7.5k views · view the original post on X. Kept by the AI Radar as AI infra & evals.
More AI work like this
- This was so much fun. — @jvr0x
- What makes an AI factory energy efficient? — @NVIDIAAIInfra
- $QCOM UNVEILS NEW 2NM SNAPDRAGON CHIPS WITH BIG ON-DEVICE AI PUSH — @wallstengine
- GPT-6 Sol and GPT-6 Luna from OpenAI (@OpenAI) are now available on Merge Gateway. — @merge_api
- LangSmith now supports decision models including Jev and SemIf, giving you visibility… — @LangChain
- What the DPU Is in the Agentic Era ? we give the… — @zartbotF
- We just pushed a big update to Ramp AI Index. — @arakharazian
- simultaneous live detection of: — @yoheinakajima
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 43.2k posts from 5k X accounts over the last 14 days, 1.8k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-22 21:33 UTC. Full method.