2 comments
OutOfHere8 days ago
The biggest continuing limitation I see is that the cached input has to be at least 1024 tokens. This is terrible. It means a lot of good prefixes that are smaller will go uncached for no good reason. The threshold should have been 128.
Read the full thread on Hacker News →
Related stories
- Better prompt caching for GPT-6openai.comHacker News · 1 points · 8 days ago
- OpenAI says planned GPT-6.1 is too insecure to releasearstechnica.comArs Technica · 0 points · 1 day ago
- Hacker News · 4 points · 6 days ago
- Hacker News · 2 points · 7 days ago
- Please add prompt caching to Jev-style modelsemschwartz.meHacker News · 1 points · 1 day ago
- GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligenceartificialanalysis.aiHacker News · 47 points · about 18 hours ago