Hacker News
new
top
best
ask
show
job
Better prompt caching for GPT‑6
(
openai.com
)
3 points
by
mehrdadrad
13 hours ago
2 comments
OutOfHere
11 hours ago
The biggest continuing limitation I see is that the cached input has to be at least 1024 tokens. This is terrible. It means a lot of good prefixes that are smaller will go uncached for no good reason. The threshold should have been 128.
OutOfHere
11 hours ago
Guide:
https://developers.openai.com/api/docs/guides/prompt-caching