Video wird geladen...
Video konnte nicht geladen werden
John: why aren't LLMs caching answers to common questions to increase response time? "I've asked ChatGPT, ‘When was OpenAI founded?’ three different times. It's the exact same query." "It doesn't need to light the GPUs on fire for that question. So cache those results and give them to the... show more
301,136 Aufrufe • vor 5 Monaten •via X (Twitter)
0 Kommentare
Keine Kommentare verfügbar
Kommentare vom Original-Post werden hier angezeigt
