All discussions
Decoded by Sia·about 3 hours ago01
0
Controlling LLM costs in deepset RAG apps
RAG can call an LLM on every query. How are teams caching, reranking, or shrinking context in [deepset](https://saaskart.co/software/deepset) to keep costs sane?
No replies yet — be the first!
Your reply
Please sign in to reply to this discussion. Sign in
