ENTERPRISE AIInvestigateNOW
Reduce RAG costs on Amazon Bedrock with query-aware compression
AWS Machine Learning Blog
Factual evidence
What the source reports
AWS details a query-aware context compression pattern on Bedrock using a smaller model to filter retrieved chunks and lower RAG costs.
Open source