Skip to content
#

prompt-caching

Here are 290 public repositories matching this topic...

Stop paying to re-read the same output. OMNI turns repeated bytes into retrievable handles: 97.2% off a file your agent reads twice, and across 5,984 real commands 69.6% on a heavy week, 14.9% on an ordinary one. Nothing deleted, nothing invented, every number replays on your own corpus.

  • Updated Aug 17, 2026
  • Rust

The Multi-Agent Reasoning framework creates an interactive chatbot where AI agents collaborate via structured reasoning and Swarm Integration for optimal answers. Simulating a team that discusses, debates, and refines responses, it enables complex problem-solving and precise results. Now with Prompt Caching to reduce latency and costs.

  • Updated Jan 23, 2025
  • Python

A curated list of strategies, tools, papers, and resources for reducing LLM token costs and improving efficiency in production.

  • Updated Aug 16, 2026

Improve this page

Add a description, image, and links to the prompt-caching topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the prompt-caching topic, visit your repo's landing page and select "manage topics."

Learn more