Skip to content
#

prompt-caching

Here are 451 public repositories matching this topic...

The Multi-Agent Reasoning framework creates an interactive chatbot where AI agents collaborate via structured reasoning and Swarm Integration for optimal answers. Simulating a team that discusses, debates, and refines responses, it enables complex problem-solving and precise results. Now with Prompt Caching to reduce latency and costs.

  • Updated Jan 23, 2025
  • Python
awesome-ai-tokenomics

A curated list on AI token economics: what tokens cost, where they get wasted, and how to cut the bill. Tools, benchmarks, papers, and copy-paste configs for the token economy of LLMs and coding agents.

  • Updated Sep 25, 2026
  • Python
CluxMate

AI coding agent with one Python core and three front-ends — headless CLI, Textual TUI, and an Electron desktop. Works with any OpenAI-compatible API, with risk-tiered permissions, event-sourced replayable sessions, and a fail-closed OS-level sandbox.

  • Updated Sep 22, 2026
  • Python

A curated list of strategies, tools, papers, and resources for reducing LLM token costs and improving efficiency in production.

  • Updated Sep 20, 2026
  • Shell
lemoncrow

A Complete faster runtime for coding agents. Optimize the journey, not the hop. Make coding agents 25% faster and 30% cheaper on average while keeping the quality same or more. Same Task, Same Quality, Faster and Cheaper.

  • Updated Sep 24, 2026
  • Python

Add this topic to your repo

To associate your repository with the prompt-caching topic, visit your repo's landing page and select "manage topics."

Learn more