14-stage Fusion Pipeline for LLM token compression — reversible compression, AST-aware code analysis, intelligent content routing. Zero LLM inference cost. MIT licensed.
-
Updated
Apr 1, 2026 - Python
14-stage Fusion Pipeline for LLM token compression — reversible compression, AST-aware code analysis, intelligent content routing. Zero LLM inference cost. MIT licensed.
Staged context pruning for OpenCode, Oh My Pi, pi, and Claude Code.
Biological code organization system with 1,029+ production-ready snippets - 95% token reduction for Claude/GPT with AI-powered discovery & offline packs
Pi extension for dynamic context pruning
the research synthethizer outer loop
Reversible context pruning for Pi, powered by TypeSafe Jev. Keep useful context without deleting session history.
Self-hosted gateway for LLMs and decision models. Claude Code, Codex, OpenCode and any OpenAI/Anthropic SDK app reach any provider (OpenRouter, Groq, Ollama…) with context pruning; Jev and open-rlcd System One decisions get audit and calibration. Live dashboard, single Go binary.
Local gateway that prunes Claude Code's API context in-flight — dedup, purge failed calls, trim bloat, strip old screenshots. Anthropic-sanctioned (ANTHROPIC_BASE_URL), no CA cert, no restart.
DeepSeek Harness 动态上下文管理插件(Dynamic Context Pruning for dsh),对标 opencode-dcp
Stop LLM context echoing, repetition spirals, and memory poisoning in multi-turn chats.
AI-powered tutoring system for Indian state-board students. Upload textbook PDFs and get curriculum-aligned answers instantly. Uses Context Pruning to score and filter chapters before querying Gemini LLM — reducing API costs by ~80%. Built with Python, Flask, FAISS, sentence-transformers, and Gemini 2.5 Flash.
Local Ollama RAG memory proxy using ChromaDB and nomic-embed-text to retrieve relevant chat history, compress context, and reduce LLM tokens.
Enterprise-grade Context Pruning for RAG. Hierarchical 2-stage architecture (MiniLM-L6 ONNX INT8) achieving 60-75% token compression at <20ms latency.
opencode V2 adapter for fast-jev-compaction: prunes stale tool calls and results from the outgoing request using a typed-decision endpoint
Jev Prune for Claude Code: lossless context pruning, compaction and filtering
To associate your repository with the context-pruning topic, visit your repo's landing page and select "manage topics."