14-stage Fusion Pipeline for LLM token compression — reversible compression, AST-aware code analysis, intelligent content routing. Zero LLM inference cost. MIT licensed.
-
Updated
Apr 1, 2026 - Python
14-stage Fusion Pipeline for LLM token compression — reversible compression, AST-aware code analysis, intelligent content routing. Zero LLM inference cost. MIT licensed.
Staged context pruning for OpenCode, Oh My Pi, pi, and Claude Code.
Biological code organization system with 1,029+ production-ready snippets - 95% token reduction for Claude/GPT with AI-powered discovery & offline packs
Pi extension for dynamic context pruning
the research synthethizer outer loop
Reversible context pruning for Pi, powered by TypeSafe Jev. Keep useful context without deleting session history.
Local gateway that prunes Claude Code's API context in-flight — dedup, purge failed calls, trim bloat, strip old screenshots. Anthropic-sanctioned (ANTHROPIC_BASE_URL), no CA cert, no restart.
DeepSeek Harness 动态上下文管理插件(Dynamic Context Pruning for dsh),对标 opencode-dcp
AI-powered tutoring system for Indian state-board students. Upload textbook PDFs and get curriculum-aligned answers instantly. Uses Context Pruning to score and filter chapters before querying Gemini LLM — reducing API costs by ~80%. Built with Python, Flask, FAISS, sentence-transformers, and Gemini 2.5 Flash.
Local Ollama RAG memory proxy using ChromaDB and nomic-embed-text to retrieve relevant chat history, compress context, and reduce LLM tokens.
Enterprise-grade Context Pruning for RAG. Hierarchical 2-stage architecture (MiniLM-L6 ONNX INT8) achieving 60-75% token compression at <20ms latency.
To associate your repository with the context-pruning topic, visit your repo's landing page and select "manage topics."