MCP server that minimizes LLM token usage by compressing, summarizing, filtering, chunk-referencing, and pruning large …
MCP server that minimizes LLM token usage by compressing, summarizing, filtering, chunk-referencing, and pruning large context before it reaches the model, with heuristic or local-SLM smart actions, caching, and token counting.
hocestnonsatis
mcp
free
No benchmark results have been added yet.