Smart token compression for LLM apps. Save 70-90% on API costs with Gemma 4 local compression, multi-model cost tracking, and intelligent model routing.
promptthrift-mcp
PROMPTTHRIFT_OLLAMA_URL