2 servers tagged with "cost-reduction"
MCP server for TOON (Token-Oriented Object Notation) - Reduce LLM token usage by 50-70%
Check whether a task can run on a local model instead of cloud before every inference call. Saves money on every call that does not need cloud.