Check if a task runs locally vs cloud. Save money on calls that don't need cloud inference.
Cloud inference is expensive. Everything that can run locally should. This MCP server tells your agent — before every cloud API call — whether the task can be handled by a local model instead. Route to Ollama, LM Studio, or llama.cpp when you can. Only pay for cloud when you must. Call this BEFORE every cloud inference call. If verdict is , skip the cloud call entirely and route to your local…
Abgeleitet aus den Transporten, die dieser Eintrag deklariert (stdio, streamable-http). Ein Client, der hier nicht steht, ist damit nicht ausgeschlossen — Forge kann ihn nur nicht bestätigen.
Die Verifizierung bestätigt die Identität des Publishers (die Inhaberschaft am Repo), nicht die Sicherheit des Codes. Der Sicherheits-Scan deckt bekannte CVEs und verdächtige Installationsskripte ab.
Forge hat gegen diesen Endpunkt keinen tools/list-Handshake abgeschlossen und hat daher keine Beobachtung dessen, was der Server offenlegt. Nichts hiervon sagt, dass er nichts offenlegt.
Cloud inference is expensive. Everything that can run locally should. This MCP server tells your agent — before every cloud API call — whether the task can be handled by a local model instead. Route to Ollama, LM Studio, or llama.cpp when you can. Only pay for cloud when you must. Call this BEFORE every cloud inference call. If verdict is , skip the cloud call entirely and route to your local model. Only use cloud when this tool returns . Field | Required | Description | | ✅ | The exact task…