MCP server + Playwright reporter that builds a flakiness knowledge graph from test run history
A Playwright custom reporter + MCP server that builds a local flakiness knowledge graph from your test run history. Ask your AI agent which tests are unreliable, on which browser, and whether they're getting worse. A single Playwright trace tells you what failed right now. It doesn't tell you whether this test has been silently flaking for two weeks, or only fails on Firefox in CI, or is getting…
Inferred from the transports this listing declares (stdio). A client not listed here hasn’t been ruled out — it just isn’t something Forge can confirm.
Verification confirms publisher identity (repo ownership), not code safety. The security scan covers known CVEs and suspicious install scripts.
Read out of the source npm actually ships, at scan time. The package was never executed. Tools registered dynamically at runtime, or hidden inside bundled or minified code, can be missed — so this is a floor on the tool surface, not a complete census of it.
get_flaky_testsNo description publishedThis tool published no description. Forge does not invent one.
get_test_historyNo description publishedThis tool published no description. Forge does not invent one.
get_failure_patternsNo description publishedThis tool published no description. Forge does not invent one.
get_slow_testsNo description publishedThis tool published no description. Forge does not invent one.
get_error_groupsNo description publishedThis tool published no description. Forge does not invent one.
get_flakiness_trendNo description publishedThis tool published no description. Forge does not invent one.
cluster_semantic_error_treesNo description publishedThis tool published no description. Forge does not invent one.
correlate_git_commit_flakinessNo description publishedThis tool published no description. Forge does not invent one.
0 of 8 tools published a description.
Tool names and descriptions are written by the publisher and shown verbatim as inert text. They are the strings an MCP client passes to a model, so Forge scans them for prompt-injection patterns — any finding appears with the security scan above. “Privileged” is a keyword match on the tool name, not an audit of what the tool does: a benign-sounding name can still do anything.
A Playwright custom reporter + MCP server that builds a local flakiness knowledge graph from your test run history. Ask your AI agent which tests are unreliable, on which browser, and whether they're getting worse. A single Playwright trace tells you what failed right now. It doesn't tell you whether this test has been silently flaking for two weeks, or only fails on Firefox in CI, or is getting slower with every release. This tool fixes that by accumulating run history into a SQLite database…
Linked names open Forge’s index of every entry observed exposing that tool. Browse all indexed tools.
This package was last scanned before Forge began storing the resolved tree. The next scan will record it.