Measure whether your AI agrees with itself using statistical consensus metrics.
One command. Find out if your AI agrees with itself. []( [](./LICENSE.md) []() ConKurrence measures whether multiple AI models produce consistent outputs on your evaluation tasks — using the same psychometric methods trusted in clinical research (Fleiss' κ, Kendall's W, bootstrap confidence intervals). Stop guessing whether your golden dataset is reliable. Know statistically. You now know is…
Dedotto dai trasporti dichiarati da questo annuncio (stdio). Un client che non compare qui non è escluso — semplicemente Forge non è in grado di confermarlo.
La verifica conferma l’identità del publisher (la proprietà del repo), non la sicurezza del codice. L’analisi di sicurezza copre i CVE noti e gli script di installazione sospetti.
Forge non ha alcuna analisi registrata per questa voce, quindi non ha alcuna osservazione della sua superficie di strumenti. È assenza di prove, non prova che non esponga alcuno strumento.
One command. Find out if your AI agrees with itself. []( [](./LICENSE.md) []() ConKurrence measures whether multiple AI models produce consistent outputs on your evaluation tasks — using the same psychometric methods trusted in clinical research (Fleiss' κ, Kendall's W, bootstrap confidence intervals). Stop guessing whether your golden dataset is reliable. Know statistically. You now know is reliable, needs review, and should be redesigned before trusting it. Your golden dataset has a hidden…