AI evaluation toolkit — measure inter-rater agreement across multiple LLM providers
- ✓Actively maintained (<30d)
- ✓Clear description
- ✓Documented (README)
- !Licence file present but not machine-readable
/plugin marketplace add AlligatorC0der/conkurrence
/plugin install conkurrencePlugins overview
# ConKurrence
Multi-model rating for AI evaluation: measures agreement among model raters (Fleiss' κ, Kendall's W) to find contested items for expert review.
**What agreement tells you:** where your raters and your criteria are consistent, and where they are not — which items are contested, and whether disagreement comes from the criteria or the raters.
**What it does not tell you:** that the raters are right. Agreement is reliability, not validity. Models can agree confidently on a wrong answer. Use ConKurrence to decide where an expert's judgment is needed and to diagnose your evaluation criteria — not to replace expert-labelled ground truth.
**Status: maintenance mode.** ConKurrence is developed by JE Vectors LLC as the evaluation instrument behind Inqura. It is not accepting purchases and carries no support commitment.
## Install
```bash
npm install -g conkurrence
```
## MCP Server
Use ConKurrence as an MCP server in Claude Desktop or any MCP-compatible client:
```bash
npx conkurrence mcp
```
### Claude Desktop Configuration
Add to your `claude_desktop_config.json`:
```json
{
"mcpServers": {
"conkurrence": {
"command": "npx",
"args": ["-y", "conkurrence", "mcp"]
}
}
}
```
### Claude Code Plugin
```
/plugin marketplace add AlligatorC0der/conkurrence
```
## Features
- **Multi-model evaluation** — Run your schema against Bedrock, OpenAI, and Gemini models simultaneously
- **Agreement statistics** — Fleiss' kappa with bootstrap confidence intervals; Kendall's W against expert-labelled anchor items
- **Self-consistency mode** — No API keys needed; uses the host model via MCP Sampling
- **Schema suggestion** — AI-powered schema design from your data
- **Trend tracking** — Compare runs over time, detect agreement degradation
- **Cost estimation** — Know the cost before running
## MCP Tools
| Tool | Description |
|------|-------------|
| `conkurrence_run` | Execute an evaluation across multiple AI raters |
| `conkurrence_report` | Generate a detailed markdown report |
| `conkurrence_compare` | Side-by-side comparison of two runs |
| `conkurrence_trend` | Track agreement over multiple runs |
| `conkurrence_suggest` | AI-powered schema suggestion from your data |
| `conkurrence_validate_schema` | Validate a schema before running |
| `conkurrence_estimate` | Estimate cost and token usage |
## Links
- **Homepage:** [conkurrence.com](https://conkurrence.com)
- **npm:** [npmjs.com/package/conkurrence](https://www.npmjs.com/package/conkurrence)
- **Terms of Service:** [app.conkurrence.com/terms](https://app.conkurrence.com/terms)
- **Privacy Policy:** [app.conkurrence.com/privacy](https://app.conkurrence.com/privacy)
## License
[BUSL-1.1](LICENSE.md) — Business Source License 1.1
What people ask about conkurrence
What is AlligatorC0der/conkurrence?
+
AlligatorC0der/conkurrence is plugins for the Claude AI ecosystem. AI evaluation toolkit — measure inter-rater agreement across multiple LLM providers It has 0 GitHub stars and its last recorded update is dated 2026-10-03.
How do I install conkurrence?
+
You can install conkurrence by cloning the repository (https://github.com/AlligatorC0der/conkurrence) or following the README instructions on GitHub. ClaudeWave also provides quick install blocks on this page.
Is AlligatorC0der/conkurrence safe to use?
+
Our security agent has analyzed AlligatorC0der/conkurrence and assigned a Trust Score of 72/100 (tier: OK). See the full breakdown of passed checks and flags on this page.
Who maintains AlligatorC0der/conkurrence?
+
AlligatorC0der/conkurrence is maintained by AlligatorC0der. The last recorded GitHub activity is dated 2026-10-03, with 0 open issues.
Are there alternatives to conkurrence?
+
Yes. On ClaudeWave you can browse similar plugins at /categories/plugins, sorted by popularity or recent activity.
Deploy conkurrence to your cloud
Ship this repo to production in minutes. Each platform spins up its own environment with editable env vars.
Maintain this repo? Add a badge to your README
Drop the badge into your GitHub README to show it's tracked on ClaudeWave. Each badge links back to this page and reflects the live Trust Score.
[](https://claudewave.com/repo/alligatorc0der-conkurrence)<a href="https://claudewave.com/repo/alligatorc0der-conkurrence"><img src="https://claudewave.com/api/badge/alligatorc0der-conkurrence" alt="Featured on ClaudeWave: AlligatorC0der/conkurrence" width="320" height="64" /></a>More Plugins
Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands.
AI agent skill that researches any topic across Reddit, X, YouTube, HN, Polymarket, and the web - then synthesizes a grounded summary
Write HTML. Render video. Built for agents.
Agent skill that removes signs of AI-generated writing from text
Academic Research Skills for Claude Code: research → write → review → revise → finalize
Create beautiful slides on the web using a coding agent's frontend skills