AI evaluation toolkit — measure inter-rater agreement across multiple LLM providers
- ✓Actively maintained (<30d)
- ✓Clear description
- ✓Documented (README)
- !Licence file present but not machine-readable
/plugin marketplace add AlligatorC0der/conkurrence
/plugin install conkurrenceResumen de Plugins
# ConKurrence
Multi-model rating for AI evaluation: measures agreement among model raters (Fleiss' κ, Kendall's W) to find contested items for expert review.
**What agreement tells you:** where your raters and your criteria are consistent, and where they are not — which items are contested, and whether disagreement comes from the criteria or the raters.
**What it does not tell you:** that the raters are right. Agreement is reliability, not validity. Models can agree confidently on a wrong answer. Use ConKurrence to decide where an expert's judgment is needed and to diagnose your evaluation criteria — not to replace expert-labelled ground truth.
**Status: maintenance mode.** ConKurrence is developed by JE Vectors LLC as the evaluation instrument behind Inqura. It is not accepting purchases and carries no support commitment.
## Install
```bash
npm install -g conkurrence
```
## MCP Server
Use ConKurrence as an MCP server in Claude Desktop or any MCP-compatible client:
```bash
npx conkurrence mcp
```
### Claude Desktop Configuration
Add to your `claude_desktop_config.json`:
```json
{
"mcpServers": {
"conkurrence": {
"command": "npx",
"args": ["-y", "conkurrence", "mcp"]
}
}
}
```
### Claude Code Plugin
```
/plugin marketplace add AlligatorC0der/conkurrence
```
## Features
- **Multi-model evaluation** — Run your schema against Bedrock, OpenAI, and Gemini models simultaneously
- **Agreement statistics** — Fleiss' kappa with bootstrap confidence intervals; Kendall's W against expert-labelled anchor items
- **Self-consistency mode** — No API keys needed; uses the host model via MCP Sampling
- **Schema suggestion** — AI-powered schema design from your data
- **Trend tracking** — Compare runs over time, detect agreement degradation
- **Cost estimation** — Know the cost before running
## MCP Tools
| Tool | Description |
|------|-------------|
| `conkurrence_run` | Execute an evaluation across multiple AI raters |
| `conkurrence_report` | Generate a detailed markdown report |
| `conkurrence_compare` | Side-by-side comparison of two runs |
| `conkurrence_trend` | Track agreement over multiple runs |
| `conkurrence_suggest` | AI-powered schema suggestion from your data |
| `conkurrence_validate_schema` | Validate a schema before running |
| `conkurrence_estimate` | Estimate cost and token usage |
## Links
- **Homepage:** [conkurrence.com](https://conkurrence.com)
- **npm:** [npmjs.com/package/conkurrence](https://www.npmjs.com/package/conkurrence)
- **Terms of Service:** [app.conkurrence.com/terms](https://app.conkurrence.com/terms)
- **Privacy Policy:** [app.conkurrence.com/privacy](https://app.conkurrence.com/privacy)
## License
[BUSL-1.1](LICENSE.md) — Business Source License 1.1
Lo que la gente pregunta sobre conkurrence
¿Qué es AlligatorC0der/conkurrence?
+
AlligatorC0der/conkurrence es plugins para el ecosistema de Claude AI. AI evaluation toolkit — measure inter-rater agreement across multiple LLM providers Tiene 0 estrellas en GitHub y su última actualización registrada es del 2026-10-03.
¿Cómo se instala conkurrence?
+
Puedes instalar conkurrence clonando el repositorio (https://github.com/AlligatorC0der/conkurrence) o siguiendo las instrucciones del README en GitHub. ClaudeWave también te ofrece bloques de instalación rápida en esta misma página.
¿Es seguro usar AlligatorC0der/conkurrence?
+
Nuestro agente de seguridad ha analizado AlligatorC0der/conkurrence y le ha asignado un Trust Score de 72/100 (tier: OK). Revisa el desglose completo de comprobaciones superadas y flags en esta página.
¿Quién mantiene AlligatorC0der/conkurrence?
+
AlligatorC0der/conkurrence es mantenido por AlligatorC0der. La última actividad registrada en GitHub es del 2026-10-03, con 0 issues abiertos.
¿Hay alternativas a conkurrence?
+
Sí. En ClaudeWave puedes explorar plugins similares en /categories/plugins, ordenados por popularidad o actividad reciente.
Despliega conkurrence en tu cloud
Lleva este repo a producción en minutos. Cada plataforma genera su propio entorno con variables de entorno editables.
¿Mantienes este repo? Añade un badge a tu README
Pega el badge en tu README de GitHub para mostrar que está auditado por ClaudeWave. Cada badge enlaza de vuelta a esta página y muestra el Trust Score actual.
[](https://claudewave.com/repo/alligatorc0der-conkurrence)<a href="https://claudewave.com/repo/alligatorc0der-conkurrence"><img src="https://claudewave.com/api/badge/alligatorc0der-conkurrence" alt="Featured on ClaudeWave: AlligatorC0der/conkurrence" width="320" height="64" /></a>Más Plugins
Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands.
AI agent skill that researches any topic across Reddit, X, YouTube, HN, Polymarket, and the web - then synthesizes a grounded summary
Write HTML. Render video. Built for agents.
Agent skill that removes signs of AI-generated writing from text
Academic Research Skills for Claude Code: research → write → review → revise → finalize
Create beautiful slides on the web using a coding agent's frontend skills