AI design system scanner + mcp server. Scores repos 0-100, writes fix prompts. 27K+ npm downloads.
- ✓Open-source license (MIT)
- ✓Actively maintained (<30d)
- ✓Clear description
- ✓Topics declared
- ✓Documented (README)
claude mcp add roast-my-design-system -- npx -y roast-my-design-system{
"mcpServers": {
"roast-my-design-system": {
"command": "npx",
"args": ["-y", "roast-my-design-system"]
}
}
}MCP Servers overview
<img src="assets/roastmds.svg" width="100" alt="roast-my-design-system">
# roast-my-design-system
[](https://www.npmjs.com/package/roast-my-design-system) [](https://www.npmjs.com/package/roast-my-design-system) [](https://socket.dev/npm/package/roast-my-design-system) [](LICENSE) [](https://www.npmjs.com/package/roast-my-design-system?activeTab=dependencies) [](#what-makes-the-numbers-trustworthy)
[](#live-answers-over-mcp) [](#live-answers-over-mcp) [](#live-answers-over-mcp)
## Find where your AI agent will invent UI.
Your AI can write the UI. This makes sure it writes *your* UI.
```bash
npx roast-my-design-system@latest
```
Run it at the root of a UI repo. One second later, a report opens.
No account. No network. No telemetry. Nothing in your repo changes.
Version history is in [CHANGELOG.md](CHANGELOG.md).
## The idea
### An agent copies what it finds.
### It invents where the repo has no answer.
We tested this on 10 real open-source products, in 355 agent sessions, with and without this tool.
Findings each session added, counted by this tool's own rules:
| Task | Model | When | Without roast | Roast MCP and rules installed | Roast plugin with its edit hook |
|---|---|---|---|---|---|
| A new chart and a new component, 4 products, 3 runs each | Sonnet 5 | Oct 2026, roast 10.1 | 21 in 24 runs | not run | 7 in 24 runs |
| A new chart and a new component, 4 products, 3 runs each | Haiku 4.5 | Oct 2026, roast 10.1 | 72 in 24 runs | not run | 1 in 24 runs |
| Add a panel, build a dashboard, tighten a list (80 sessions) | Sonnet 5 | Sept 2026 | 7 | 9 | not run |
| A Christmas theme, 10 products | Sonnet 5 | Sept 2026 | 28 | 2 | not run |
| A new chart, 4 products, 5 runs each | Sonnet 5 | Sept 2026 | 10 in 20 runs | not run | 0 in 20 runs |
| A new chart, 4 products | Haiku 4.5 | Sept 2026 | 25 | 39 | 0 in 14 of 15 runs |
| Chart, status colour, empty state, new component | Haiku 4.5 | Sept 2026 | 42 in 16 runs | 49 in 16 runs | 3 in 51 runs |
Routine work stayed on-system with or without the tool. Work that needed something new did not.
An agent does not always call a tool when it should. In October, Sonnet called the MCP tools in 7 of 24 sessions and Haiku in 2 of 24. The edit hook runs without being asked. That is the column with the lowest numbers.
Where did the agent invent? Where the repo had nothing to copy. Dub has no chart palette. Its own charts hardcode 17 colours. Asked for a chart, Haiku hardcoded 9 more. In October, 5 of the 7 findings Sonnet left with the plugin on were chart colours on Dub. The agent kept them and wrote a comment saying the repo has no chart palette.
So the mess an agent adds is a map of the gaps in your system.
This tool draws the map.
## Roast. Teach. Guard.
**Roast** the repo to find its real design system, the mess in it, and the gaps.
**Teach** the agent, through generated rules and a local MCP server.
**Guard** every edit, so the agent does not have to remember to ask.
ROAST what is in the repo, and what is missing
↓
TEACH rules in the agent files, answers over MCP
↓
GUARD every edit checked, the PR gated
## What you get
### Where will your agent have to guess?
The first thing the report says: the places where the repo has no answer yet, or "No gaps found".
- the gap, in one line
- the files that prove it
- the one move that closes it
One gap is known for certain today, because the runs found it: charts hardcoding colours in a repo with no chart palette. 43 of 126 public repos look like that. A gap joins the list when it has been measured.
### Fixes you can make right now to increase the health score
The mess the agent will copy, ranked by what fixing it is worth.
- what was found
- where, with the file path
- why it matters
- what to change
- a copy button with the fix prompt for your agent
Fix, rescan, press the next button.
### A score you can defend in a meeting
0 to 100. The same number every run.
Measured against three yardsticks: the ideal norms of a design system, the median of 34 product repos at the core of a 119-repo benchmark, and 10 reputable systems (Primer, Polaris, Carbon, shadcn/ui and others).
Monorepos get a score per package. `packages/ui` at 80 stops hiding `apps/web` at 40.
### Every finding, with its file path
- every hardcoded colour, and its near-identical twin
- every spacing value off the scale
- every duplicated component
And 8 more kinds, all in [What it measures](#what-it-measures). One HTML file. Open it, Slack it, email it.
The colour usage bar shows where the repo's colour comes from: of every 100 colour uses, how many read a theme colour by name, how many use Tailwind's palette, and how many are strays written by hand. It follows CSS variables, Tailwind themes, Sass and Less variables and JavaScript theme objects to do the counting.
The header names what the repo is built on: shadcn/ui, a Tailwind theme, MUI, Mantine, Chakra or Ant Design, and both kits when a repo uses two.
### Rules for your agent
Generated from your repo, into `design-system-rules.md`:
canonical components
the token file
known duplicates to avoid
spacing steps
typefaces
the kit's own vocabulary
`--apply` writes them into every agent file you have: Claude, Cursor, GitHub Copilot, Windsurf. Every scan also checks the rules you already have for stale references.
### A script does the counting
Every number comes from a deterministic read of your files. Claude writes the explanation, labelled as written by AI and kept apart from the numbers.
## How the runs were done
Claude Code, headless, on 10 public products pinned to one commit each: cal.com, Dub, Metabase, Plausible, SigNoz, trigger.dev and four more. 259 sessions in September 2026 and 96 in October 2026 on roast 10.1, with Sonnet 5 and Haiku 4.5. The October sessions repeat the chart and new-component tasks on Plausible, SigNoz, trigger.dev and Dub at the same commits. Every changed file was judged by this tool's rules at the end of the session and at the pinned commit. A finding counts only if the session added it.
Nothing was rendered. Zero findings means the code follows these rules. It does not mean the design was reviewed. Method, tables and limits are in the research write-up, which will be published separately.
## Live examples
Eleven reports, hosted exactly as the tool writes them. Every number deterministic, every path real.
- **[npx shadcn create, fresh](https://gregkozakiewicz.github.io/roast-my-design-system/examples/shadcn-create-fresh.html)** (factory install, all 61 components): read as a fresh install, "the score is the kit's, not yours"; 13 colours, every theme variable in place, shadcn's own 24 bracket values named and not counted. No score.
- **[Unleash](https://gregkozakiewicz.github.io/roast-my-design-system/examples/unleash-mui.html)** (MUI): 1,109 files import the kit and the theme is read 6,166 times; 4 colours and 2 spacings per 100 kit files are written onto components. Score 60.
- **[Metabase](https://gregkozakiewicz.github.io/roast-my-design-system/examples/metabase-mantine.html)** (Mantine): its own wrapper over Mantine counts as the kit, so 2,679 files are read instead of 392; nothing written onto components, 2 spacings per 100 kit files. Score 47.
- **[SigNoz](https://gregkozakiewicz.github.io/roast-my-design-system/examples/signoz-antd.html)** (Ant Design): 2 colours and 7 spacings per 100 kit files written in style objects where a token exists. Score 51.
- **[Apache Airflow](https://gregkozakiewicz.github.io/roast-my-design-system/examples/airflow-chakra.html)** (Chakra UI): no colours written onto components, 4 spacings per 100 kit files as pixel strings where a space step exists. Score 43.
- **[vercel/ai-chatbot](https://gregkozakiewicz.github.io/roast-my-design-system/examples/vercel-ai-chatbot.html)** (shadcn install): 71 values like [13px] written outside the Tailwind scale, and 66 palette colours per 100 files where a theme variable exists; Claude's notes embedded. Score 80.
- **[excalidraw/excalidraw](https://gregkozakiewicz.github.io/roast-my-design-system/examples/excalidraw-excalidraw.html)**: 78 off-scale spacing values and 90 !important declarations. Score 55.
- **[dubinc/dub](https://gregkozakiewicz.github.io/roast-my-design-system/examples/dubinc-dub.html)**: 642 arbitrary bracket values, 21 duplicated components, and the chart-palette gap named. Score 20.
- **[telekom/scale](https://gregkozakiewicz.github.io/roast-my-design-system/examples/telekom-scale.html)** (Stencil): 95 Stencil components read by tag; 66 spacing values outside the scale where about 12 would do; Claude's notes embedded. Score 55.
- **[magicuidesign/magicui](https://gregkozakiewicz.github.io/roast-my-design-system/examples/magicui.html)** (registry): counted on the components it publishes, 52 off-theme colours per 100 files in the code it ships, its docs site kept out and named. Score 78.
- **[adobe/spectrum-web-components](https://gregkozakiewicz.github.io/roast-my-design-system/examples/adobe-spectrum.html)** (Lit): 740 colour tokens with 8 hardcoded colours beside them, and 37 !important declarations. Score 83.
The full report for vercel/ai-chatbot. The verdict answers wWhat people ask about roast-my-design-system
What is gregkozakiewicz/roast-my-design-system?
+
gregkozakiewicz/roast-my-design-system is mcp servers for the Claude AI ecosystem. AI design system scanner + mcp server. Scores repos 0-100, writes fix prompts. 27K+ npm downloads. It has 26 GitHub stars and its last recorded update is dated 2026-10-05.
How do I install roast-my-design-system?
+
You can install roast-my-design-system by cloning the repository (https://github.com/gregkozakiewicz/roast-my-design-system) or following the README instructions on GitHub. ClaudeWave also provides quick install blocks on this page.
Is gregkozakiewicz/roast-my-design-system safe to use?
+
Our security agent has analyzed gregkozakiewicz/roast-my-design-system and assigned a Trust Score of 95/100 (tier: Verified). See the full breakdown of passed checks and flags on this page.
Who maintains gregkozakiewicz/roast-my-design-system?
+
gregkozakiewicz/roast-my-design-system is maintained by gregkozakiewicz. The last recorded GitHub activity is dated 2026-10-05, with 0 open issues.
Are there alternatives to roast-my-design-system?
+
Yes. On ClaudeWave you can browse similar mcp servers at /categories/mcp, sorted by popularity or recent activity.
Deploy roast-my-design-system to your cloud
Ship this repo to production in minutes. Each platform spins up its own environment with editable env vars.
Maintain this repo? Add a badge to your README
Drop the badge into your GitHub README to show it's tracked on ClaudeWave. Each badge links back to this page and reflects the live Trust Score.
[](https://claudewave.com/repo/gregkozakiewicz-roast-my-design-system)<a href="https://claudewave.com/repo/gregkozakiewicz-roast-my-design-system"><img src="https://claudewave.com/api/badge/gregkozakiewicz-roast-my-design-system" alt="Featured on ClaudeWave: gregkozakiewicz/roast-my-design-system" width="320" height="64" /></a>More MCP Servers
Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or cloud, 400+ integrations.
User-friendly AI Interface (Supports Ollama, OpenAI API, ...)
An open-source AI agent that brings the power of Gemini directly into your terminal.
Real-time global intelligence dashboard. AI-powered news aggregation, geopolitical monitoring, and infrastructure tracking in a unified situational awareness interface
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl! Don't be shy, join here: https://discord.gg/EMgGbDceNQ and follow here for daily tips and tricks: https://x.com/Scrapling_dev
The fastest path to AI-powered full stack observability, even for lean teams.