TokenMark MCP server and CLI: measured local-LLM speeds and model picks for Strix Halo, DGX Spark and Mac
- ✓Open-source license (MIT)
- ✓Actively maintained (<30d)
- ✓Clear description
- ✓Topics declared
- ✓Documented (README)
claude mcp add tokenmark-mcp -- npx -y @altronis/tokenmark-mcp{
"mcpServers": {
"tokenmark-mcp": {
"command": "npx",
"args": ["-y", "@altronis/tokenmark-mcp"]
}
}
}MCP Servers overview
# @altronis/tokenmark-mcp
An [MCP](https://modelcontextprotocol.io) server that gives Claude (and other agents) real local-LLM benchmark data + hardware-aware model recommendations from [TokenMark](https://tokenmark.app).
Ask *"what should I run on a Strix Halo for coding?"* and the agent answers from **measured** configs (decode tok/s, quant, backend), each with a source link. It never invents numbers.
## Add to Claude Code
```sh
claude mcp add tokenmark -- npx -y @altronis/tokenmark-mcp
```
## Add to Cline
Cline CLI:
```sh
cline mcp add tokenmark --yes -- npx -y @altronis/tokenmark-mcp
```
Cline in VS Code: add the `tokenmark` entry from the JSON below to `cline_mcp_settings.json`. Step-by-step notes for agents are in [llms-install.md](llms-install.md).
## Add to any MCP client
Run the server over stdio:
```sh
npx -y @altronis/tokenmark-mcp
```
Or in a client config:
```json
{
"mcpServers": {
"tokenmark": {
"command": "npx",
"args": ["-y", "@altronis/tokenmark-mcp"]
}
}
}
```
## Tools
- **`tokenmark_recommend`**: `{ hardware, tasks?, prefer?, limit? }` → ranked model + best-config picks (5 by default, 10 max) with measured tok/s + why.
- **`tokenmark_configs`**: `{ model?, hardware?, limit? }` → tracked benchmark configs, fastest first (25 by default, 50 max; `matched` gives the full count), with the run mode (`speculative: true` for MTP/DFlash/draft-model runs).
- **`tokenmark_hardware`**: `{ platform? }` → the hardware catalogue (Strix Halo, Gorgon Halo, DGX Spark, Mac Max/Ultra). With a platform: chip specs with a source per value, the boxes that ship it, Singapore prices.
- **`tokenmark_search`**: `{ term }` → matching models/configs.
- **`tokenmark_submit`**: `{ repo, source?, note? }` → queues a GitHub/Hugging Face repo with benchmark numbers for human review.
- **`tokenmark_submission_status`**: `{ id }` → where a submission is.
A speed someone measured themselves, or hardware missing from the catalogue, goes through the signed-in form at https://tokenmark.app/submit.
Data is pulled live from `https://tokenmark.app` (override with `TOKENMARK_URL`). Zero runtime dependencies.
## The same data in your terminal
The CLI lives in [`cli/`](https://github.com/sypherin/tokenmark-mcp/tree/main/cli) of this repo:
```sh
npx @altronis/tokenmark-cli recommend --hardware "Strix Halo" --tasks coding
```
The files in `bin/` are the published builds, made from the TokenMark tracker that runs https://tokenmark.app.
MIT licensed. Data aggregated from public community benchmarks with attribution.
What people ask about tokenmark-mcp
What is sypherin/tokenmark-mcp?
+
sypherin/tokenmark-mcp is mcp servers for the Claude AI ecosystem. TokenMark MCP server and CLI: measured local-LLM speeds and model picks for Strix Halo, DGX Spark and Mac It has 0 GitHub stars and its last recorded update is dated 2026-10-05.
How do I install tokenmark-mcp?
+
You can install tokenmark-mcp by cloning the repository (https://github.com/sypherin/tokenmark-mcp) or following the README instructions on GitHub. ClaudeWave also provides quick install blocks on this page.
Is sypherin/tokenmark-mcp safe to use?
+
Our security agent has analyzed sypherin/tokenmark-mcp and assigned a Trust Score of 95/100 (tier: Verified). See the full breakdown of passed checks and flags on this page.
Who maintains sypherin/tokenmark-mcp?
+
sypherin/tokenmark-mcp is maintained by sypherin. The last recorded GitHub activity is dated 2026-10-05, with 0 open issues.
Are there alternatives to tokenmark-mcp?
+
Yes. On ClaudeWave you can browse similar mcp servers at /categories/mcp, sorted by popularity or recent activity.
Deploy tokenmark-mcp to your cloud
Ship this repo to production in minutes. Each platform spins up its own environment with editable env vars.
Maintain this repo? Add a badge to your README
Drop the badge into your GitHub README to show it's tracked on ClaudeWave. Each badge links back to this page and reflects the live Trust Score.
[](https://claudewave.com/repo/sypherin-tokenmark-mcp)<a href="https://claudewave.com/repo/sypherin-tokenmark-mcp"><img src="https://claudewave.com/api/badge/sypherin-tokenmark-mcp" alt="Featured on ClaudeWave: sypherin/tokenmark-mcp" width="320" height="64" /></a>More MCP Servers
Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or cloud, 400+ integrations.
User-friendly AI Interface (Supports Ollama, OpenAI API, ...)
An open-source AI agent that brings the power of Gemini directly into your terminal.
Real-time global intelligence dashboard. AI-powered news aggregation, geopolitical monitoring, and infrastructure tracking in a unified situational awareness interface
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl! Don't be shy, join here: https://discord.gg/EMgGbDceNQ and follow here for daily tips and tricks: https://x.com/Scrapling_dev
The fastest path to AI-powered full stack observability, even for lean teams.