Mcp Turboquant

media MCP Server

MCP server for LLM quantization. Compress any model to GGUF/GPTQ/AWQ in one tool call. First MCP server for model compression.

Install Ready
mediamedia
7 views3 stars1 forksMIT

Why This Matters

Discovered via unknown and last synced 3mo ago.

Install Ready
Source
unknown
Stars
3
Last synced
3mo ago
Install
Instructions detected

Install

1. Install the package

uvx mcp-turboquant

2. Add to claude_desktop_config.json

{
  "mcpServers": {
    "mcp-turboquant": {
      "command": "uvx",
      "args": [
        "mcp-turboquant"
      ]
    }
  }
}

Config file location: ~/Library/Application Support/Claude/claude_desktop_config.json (macOS) / %APPDATA%\Claude\claude_desktop_config.json (Windows)

7
Tools
0
Resources
0
Prompts
Standard I/O
Transport

Available Tools (7)

quantize

Quantize a model to GGUF/GPTQ/AWQ

push

Push quantized model to HuggingFace Hub

evaluate

Run perplexity evaluation on a quantized model

info

Get model info from HuggingFace (params, size, architecture)

recommend

Hardware-aware recommendation for best format + bits

check

Check available quantization backends on the system

Tool

Description