media MCP Server
MCP server for LLM quantization. Compress any model to GGUF/GPTQ/AWQ in one tool call. First MCP server for model compression.
Discovered via unknown and last synced 3mo ago.
1. Install the package
uvx mcp-turboquant
2. Add to claude_desktop_config.json
{
"mcpServers": {
"mcp-turboquant": {
"command": "uvx",
"args": [
"mcp-turboquant"
]
}
}
}Config file location: ~/Library/Application Support/Claude/claude_desktop_config.json (macOS) / %APPDATA%\Claude\claude_desktop_config.json (Windows)
Quantize a model to GGUF/GPTQ/AWQ
Push quantized model to HuggingFace Hub
Run perplexity evaluation on a quantized model
Get model info from HuggingFace (params, size, architecture)
Hardware-aware recommendation for best format + bits
Check available quantization backends on the system
Description
The cheapest AI media API on the market. Generate images (Flux), music (AceStep), speech with voice cloning, transcribe video/audio, OCR, video generation, background removal, upscale, style transfer, and prompt enhancement — all through one unified API. Free $5 credit on signup.
Write professional press releases that get media attention and coverage
Learn how to use the deAPI AI Media Suite (Community) Claude skill. Complete guide with installation instructions and examples.
Learn how to use the Press Release Writer Claude skill. Complete guide with installation instructions and examples.