Audio Transcriber

media MCP Server

Transcribe audio into text. Agentic AI supported through MCP Server.

VerifiedInstall Ready
mediamedia
9 views2 stars0 forksMIT

Why This Matters

Discovered via github-topic:mcp-server and last synced 3mo ago.

VerifiedInstall Ready
Source
github-topic:mcp-server
Stars
2
Last synced
3mo ago
Install
Instructions detected

Install

1. Install the package

uvx --from audio-transcriber audio-transcriber-mcp

2. Add to claude_desktop_config.json

{
  "mcpServers": {
    "audio-transcriber": {
      "command": "npx",
      "args": [
        "audio-transcriber"
      ]
    }
  }
}

Config file location: ~/Library/Application Support/Claude/claude_desktop_config.json (macOS) / %APPDATA%\Claude\claude_desktop_config.json (Windows)

20
Tools
0
Resources
0
Prompts
Standard I/O
Transport

Available Tools (20)

EUNOMIA_POLICY_FILE

Path to the Eunomia security guardrail policies JSON file.

base

~1 GB

transcribe_audio

Transcribes audio from a provided file or by recording from the microphone.

Description

Default / Example

medium

~5 GB

Size

Parameters

large

~10 GB

AG-UI

MCP ![PyPI - Version](https://img.shields.io/pypi/v/audio-transcriber) ![MCP Server](https://badge.mcpx.dev?type=server 'MCP Server') ![PyPI - Downloads](https://img.shields.io/pypi/dd/audio-transcriber) ![GitHub Repo stars](https://img.shields.io/github/stars/Knuckles-Team/audio-transcriber) ![GitHub forks](https://img.shields.io/github/forks/Knuckles-Team/audio-transcriber) ![GitHub contributors](https://img.shields.io/github/contributors/Knuckles-Team/audio-transcriber) ![PyPI - License](https://img.shields.io/pypi/l/audio-transcriber) ![GitHub](https://img.shields.io/github/license/Knuckles-Team/audio-transcriber) ![GitHub last commit (by committer)](https://img.shields.io/github/last-commit/Knuckles-Team/audio-transcriber) ![GitHub pull requests](https://img.shields.io/github/issues-pr/Knuckles-Team/audio-transcriber) ![GitHub closed pull requests](https://img.shields.io/github/issues-pr-closed/Knuckles-Team/audio-transcriber) ![GitHub issues](https://img.shields.io/github/issues/Knuckles-Team/audio-transcriber) ![GitHub top language](https://img.shields.io/github/languages/top/Knuckles-Team/audio-transcriber) ![GitHub language count](https://img.shields.io/github/languages/count/Knuckles-Team/audio-transcriber) ![GitHub repo size](https://img.shields.io/github/repo-size/Knuckles-Team/audio-transcriber) ![GitHub repo file count (file type)](https://img.shields.io/github/directory-file-count/Knuckles-Team/audio-transcriber) ![PyPI - Wheel](https://img.shields.io/pypi/wheel/audio-transcriber) ![PyPI - Implementation](https://img.shields.io/pypi/implementation/audio-transcriber) *Version: 0.14.0* ## Overview Transcribe your .wav .mp4 .mp3 .flac files to text or record your own audio! This repository is actively maintained - Contributions are welcome! Contribution Opportunities: - Support new models Wrapped around [OpenAI Whisper](https://pypi.org/project/openai-whisper) ## MCP ## MCP Tools

tiny

~1 GB

small

~2 GB

AUDIO_PROCESSINGTOOL

Boolean flag for enabling internal audio processing tools.

MCP

Agent ![PyPI - Version](https://img.shields.io/pypi/v/audio-transcriber) ![MCP Server](https://badge.mcpx.dev?type=server 'MCP Server') ![PyPI - Downloads](https://img.shields.io/pypi/dd/audio-transcriber) ![GitHub Repo stars](https://img.shields.io/github/stars/Knuckles-Team/audio-transcriber) ![GitHub forks](https://img.shields.io/github/forks/Knuckles-Team/audio-transcriber) ![GitHub contributors](https://img.shields.io/github/contributors/Knuckles-Team/audio-transcriber) ![PyPI - License](https://img.shields.io/pypi/l/audio-transcriber) ![GitHub](https://img.shields.io/github/license/Knuckles-Team/audio-transcriber) ![GitHub last commit (by committer)](https://img.shields.io/github/last-commit/Knuckles-Team/audio-transcriber) ![GitHub pull requests](https://img.shields.io/github/issues-pr/Knuckles-Team/audio-transcriber) ![GitHub closed pull requests](https://img.shields.io/github/issues-pr-closed/Knuckles-Team/audio-transcriber) ![GitHub issues](https://img.shields.io/github/issues/Knuckles-Team/audio-transcriber) ![GitHub top language](https://img.shields.io/github/languages/top/Knuckles-Team/audio-transcriber) ![GitHub language count](https://img.shields.io/github/languages/count/Knuckles-Team/audio-transcriber) ![GitHub repo size](https://img.shields.io/github/repo-size/Knuckles-Team/audio-transcriber) ![GitHub repo file count (file type)](https://img.shields.io/github/directory-file-count/Knuckles-Team/audio-transcriber) ![PyPI - Wheel](https://img.shields.io/pypi/wheel/audio-transcriber) ![PyPI - Implementation](https://img.shields.io/pypi/implementation/audio-transcriber) *Version: 0.33.0* > **Documentation** — Installation, deployment, and usage across the CLI, Python API, > MCP server, and A2A agent are maintained in the > [official documentation](https://knuckles-team.github.io/audio-transcriber/). --- ## Overview **Audio Transcriber** is a production-grade Agent and Model Context Protocol (MCP) server designed to interface directly with Transcribe your .wav .mp4 .mp3 .flac files to text or record your own audio!. --- ## Key Features - **Consolidated Action-Routed MCP Tools:** Minimizes token overhead and eliminates tool bloat in LLM contexts by grouping methods into optimized, togglable tool modules. - **Enterprise-Grade Security:** Comprehensive support for Eunomia policies, OIDC token delegation, and granular execution context tracking. - **Integrated Graph Agent:** Built-in Pydantic AI agent supporting the Agent Control Protocol (ACP) and standard Web interfaces (AG-UI). - **Native Telemetry & Tracing:** Out-of-the-box OpenTelemetry exports and native Langfuse tracing. --- ## CLI or API This agent wraps the Transcribe your .wav .mp4 .mp3 .flac files to text or record your own audio! API. You can interact with it programmatically or via its integrated execution entrypoints. Detailed instructions on how to use the underlying API wrappers, extended schema bindings, and developer SDK references are maintained in [docs/index.md](docs/index.md). --- ## MCP This server utilizes dynamic Action-Routed tools to optimize token overhead and maximize IDE compatibility. ### Available MCP Tools

Feature

Functionality

EUNOMIA_TYPE

Eunomia guardrail deployment type (e.g., `none`, `embedded`, `remote`).

Page

Contents

AUDIO_PROCESSING_TOOL

Toggle the audio processing tool module.

WHISPER_MODEL

Standard OpenAI Whisper model to use for local transcription (e.g., `base`, `tiny`, `small`).

AUTH_TYPE

Security authentication type to apply (e.g., `jwt`, `none`).

MISC_TOOL

`True`

OTEL_EXPORTER_OTLP_ENDPOINT

OpenTelemetry collector endpoint for exporting traces.