Is Crispasr Agent Transcriber free?

Yes, Crispasr Agent Transcriber is free to use.

How do I install Crispasr Agent Transcriber?

Crispasr Agent Transcriber is a local plugin. Install it using PyPI package: crispasr-agent-transcriber and add the generated configuration snippet to your AI app's MCP config file. Then restart your AI app.

What AI apps work with Crispasr Agent Transcriber?

Crispasr Agent Transcriber uses the Model Context Protocol (MCP) and works with any MCP-compatible AI app, including Claude, ChatGPT / Codex, Gemini, Copilot, Cursor, and more.

Back to Browse

Crispasr Agent Transcriber MCP Server

by EmiyaKatuz

Developer ToolsLow Risk9.5MCP RegistryLocal

Free

Server data from the Official MCP Registry

Transcribe local audio and video with CrispASR and local models only.

About

Transcribe local audio and video with CrispASR and local models only.

Security Report

9.5

Low Risk9.5Low Risk

Valid MCP server (2 strong, 3 medium validity signals). 1 known CVE in dependencies Imported from the Official MCP Registry. 1 finding(s) downgraded by scanner intelligence.

12 files analyzed · 2 issues found

Security scores are indicators to help you make informed decisions, not guarantees. Always review permissions before connecting any MCP server.

Permissions Required

This plugin requests these system permissions. Most are normal for its category.

file_system

Check that this permission is expected for this type of plugin.

Shell Command Execution

Runs commands on your machine. Be cautious — only use if you trust this plugin.

How to Install

Add this to your MCP configuration file:

{
  "mcpServers": {
    "io-github-emiyakatuz-crispasr-agent-transcriber": {
      "args": [
        "crispasr-agent-transcriber"
      ],
      "command": "uvx"
    }
  }
}

Documentation

View on GitHub

From the project's GitHub README.

crispasr-agent-transcriber

Local-only transcription for Codex and MCP-based AI agents, powered by CrispASR. No cloud uploads, no API keys required for transcription.

GitHub Release | npm installer | PyPI package | MCP Registry

What it does

Give it a local audio or video file. It:

Probes the spoken language (English or Chinese) using CrispASR's FireRed LID.
Starts a local CrispASR server with the right backend -- Cohere Transcribe for English, Qwen3-ASR for Chinese.
Extracts audio from video with ffmpeg when needed.
Calls CrispASR's /v1/audio/transcriptions endpoint.
Writes the transcript and metadata to disk.

Everything runs on your machine. Media never leaves it.

Quick install for Codex

The plugin includes the Codex Skill, command-line tool, and MCP server. Media stays on your computer. Model files are never downloaded automatically.

1. Install prerequisites

Install Node.js 20 or newer, uv, and ffmpeg. The installer uses uv to provide Python.

node --version
uv --version
ffmpeg -version

2. Run the installer

npx @emiyakatuz/crispasr-agent-transcriber@latest install

The installer:

downloads the matching GitHub Release and verifies its SHA-256 checksum;
installs the plugin under ~/plugins/crispasr-agent-transcriber;
installs the Python and MCP dependencies;
detects CUDA, Vulkan, or CPU and installs the best CrispASR build;
registers the plugin in the Codex Personal marketplace;
preserves existing models, binaries, and outputs during updates.

3. Add the local models

When a model is missing, the installer prints its official source and stops. Download the three files listed under Required models into:

~/plugins/crispasr-agent-transcriber/models/

Then verify the complete installation:

npx @emiyakatuz/crispasr-agent-transcriber@latest doctor

4. Enable the plugin

With a Codex build that supports plugin commands, run:

codex plugin add crispasr-agent-transcriber@personal

If the CLI has no codex plugin command, open the Codex desktop Plugins view and install CrispASR Transcriber from the Personal marketplace. Start a new conversation, then ask:

Transcribe C:\path\to\sample.mp4 with CrispASR using auto language detection.
Save a verbose JSON transcript and an SRT subtitle file.

Update or uninstall

npx @emiyakatuz/crispasr-agent-transcriber@latest update
npx @emiyakatuz/crispasr-agent-transcriber@latest uninstall

Uninstall preserves local models, CrispASR binaries, and outputs. Use uninstall --purge-data only when those files should also be deleted. See Plugin installation for manual installation and troubleshooting.

Direct command-line use

After installation, you can run the transcription script without Codex:

Set-Location (Join-Path $HOME "plugins\crispasr-agent-transcriber")
uv run python scripts/transcribe.py sample.mp4 --profile auto `
  --manage-server `
  --english-model models\cohere-transcribe.gguf `
  --chinese-model models\qwen3-asr-1.7b-q4_k.gguf `
  --lid-backend firered --lid-model models\firered-lid-q2_k.gguf `
  --format verbose_json

Use with other AI agents

The MCP server is the cross-agent interface. Any agent that supports MCP stdio can run the released package directly from GitHub:

uvx --from "crispasr-agent-transcriber[mcp] @ git+https://github.com/EmiyaKatuz/crispasr-agent-transcriber.git@v0.3.5" crispasr-agent-mcp

Use the same command and arguments in Claude Desktop, Cursor, or another MCP client. See AI agent integrations for a generic MCP configuration and Codex CLI command.

Maintainer publishing

End users do not need the release steps. Maintainers should follow the publishing guide for Codex Marketplace, PyPI, MCP Registry, and cross-agent distribution.

Required models

This tool does not download models automatically. Download these three GGUF files and keep them in a local directory (the repo's models/ folder works well):

Purpose	Local file	Variant / size	Model page	File page
English ASR	`cohere-transcribe.gguf`	F16, ~3.85 GB	Cohere Transcribe 03-2026 GGUF	Download
Chinese ASR	`qwen3-asr-1.7b-q4_k.gguf`	Q4_K, ~1.33 GB	Qwen3-ASR 1.7B GGUF	Download
Language detection	`firered-lid-q2_k.gguf`	Q2_K, ~350 MB	FireRed LID GGUF	Download

The English F16 file is the highest-quality Cohere option and preserves the existing default. The same repository also provides smaller quantized files, including cohere-transcribe-q4_k.gguf; pass its exact local path if you choose that variant. All three upstream model families are Apache 2.0 licensed.

For automatic English/Chinese routing, pass both ASR paths. The language probe runs first, and only the matching model is loaded:

--english-model models\cohere-transcribe.gguf
--chinese-model models\qwen3-asr-1.7b-q4_k.gguf
--lid-backend firered --lid-model models\firered-lid-q2_k.gguf

For an explicit english or chinese profile, --model remains available as a single-model override.

CrispASR binary management

The tool auto-detects, installs, and updates the CrispASR binary from GitHub releases.

Flag	Effect
`--install-crispasr`	Download latest platform binary to `bin/`
`--update-crispasr`	Upgrade to newest release
`--crispasr-status`	Show installed version + update availability
`--crispasr-bin-dir PATH`	Custom directory (default `./bin`)
`--crispasr-bin PATH`	Exact path to `crispasr.exe`

When --manage-server is set and no binary is found, it auto-installs before starting the server.

GPU detection

On install and update, the tool checks your hardware:

CUDA -- nvidia-smi available, or CUDA_PATH / CUDA_HOME set, or CUDA in PATH -> downloads crispasr-*-cuda variant.
Vulkan -- vulkaninfo or VULKAN_SDK set (only when CUDA is absent) -> downloads crispasr-*-vulkan variant.
CPU -- fallback when no GPU toolkit is detected.

macOS always uses the universal binary.

Profiles

Profile	Backend	ASR model	Language hint
`english`	`cohere`	Cohere Transcribe 03-2026	`en`
`chinese`	`qwen3-1.7b`	Qwen3-ASR 1.7B	`zh`
`auto`	determined by LID	determined by LID	detected

auto mode runs FireRed language detection on the media, then routes English to Cohere or Chinese to Qwen3-1.7B. Mixed or uncertain content stops with a clear error asking you to re-run with --profile english or --profile chinese.

Usage

Managed server (tool starts CrispASR for you)

uv run python scripts/transcribe.py sample.wav `
  --profile auto `
  --manage-server `
  --english-model models\cohere-transcribe.gguf `
  --chinese-model models\qwen3-asr-1.7b-q4_k.gguf `
  --lid-backend firered --lid-model models\firered-lid-q2_k.gguf `
  --format srt `
  --out-dir outputs

Add --keep-server to leave the server running after transcription.

Manual server (you start CrispASR)

# Terminal 1 -- start the server
crispasr --server --backend cohere `
  -m models\cohere-transcribe.gguf `
  --port 8080

# Terminal 2 -- transcribe
uv run python scripts/transcribe.py sample.mp4 `
  --profile english `
  --server-url http://127.0.0.1:8080 `
  --format verbose_json

If the running server's backend doesn't match the selected profile, the tool prints the exact command you need to start the correct server.

Output formats

`--format`	File extension	Contents
`text`	`.txt`	Plain transcript
`verbose_json`	`.json`	Full response with segments
`srt`	`.srt`	SubRip subtitles
`vtt`	`.vtt`	WebVTT subtitles

A .metadata.json sidecar is always written alongside the transcript.

Video files

Video files are detected automatically. ffmpeg extracts the audio track to a temporary mono 16 kHz WAV before sending it to CrispASR. The temporary file is deleted when transcription finishes.

All CLI flags

--profile auto|english|chinese
--format text|verbose_json|srt|vtt
--out-dir PATH
--server-url URL
--allow-remote-server
--manage-server
--keep-server
--model PATH               Local GGUF override for an explicit profile
--english-model PATH       Cohere model selected after English detection
--chinese-model PATH       Qwen3-ASR model selected after Chinese detection
--allow-model-auto-download
--lid-model PATH           Local LID model path
--lid-backend firered|silero|ecapa|whisper
--host HOST                Managed server host (default 127.0.0.1)
--port PORT                Managed server port (default 8080)
--language CODE            Language hint for transcription
--prompt TEXT              Initial prompt/context
--vad                      Enable voice activity detection
--diarize                  Enable speaker diarization
--diarize-method METHOD
--hotwords WORD,WORD       Comma-separated hotwords
--no-timestamps
--preprocess auto|always|never
--api-key KEY              If CRISPASR_API_KEYS is enabled
--crispasr-bin-dir PATH
--crispasr-bin PATH
--install-crispasr
--update-crispasr
--crispasr-status

MCP server

uv sync --extra mcp
uv run --extra mcp crispasr-agent-mcp

Exposed tools:

Tool	Description
`crispasr_health`	Check CrispASR server health
`crispasr_backends`	List available backends
`crispasr_detect_language`	Run language detection on a file
`transcribe_audio`	Transcribe an audio file
`transcribe_video`	Transcribe a video file
`transcribe_folder`	Batch-transcribe a folder

Security model

No cloud uploads. Media files stay on the local filesystem.
No remote servers by default. --server-url only accepts localhost unless --allow-remote-server is explicitly passed.
No URL inputs. Only local file paths are accepted. URLs, S3, and other remote schemes are rejected.
No shell injection. ffmpeg is called with argument lists and shell=False. No user-controlled strings are interpolated into shell commands.
No model downloads by default. CrispASR model auto-download (-m auto) requires --allow-model-auto-download. The same guard applies to language detection models.
Temporary files are cleaned up. Converted WAV files and LID probe windows are deleted when transcription finishes.
Binary downloads are explicit. CrispASR binary installs only from the official CrispStrobe/CrispASR GitHub releases.
Verified plugin releases. The npm installer requires the plugin ZIP to match the SHA-256 value published in the same GitHub Release.
Narrow installer writes. The installer manages only its plugin directory and the named Personal marketplace entry. Updates preserve local models, binaries, and outputs.

Verify

uv run pytest
uv run ruff check .  # zero lint warnings

License

This project is licensed under the MIT License.

Third-party components and attribution

This tool orchestrates several independently-licensed projects. It does not bundle, fork, or redistribute their code -- it downloads pre-built binaries and calls them as subprocesses or HTTP services at runtime.

Component	License	Role
CrispASR	MIT	ASR engine, server, language detection
ffmpeg	LGPL 2.1+ / GPL 2+	Media decoding and audio extraction
Cohere Transcribe 03-2026	Apache 2.0	English ASR model (loaded by CrispASR)
Qwen3-ASR 1.7B	Apache 2.0	Chinese ASR model (loaded by CrispASR)
FireRed LID	Apache 2.0	Language detection model (loaded by CrispASR)
httpx	BSD	HTTP client for CrispASR API
MCP Python SDK	MIT	MCP server framework
Node.js	MIT	npm installer runtime
adm-zip	MIT	Verified plugin ZIP extraction

Model files must be downloaded separately by the user from their respective HuggingFace repositories. See Required models above.

Related projects

CrispASR -- the ASR engine this tool wraps
CrisperWeaver -- CrispASR's desktop GUI (not used by this tool)

Reviews

No reviews yet

Be the first to review this server!

More Developer Tools MCP Servers

Fetch

Free

by Modelcontextprotocol · Developer Tools

Web content fetching and conversion for efficient LLM usage

Toleno

Free

by Toleno · Developer Tools

Toleno Network MCP Server — Manage your Toleno mining account with Claude AI using natural language.

mcp-creator-python

Free

by mcp-marketplace · Developer Tools

Create, build, and publish Python MCP servers to PyPI — conversationally.

Crispasr Agent Transcriber MCP Server

About

Security Report

Findings (2)

Permissions Required

How to Install

Documentation

crispasr-agent-transcriber

What it does

Quick install for Codex

1. Install prerequisites

2. Run the installer

3. Add the local models

4. Enable the plugin

Update or uninstall

Direct command-line use

Use with other AI agents

Maintainer publishing

Required models

CrispASR binary management

GPU detection

Profiles

Usage

Managed server (tool starts CrispASR for you)

Manual server (you start CrispASR)

Output formats

Video files

All CLI flags

MCP server

Security model

Verify

License

Third-party components and attribution

Related projects

Reviews

No reviews yet

More Developer Tools MCP Servers

Fetch

Toleno

mcp-creator-python

MarkItDown

FinAgent

mcp-creator-typescript