API Development stdio

FunASR

Created by modelscope

Provides an industrial-grade, end-to-end speech recognition toolkit featuring ultra-fast transcription, multi-language support, speaker diarization, and emotion detection.

How to Connect & Install

Select your client or environment below to copy the verified, ready-to-run installation configuration.

1-line autonomous agent setup instruction. Paste into Claude Code, Codex CLI, or Cursor Agent:
Fetch and configure Model Context Protocol (MCP) server FunASR from https://www.fastmcp.dev/server/funasr

Agents automatically discover, install, and execute the server spec without manual token setup.

Connect via direct Stateless MCP v2 / Streamable HTTP endpoint or JSON-RPC:
https://github.com/modelscope/funasr

Stateless edge endpoint compliant with MCP 2024-11-05 and v2 draft protocol specs.

Includes server instructions and tool definitions for Claude:
claude mcp add funasr -- npx -y funasr

Run in terminal to register server in current project or pass --scope user for global availability.

Configures server instructions and tools for OpenAI Codex & ChatGPT Desktop:
[mcp_servers.funasr] command = "npx" args = ["-y","funasr"]

Applies to ChatGPT Desktop app, Codex CLI, and IDE extension across your Codex host.

Add to Cursor project settings (.cursor/mcp.json) or Cursor Settings > Features > MCP:
{ "mcpServers": { "funasr": { "type": "stdio", "command": "npx", "args": [ "-y", "funasr" ] } } }

Commit to version control to allow teammates to connect immediately.

Install as an autonomous Agent Skill using the standard skills CLI:
npx skills add https://github.com/modelscope/funasr

Standard agent skill installer compatible with Claude Code, Cursor, and agy harnesses.

Server Guidance & Instructions OpenAI & Claude Spec Compliant
FunASR: Provides an industrial-grade, end-to-end speech recognition toolkit featuring ultra-fast transcription, multi-language support, speaker diarization, and emotion detection. Key capabilities: Support for 50+ Languages; 170x Realtime Speech Recognition (GPU).

Codex and Claude read this instruction prompt during initialization to orchestrate cross-tool workflows and rate limits.

Top Core Functions
01 Support for 50+ Languages
02 170x Realtime Speech Recognition (GPU)
03 Emotion Detection and Audio Events