FunASR · emploidai Marketplace
emploidai Marketplace
Add-onsAppletsPlugins
Search tools, teams, and capabilitiesPublish
MarketplacePluginsFunASR
Plugin
Limited listing

FunASR

by langgenius · v0.1.1

Open-source speech recognition with OpenAI-compatible serving and multilingual models.

399 installsUpdated Jul 24, 2026
Publisher information is incomplete

This community listing does not yet include every recommended support, privacy, pricing, and permission disclosure. Review the available package permissions before installing.

Capabilities

Models

Available inside your emploidai workspace after installation.

Data sources

Available inside your emploidai workspace after installation.

Category

model

Version

0.1.1langgenius

Requirements

Maximum memory 256MB

Pricing

Not disclosed by publisher

Security & access

Review before installing

CompatibleRequires emploidai 1.0.0+

Permissions

  • Uses model capability
  • Uses tool capability

Dependencies

No additional dependencies

emploidai Marketplace

Discover capabilities. Review access. Install inside your workspace.

DocumentationSecuritySupportPrivacyTerms

FunASR Speech Recognition Plugin

Self-hosted speech recognition through FunASR's OpenAI-compatible API. The plugin supports multilingual ASR models for private Dify deployments without sending audio to a hosted transcription service.

In FunASR's 192-minute benchmark, SenseVoiceSmall reached 170x real-time on GPU and 17x real-time on CPU. Whisper-large-v3 reached 13x real-time on GPU in the same benchmark. Results depend on the hardware, audio, model, and concurrency settings.

Setup

  1. Deploy a SenseVoice server:
pip install "funasr>=1.3.26" fastapi uvicorn python-multipart
funasr-server --device cuda --model sensevoice --host 0.0.0.0 --port 8000

Use --device cpu when CUDA is unavailable. Fun-ASR-Nano deployments use vLLM and can be installed separately:

pip install "funasr>=1.3.26" vllm fastapi uvicorn python-multipart
funasr-server --device cuda --model fun-asr-nano --host 0.0.0.0 --port 8000
  1. In Dify, configure this plugin with the server URL, for example http://your-server:8000 or http://your-server:8000/v1. The plugin normalizes both forms to a base URL ending in /v1, such as http://localhost:8000/v1. Leave the API key empty when the server does not require one.

Predefined models accept audio files up to 25 MB.

Supported Models

  • sensevoice - Mandarin, Cantonese, English, Japanese, and Korean ASR with emotion and audio-event tags (default)
  • paraformer - Chinese ASR with punctuation
  • paraformer-en - English ASR
  • fun-asr-nano - Chinese, English, Japanese, and Chinese dialect/accent ASR with an encoder + LLM architecture