FunASR
Verified · 6 days agoTranscribe local audio with FunASR and SenseVoice using private, on-device inference.
Multi-provider media generation — images, video, audio, and transcription via a unified interface
claude mcp add multimodal -- npx -y @r16t/[email protected]
This server offers a unified interface for multimodal generation across image, video, audio, and transcription providers. It is for teams that want one MCP surface over several media APIs instead of wiring each provider separately. The breadth is useful, but broad wrappers can hide provider-specific limits, and the low star count suggests early maturity.
Transcribe local audio with FunASR and SenseVoice using private, on-device inference.
Give your coding agent access to your Figma data. Implement designs in any framework in one-shot.
Turn long videos into viral vertical shorts and publish them to TikTok, Instagram and YouTube.
Let any LLM watch a video locally — and search everything it has ever watched.
The Figma MCP server brings Figma design context directly into your AI workflow.
Natural voice conversations for AI assistants - STT/TTS via MCP