FunASR
Verified · 6 days agoTranscribe local audio with FunASR and SenseVoice using private, on-device inference.
Turn any LLM multimodal; generate images, voices, videos, 3D models, music, and more.
claude mcp add --transport http rostro https://proto.rostro.dev/mcp
A multimodal generation server that aims to add images, voice, video, music, and 3D model creation behind one interface. It is for users who want broad media-generation coverage without stitching together separate tools. The ambition is high, but the very wide scope and lack of visible adoption signals make it look more experimental than mature.
Transcribe local audio with FunASR and SenseVoice using private, on-device inference.
Give your coding agent access to your Figma data. Implement designs in any framework in one-shot.
Turn long videos into viral vertical shorts and publish them to TikTok, Instagram and YouTube.
Let any LLM watch a video locally — and search everything it has ever watched.
The Figma MCP server brings Figma design context directly into your AI workflow.
Natural voice conversations for AI assistants - STT/TTS via MCP