FunASR
Verified · 23 days agoTranscribe local audio with FunASR and SenseVoice using private, on-device inference.
Evidence-first image MCP. Agent Media Twin with metadata and citeable OCR evidence.
claude mcp add image-reader-mcp -- npx -y @sylphx/[email protected]
This server is built for extracting structured information from images, with emphasis on OCR, metadata, and evidence-linked outputs. It is best suited to document-heavy or verification-sensitive workflows where an agent needs traceable image reading rather than generic vision summaries. The positioning is specific and potentially useful, but the lack of adoption signals suggests a niche tool that should be tested carefully before broader use.
Transcribe local audio with FunASR and SenseVoice using private, on-device inference.
Give your coding agent access to your Figma data. Implement designs in any framework in one-shot.
Turn long videos into viral vertical shorts and publish them to TikTok, Instagram and YouTube.
Trim, convert, resize, compress, and remix audio and video.
Screen recording, meeting notes, and voice dictation - all with AI
Let any LLM watch a video locally — and search everything it has ever watched.