FunASR
Verified · 6 days agoTranscribe local audio with FunASR and SenseVoice using private, on-device inference.
Give an AI agent eyes for video: turn a clip into a numbered frame grid + transcript.
claude mcp add vidgrid -- uvx vidgrid-mcp
This server turns video into a numbered frame grid plus transcript, giving agents a lightweight way to inspect clips without full video-native tooling. It is for media analysis, QA, annotation, and multimodal workflows that need a quick visual index of a video. The concept is practical and understandable, but the zero-star footprint suggests an early-stage utility rather than a widely tested media stack component.
Transcribe local audio with FunASR and SenseVoice using private, on-device inference.
Give your coding agent access to your Figma data. Implement designs in any framework in one-shot.
Turn long videos into viral vertical shorts and publish them to TikTok, Instagram and YouTube.
Let any LLM watch a video locally — and search everything it has ever watched.
The Figma MCP server brings Figma design context directly into your AI workflow.
Natural voice conversations for AI assistants - STT/TTS via MCP