FunASR
Verified · 21 days agoTranscribe local audio with FunASR and SenseVoice using private, on-device inference.
Give an AI agent eyes for video: turn a clip into a numbered frame grid + transcript.
claude mcp add vidgrid -- uvx vidgrid-mcp
This server turns video into a numbered frame grid plus transcript, giving agents a lightweight way to inspect clips without full video-native tooling. It is for media analysis, QA, annotation, and multimodal workflows that need a quick visual index of a video. The concept is practical and understandable, but the zero-star footprint suggests an early-stage utility rather than a widely tested media stack component.
Transcribe local audio with FunASR and SenseVoice using private, on-device inference.
Give your coding agent access to your Figma data. Implement designs in any framework in one-shot.
Turn long videos into viral vertical shorts and publish them to TikTok, Instagram and YouTube.
Trim, convert, resize, compress, and remix audio and video.
Screen recording, meeting notes, and voice dictation - all with AI
Let any LLM watch a video locally — and search everything it has ever watched.