FunASR
Verified · 6 days agoTranscribe local audio with FunASR and SenseVoice using private, on-device inference.
Image, audio and video: generation, transcription, editing, and platform APIs like YouTube or Spotify. Most generation servers proxy a paid API — check whose key you're burning before wiring one into a loop.
Transcribe local audio with FunASR and SenseVoice using private, on-device inference.
Give your coding agent access to your Figma data. Implement designs in any framework in one-shot.
Turn long videos into viral vertical shorts and publish them to TikTok, Instagram and YouTube.
Let any LLM watch a video locally — and search everything it has ever watched.
The Figma MCP server brings Figma design context directly into your AI workflow.
Natural voice conversations for AI assistants - STT/TTS via MCP
Extract brand assets (logos, colors, backdrop images, brand name) from any website URL
The AI-powered toolkit that grows your YouTube channel on autopilot
MCP server + Claude Code plugin for ComfyUI: run workflows, generate images, manage models & VRAM.
An MCP server retrieving transcripts of YouTube videos
An MCP server that provides image generation and editing capabilities
Render PyTorch architecture diagrams and animated GIF reveals from trusted model source.
MCP server for Adobe Photoshop — 102 tools (generative AI + recipes), standalone web UI. Control Ph…
Public Spotify metadata, lyrics and podcasts for LLM agents. No API key, read-only.
Publish videos to TikTok, Reels, Shorts, X, and Facebook through Taisly.
MCP server giving AI agents eyes on OpenImageDebugger buffers in live gdb/lldb sessions
Hand any social video to your AI agent — frames + transcript bundled. MCP server for watch-cli.
Bidirectional Figma MCP — AI draws UI on Figma canvas, reads designs back
Generate images, video, and audio with Glif's media-generation agent
Parse PDFs, images, doc, docx, ppt, pptx, xls, xlsx, html into Markdown using MinerU API.
Automate Google NotebookLM — Q&A with citations, audio, video, content generation
AI image generation and editing with prompt optimization and quality presets
Chat-driven AI video montage: beat-synced cuts, auto-matched music, subtitles, transitions.
Image processing pipeline for Next.js. Responsive optimization with Sharp.
Guardrailed video editing MCP server: FFmpeg tools for subtitles, audio, effects, Hyperframes.
Guardrailed video editing for AI agents: FFmpeg, captions, effects, Hyperframes, and receipts.
Text-to-speech, speech-to-text, audio-to-face lipsync, and motion-capture clips for 3D agents.
Apple Music MCP server: playlists, library, catalog, playback and Up Next, on Mac/Windows/Linux.
MCP server for programmatic image comparison and visual diff generation.
Chat with 300+ LLMs via OpenRouter. Analyze and generate images, audio, and video from MCP.
MCP Server for Video Jungle - Analyze, Search, Generate, and Edit Videos
MCP server to convert Figma designs to Flowbite UI components in Tailwind CSS
MCP server for Draw Things - local AI image generation on Mac
MCP server for OpenAI Images/Videos and Google GenAI (Veo) media generation.
Access Apple Voice Memos on macOS. List, get audio, extract and generate transcripts.
MCP server for AI-powered image recognition and description using OpenAI vision models.
MCP server for AI-powered image recognition and description using OpenAI vision models.
Local audio transcription using whisper.cpp. Transcribe with OpenAI Whisper models.
FFmpeg video/audio tools: cut, convert, concat, remove silence, and raw commands.
Turn any LLM multimodal; generate images, voices, videos, 3D models, music, and more.
InChat Image Viewer MCP - View images inline in AI chat by just providing the file path
PDF reader for vision LLMs. Auto-detects text corruption and switches to image mode.
SynClub MCP Server for AI-powered comic creation with script generation and image tools
Convert LaTeX math expressions to crisp, scalable SVG images
TikTok video data analytics and content strategy tools