FunASR
Verified · 6 days agoTranscribe local audio with FunASR and SenseVoice using private, on-device inference.
PDF reader for vision LLMs. Auto-detects text corruption and switches to image mode.
claude mcp add pdf4vllm -- uvx pdf4vllm-mcp
This server is a PDF ingestion tool for vision-capable language models, with logic to detect text corruption and fall back to image-based handling. It is for LLM pipeline builders who need more resilient document input than plain text extraction provides. The problem it addresses is real and specific, but the package appears early-stage from this listing and its quality will depend heavily on how well that fallback logic performs in practice.
Transcribe local audio with FunASR and SenseVoice using private, on-device inference.
Give your coding agent access to your Figma data. Implement designs in any framework in one-shot.
Turn long videos into viral vertical shorts and publish them to TikTok, Instagram and YouTube.
Let any LLM watch a video locally — and search everything it has ever watched.
The Figma MCP server brings Figma design context directly into your AI workflow.
Natural voice conversations for AI assistants - STT/TTS via MCP