claude-flow
Verified · 2 days agoAI orchestration with hive-mind swarms, neural networks, and 87 MCP tools for enterprise dev.
Generate QA datasets & evaluate RAG systems with failure diagnosis. Any LLM.
claude mcp add ragscore -- uvx ragscore
RAG evaluation in two commands: generate QA pairs from your documents, fire them at your RAG endpoint, and get accuracy scores with per-question failure diagnosis. Works with any LLM including local Ollama, which keeps evaluation private, and the notebook API with built-in plots makes it genuinely quick to audit a pipeline. Young project from a new org — the methodology is LLM-as-judge with all the caveats that implies, so read the scored failures rather than trusting the single headline number.
AI orchestration with hive-mind swarms, neural networks, and 87 MCP tools for enterprise dev.
Draw and visually collaborate with your agents on tldraw's canvas.
Codebase knowledge graph for AI agents — 159 languages, sub-ms queries, 99% fewer tokens.
Official OpenMetadata MCP: governed context and business semantics for AI assistants and agents.
Control real Android and iOS devices with LLM agents — tap, swipe, type, automate flows.
AI Agents Framework with Self Reflection and MCP support