claude-flow
Verified · 10 days agoAI orchestration with hive-mind swarms, neural networks, and 87 MCP tools for enterprise dev.
Grade an AI agent's transcripts on 18 reliability tests. Thin evidence is NOT TESTED, not guessed.
Evaluates AI-agent transcripts against a battery of reliability tests and explicitly distinguishes insufficient evidence from failed behavior. It is for agent developers, evaluators, and teams doing qualitative reliability review. The concept is useful but specialized, and the lack of adoption signals means its rubric quality and repeatability should be validated before using it for consequential scoring.
AI orchestration with hive-mind swarms, neural networks, and 87 MCP tools for enterprise dev.
Draw and visually collaborate with your agents on tldraw's canvas.
Codebase knowledge graph for AI agents — 162 languages, sub-ms queries, 99% fewer tokens.
Official OpenMetadata MCP: 21 read and write tools for search, lineage, data quality, governance.
Control real Android and iOS devices with LLM agents — tap, swipe, type, automate flows.
AI Agents Framework with Self Reflection and MCP support