io.github.forevercrab321-svg/leevarai

LEEVAR reliability battery

Verified · today

Grade an AI agent's transcripts on 18 reliability tests. Thin evidence is NOT TESTED, not guessed.

Our take

Evaluates AI-agent transcripts against a battery of reliability tests and explicitly distinguishes insufficient evidence from failed behavior. It is for agent developers, evaluators, and teams doing qualitative reliability review. The concept is useful but specialized, and the lack of adoption signals means its rubric quality and repeatability should be validated before using it for consequential scoring.

reviewed by hand · 2026-09-28

Something wrong or dead here? Report it

Verification record

last verified
today
github stars
0
last commit
today
archived
no
license
MIT
in registry since
2026-09-28

github.com/forevercrab321-svg/leevar-battery

claude-flowai

claude-flow

Verified · 10 days ago

AI orchestration with hive-mind swarms, neural networks, and 87 MCP tools for enterprise dev.

hand-reviewed73k stars19k dl/wkchecked 10 days ago

tldrawai

tldraw

Verified · 11 days ago

Draw and visually collaborate with your agents on tldraw's canvas.

hand-reviewed50k starschecked 11 days ago

codebase-memory-mcpai

Codebase Memory

Verified · yesterday

Codebase knowledge graph for AI agents — 162 languages, sub-ms queries, 99% fewer tokens.

hand-reviewed45k stars12k dl/wkchecked yesterday

openmetadata-mcpai

OpenMetadata

Verified · 29 days ago

Official OpenMetadata MCP: 21 read and write tools for search, lineage, data quality, governance.

hand-reviewed15k starschecked 29 days ago

mobilerunai

mobilerun

Verified · 13 days ago

Control real Android and iOS devices with LLM agents — tap, swipe, type, automate flows.

hand-reviewed9.4k starschecked 13 days ago

praisonaiai

PraisonAI

Verified · today

AI Agents Framework with Self Reflection and MCP support

hand-reviewed9.1k starschecked today