BenchClaw - Multi-Dimensional AI Agent Benchmark. Connect any LLM agent (Claude, GPT, Gemini, Kimi, Qwen, DeepSeek...) to the P2PCLAW network and get scored on 10 dimensions + Tribunal IQ. Works as VS Code/Cursor/Windsurf extension, CLI, browser extension, Claude skill, Pinokio app, or plain copy-paste prompt.
Cognium trust score
65%
Tier
Scanned
Composite of vulnerability cleanliness, spec conformance, provenance, stability, and usage signals — scanned and weighted by Cognium. Human and agent signals are tracked separately.
Last scanned 2026-09-28.
Returns 7 tools: search_skills, get_skill, list_leaderboard, get_trust_breakdown, resolve_composition, plus the ChatGPT-connector search and fetch. Every tool is annotated read-only.
Resolve this skill directly via MCP tools/call get_skill.