github verified Safe content atomic container

llm-eval-search

llm-eval-search — fvahedian-llm-eval-search. Use this tool when you need to evaluate the performance of large language models (LLMs) using a search-based approach, solving problems related to model benchmarking and comparison. It takes in model configurations and evaluation metrics as inputs and outputs performance scores and rankings. Ideal for use cases where accurate model assessment is crucial, such as in natural language processing and machine learning applications.

Cognium trust score
75%
Tier
Verified

Composite of vulnerability cleanliness, spec conformance, provenance, stability, and usage signals — scanned and weighted by Cognium. Human and agent signals are tracked separately. Last scanned 2026-09-28.

Scan details: Circle-IR · 2026-09-28 · Appeal

View full trust & usage report →

Metadata

Version
1.0.0
Skill type
atomic
Execution layer
container
Source
GitHub
Author type
human
Last scanned
2026-09-28
Updated
2026-09-28
View source Find related skills

Use via MCP

MCP

Resolve llm-eval-search from your agent

Streamable HTTP transport at https://api.skillsregistry.net/mcp. No auth for read tools. Discovery: .well-known/mcp.json.

One command in your shell — Claude Code wires it up and verifies the connection. Run /mcp in any session to confirm.

claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
Swap --scope user for --scope project to commit it to .mcp.json.

Search SkillsRegistry