Use this tool when you need to measure the uncertainty of a probability model, particularly in natural language processing tasks. It solves problems related to evaluating language models, such as assessing their ability to predict a test set. The perplexity tool takes a language model and a test dataset as input and outputs a score representing the model's performance, providing context for model selection and optimization.
Cognium trust score
30%
Tier
Unverified
Composite of vulnerability cleanliness, spec conformance, provenance, stability, and usage signals — scanned and weighted by Cognium. Human and agent signals are tracked separately.
Last scanned 2026-09-02.
Returns 7 tools: search_skills, get_skill, list_leaderboard, get_trust_breakdown, resolve_composition, plus the ChatGPT-connector search and fetch. Every tool is annotated read-only.
Resolve this skill directly via MCP tools/call get_skill.