AXe Skills HubSearch /

← All skills

Model_Watch

CH ClawHub benchmarkclidegradationllmmonitoringquality

Benchmark AI API models over time and detect quality degradation. 7 standardized tests (reasoning, coding, writing, instruction-following, hallucination). Al...

Definition

The catalogue holds this skill’s description, not a full SKILL.md — no upstream address was recorded at ingestion, so the body cannot be fetched.

Benchmark AI API models over time and detect quality degradation. 7 standardized tests (reasoning, coding, writing, instruction-following, hallucination). Al...

Metadata

Category
Software Dev
Tier
community
Version
1.0.0

Use with an agent

Fetch this skill’s definition over the open API — no key required.

curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Model_Watch