ClawHub
Google's multimodal model with strong long-context and vision understanding. Suitable for document parsing, image/video understanding, and structured output....
The catalogue holds this skill’s description, not a full SKILL.md — no upstream address was recorded at ingestion, so the body cannot be fetched.
Google's multimodal model with strong long-context and vision understanding. Suitable for document parsing, image/video understanding, and structured output....Fetch this skill’s definition over the open API — no key required.
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__多模态大模型_Gemini_3.1