ClawHub
NVIDIA LocateAnything-3B vision-language grounding model. Covers inference API (detect/ground/point/detect_text/ground_gui), data preparation (JSONL+Recipe 8...
The catalogue holds this skill’s description, not a full SKILL.md — no upstream address was recorded at ingestion, so the body cannot be fetched.
NVIDIA LocateAnything-3B vision-language grounding model. Covers inference API (detect/ground/point/detect_text/ground_gui), data preparation (JSONL+Recipe 8...Fetch this skill’s definition over the open API — no key required.
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__NVIDIA_LocateAnything-3B_vision-language_grounding_model