AXe Skills HubSearch /

← All skills

NVIDIA_LocateAnything-3B_vision-language_grounding_model

CH ClawHub 

NVIDIA LocateAnything-3B vision-language grounding model. Covers inference API (detect/ground/point/detect_text/ground_gui), data preparation (JSONL+Recipe 8...

Definition

The catalogue holds this skill’s description, not a full SKILL.md — no upstream address was recorded at ingestion, so the body cannot be fetched.

NVIDIA LocateAnything-3B vision-language grounding model. Covers inference API (detect/ground/point/detect_text/ground_gui), data preparation (JSONL+Recipe 8...

Metadata

Tier
community
Version
1.0.0

Use with an agent

Fetch this skill’s definition over the open API — no key required.

curl -s /v1/skills/community__axehub:ClawHub__ClawHub__NVIDIA_LocateAnything-3B_vision-language_grounding_model