NVIDIA imagereferring-expressionkittibounding-boxesauto-labelvlm
Four-step image referring-expression pipeline: turns images plus KITTI bounding-box labels into region descriptions, scene captions, grounded referring expressions, and (optionally) verified expressions via VLM distillation. Use when the user wants to generate referring-expres…
The catalogue holds this skill’s description, not a full SKILL.md — no upstream address was recorded at ingestion, so the body cannot be fetched.
Four-step image referring-expression pipeline: turns images plus KITTI bounding-box labels into region descriptions, scene captions, grounded referring expressions, and (optionally) verified expressions via VLM distillation. Use when the user wants to generate referring-expres…Fetch this skill’s definition over the open API — no key required.
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__tao-generate-referring-expressions