AXe Skills HubSearch /

← All skills

tao-generate-referring-expressions

NV NVIDIA imagereferring-expressionkittibounding-boxesauto-labelvlm

Four-step image referring-expression pipeline: turns images plus KITTI bounding-box labels into region descriptions, scene captions, grounded referring expressions, and (optionally) verified expressions via VLM distillation. Use when the user wants to generate referring-expres…

Definition

The catalogue holds this skill’s description, not a full SKILL.md — no upstream address was recorded at ingestion, so the body cannot be fetched.

Four-step image referring-expression pipeline: turns images plus KITTI bounding-box labels into region descriptions, scene captions, grounded referring expressions, and (optionally) verified expressions via VLM distillation. Use when the user wants to generate referring-expres…

Metadata

Category
Vision AI
Tier
community
Version
1.0.0

Use with an agent

Fetch this skill’s definition over the open API — no key required.

curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__tao-generate-referring-expressions