Search: Vision-Language
100 result(s) on page 1
Vision
See and understand images when you (the current model) have no native vision. Use this WHENEVER you need to look at, read, describe, OCR, or reason about the...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.4
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Visionvision
Indexed by skills.sh from kotot/vision
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- kotot
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__visionNVIDIA_LocateAnything-3B_vision-language_grounding_model
NVIDIA LocateAnything-3B vision-language grounding model. Covers inference API (detect/ground/point/detect_text/ground_gui), data preparation (JSONL+Recipe 8...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__NVIDIA_LocateAnything-3B_vision-language_grounding_modelblip-2-vision-language
Indexed by skills.sh from davila7/claude-code-templates
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.5
- Author
- davila7
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__blip-2-vision-languageTrio_Stream_Vision
Analyze any YouTube livestream or RTSP camera feed using natural language — ask what's happening, detect specific events, or get periodic summaries. Powered...
quarantinedaicameralivestreamClawHub- Registry
- ClawHub
- Category
- AI Agents
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Trio_Stream_VisionComputer_Vision_Expert
SOTA Computer Vision Expert (2026). Specialized in YOLO26, Segment Anything 3 (SAM 3), Vision Language Models, and real-time spatial analysis.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Computer_Vision_Expertmar-computer-vision-expert
SOTA Computer Vision Expert (2026). Specialized in YOLO26, Segment Anything 3 (SAM 3), Vision Language Models, and real-time spatial analysis.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__mar-computer-vision-expertuniversal-pdf-vision-parser
Extract multilingual document content and language learning notes (French, German, Japanese, Spanish, etc.) from PDFs using multimodal vision (Qwen-VL-Max)....
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__universal-pdf-vision-parserLanguage
A comprehensive AI agent skill for language learners at every level. Builds personalized study plans, teaches vocabulary in context, corrects your writing an...
quarantinedfluencylanguagelearningClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Languagelanguage
A1 Level English Conversation Partner Bot: Engage, Correct, and Build Confidence.
quarantinedenglish-learningconversation-practicelanguage-supportLobeHub- Registry
- LobeHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:LobeHub__LobeHub__languageAgent_Paddleocr_Vision
Multi-language document understanding with PaddleOCR
quarantinedagent-actionsbatchdocument-understandingClawHub- Registry
- ClawHub
- Category
- Productivity· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Agent_Paddleocr_VisionVision_Bot
Describe images, detect objects, extract text, and analyze webpages. Pass any image URL directly in your task. Responds in your language.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Vision_Bottao-finetune-clip
CLIP vision-language model for image-text retrieval, zero-shot classification, embedding extraction, ONNX export, and TensorRT deployment. Use when fine-tuning or training CLIP, running zero-shot classification, computing image embeddings, or deploying CLIP to ONNX/TensorRT.
quarantinedvision-languageclassificationembeddingNVIDIA- Registry
- NVIDIA
- Category
- Vision AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__tao-finetune-clipMiniMax_Vision_Analysis
Analyze, describe, and extract information from images using the MiniMax vision MCP tool. Use when: user shares an image file path or URL (any message contai...
quarantinedanalysisimagemediaClawHubchina-vision
多模态图片理解工具。Use when user wants to analyze, describe, or understand images using AI vision models. Supports scene analysis, object recognition, chart interpret...
quarantinedaianalysischinaClawHub- Registry
- ClawHub
- Category
- AI Agents
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__china-visionllava
Vision-language chat: VQA, captioning, image dialogue.
quarantinedlinuxmacoswindowsoptional- Registry
- optional
- Category
- MLOps
- Version
- 1.0.0
- Author
- Orchestra Research
- License
- MIT
curl -s /v1/skills/community__axehub:optional__optional__llavaMoltShell_Vision_Engine
Give your text-based OpenClaw agent the ability to see and describe images
quarantinedimage-to-textmoltshellscraperClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__MoltShell_Vision_EngineTrio_Vision
Turn any live camera into a smart camera — describe what to watch for in plain English, get alerts in your chat when it happens. Ask questions about any live...
quarantinedaicameramonitoringClawHub- Registry
- ClawHub
- Category
- AI Agents
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Trio_VisionVision_Tool
Image recognition using Ollama + qwen3.5:4b with think=False for reliable content extraction.
quarantinedimage-recognitionollamavisionClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Vision_Toolxiaonghongshu-vision
You can use this agent combined with multimodal models to upload images and generate Xiaohongshu-style copywriting.
quarantinedvisionLobeHub- Registry
- LobeHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:LobeHub__LobeHub__xiaonghongshu-visiontao-finetune-cosmos-embed
Cosmos-Embed1 video-text embedding for text-to-video retrieval, video-to-video search, semantic deduplication, and fine-tuning. Use when the user asks to "fine-tune Cosmos-Embed1", "run cosmos-embed inference", "export Cosmos-Embed1", "embed videos", or "search videos with text".
quarantinedvideovision-languagevlmNVIDIA- Registry
- NVIDIA
- Category
- Vision AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__tao-finetune-cosmos-embeddeepstream-import-vision-model
Use this skill to bring any vision model from HuggingFace or NVIDIA NGC into an NVIDIA DeepStream pipeline with end-to-end automation: ONNX download, SafeTensors export, TRT engine build, custom nvinfer bbox parser, multi-stream benchmark, and PDF report. Object detection mode…
quarantinedNVIDIA- Registry
- NVIDIA
- Category
- Vision AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__deepstream-import-vision-modelASCII_Vision
Fallback image viewer when vision models are unavailable. Converts images to ASCII art via ffmpeg + Python for brightness distribution, texture analysis, edg...
quarantinedClawHub- Registry
- ClawHub
- Category
- Creative· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__ASCII_VisionAgent_Vision_Scraper
Dockerized AI-powered web scraper using Playwright with virtual display and vision-based captcha solving, no third-party captcha services needed.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Agent_Vision_ScraperBaidu_Yijian_Vision
Yijian (一见) is Baidu's specialized vision AI skill for image and video analysis. Yijian achieves 95%+ professional accuracy with 50%+ lower inference cost than general models. Yijian is built for industrial quality inspection, SOP compliance, safety monitoring, and commercial …
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Baidu_Yijian_VisionLocal_GMNCODE_Vision_Pro
Advanced local vision infrastructure for agents when built-in image tools are unavailable or unreliable. Use for batch image analysis, structured JSON output...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Local_GMNCODE_Vision_ProMiao_Vision_Skill
Use when a user asks an agent such as Codex or Claude to turn an article URL, Markdown file, or long-form text into an infographic artifact with Miao Vision, or to visualize a local CSV, TSV, XLSX, or JSON data file, inspect data fields, generate an HTML chart/report, validate…
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Miao_Vision_SkillMiniMax_Vision_Skill
Analyze images using MiniMax CLI (mmx-cli) for vision tasks. Use when user wants image understanding via MiniMax instead of OpenRouter. Triggers on requests...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__MiniMax_Vision_SkillNVIDIA_Kimi_Vision
Analyze images using NVIDIA Kimi K2.5 vision model via NVIDIA NIM API. Perfect for adding vision to non-vision models like MiniMax M2.5, GLM-5, or any model...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__NVIDIA_Kimi_VisionOATDA_Vision_Analysis
Analyze images using vision-capable AI models through OATDA's unified API. Triggers when the user wants to analyze, describe, or understand images; extract t...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__OATDA_Vision_AnalysisOpenRouter_Vision_Agent
Analyze images using OpenRouter's vision API with x-ai/grok-4.1-fast. Requires OPENROUTER_API_KEY env var or user-provided key. Use when the user asks to des...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__OpenRouter_Vision_AgentPdf_Vision
Extract text content from image-based/scanned PDFs using multiple vision APIs with automatic fallback. Supports Xflow (qwen3-vl-plus) and ZhipuAI (GLM-4.6V-F...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Pdf_VisionQwen_Vision
Analyze images and videos using Qwen Vision API (Alibaba Cloud DashScope). Supports image understanding, OCR, visual reasoning.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Qwen_VisionS2-SP-OS_Vision_Cast
S2-SP-OS Vision Cast. Features a universal Protocol Sniffer (AirPlay/Chromecast/DLNA) for native casting, backed by our secure S2 ephemeral push fallback. /...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__S2-SP-OS_Vision_CastScreen_Vision
macOS screen OCR & click automation via Apple Vision + ScreenCaptureKit. Capture any window or screen region, extract text with coordinates, find text, and c...
quarantinedClawHub- Registry
- ClawHub
- Category
- Autonomous Ai Agents· inferred
- Version
- 1.0.5
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Screen_VisionTrendAI_Vision_One_Threat_Intelligence
Query TrendAI Vision One threat intelligence. Use when: looking up IOCs (IP, domain, hash, URL, email), checking threat feeds, reading intelligence reports,...
quarantinedClawHub- Registry
- ClawHub
- Category
- Security· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__TrendAI_Vision_One_Threat_IntelligenceVision_Analyzer
Analyze images using Ollama Cloud's Kimi K2.5 vision capabilities. Use when user wants to describe, understand, or get information about an image. Works with...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Vision_AnalyzerVision_Fallback
Vision/image understanding for agents whose model can't read images (returns "model does not support images", empty/unknown output, low confidence, or user-r...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Vision_FallbackVision_Helper_—_AI_Image_Analysis
Analyze images using local or cloud vision models via Ollama to identify content, UI elements, screenshots, or extract text with OCR support.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Vision_Helper_—_AI_Image_AnalysisVision_Sandbox
Agentic Vision via Gemini's native Code Execution sandbox. Use for spatial grounding, visual math, and UI auditing.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Vision_SandboxVision_Tagger
Tag and annotate images using Apple Vision framework (macOS only). Detects faces, bodies, hands, text (OCR), barcodes, objects, scene labels, and saliency re...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Vision_TaggerZai_Vision
Z.AI Vision analysis using GLM-4.6V model for image and video understanding. Use when Claude needs to analyze images (screenshots, UI designs, photos, diagra...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Zai_Visiondeepstream-import-vision-model
Use this skill to bring any vision model from HuggingFace or NVIDIA NGC into an NVIDIA DeepStream pipeline with end-to-end automation: ONNX download, SafeTensors export, TRT engine build, custom nvinfer bbox parser, multi-stream benchmark, and PDF report. Object detection mode…
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__deepstream-import-vision-modelmiao-vision
Use when a user asks an agent such as Codex or Claude to turn an article URL, Markdown file, or long-form text into an infographic artifact with Miao Vision, or to visualize a local CSV, TSV, XLSX, or JSON data file, inspect data fields, generate an HTML chart/report, validate…
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__miao-visionopenai-vision
Analyze images and multi-frame sequences using OpenAI GPT vision models
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__openai-visionqwencloud-vision
[QwenCloud] Understand images and videos with Qwen vision models. TRIGGER when: user wants to analyze, describe, or extract information from images or videos...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__qwencloud-visionvision-skill
Use this skill for computer vision tasks including image recognition (OCR, object detection) and image generation (text-to-image, image-to-image). Supports a...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__vision-skillyolo-vision-tools
Use Ultralytics YOLO to perform computer vision tasks, such as detecting people or objects in images and videos, classifying images, estimating human poses,...
quarantinedClawHub- Registry
- ClawHub
- Category
- Creative· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__yolo-vision-toolshuggingface-vision-trainer
Trains and fine-tunes vision models for object detection (D-FINE, RT-DETR v2, DETR, YOLOS), image classification (timm models — MobileNetV3, MobileViT, ResNet, ViT/DINOv3 — plus any Transformers classifier), and SAM/SAM2 segmentation using Hugging Face Transformers on Hugging …
quarantinedHuggingFace- Registry
- HuggingFace
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:HuggingFace__HuggingFace__huggingface-vision-trainerclip
Zero-shot image classification and image-text search.
quarantinedlinuxmacoswindowsoptional- Registry
- optional
- Category
- MLOps
- Version
- 1.0.0
- Author
- Orchestra Research
- License
- MIT
curl -s /v1/skills/community__axehub:optional__optional__clipNWO_Robotics
Control robots and IoT devices via natural language using the NWO Robotics API for robot commands, sensor queries, vision tasks, and task planning.
quarantinedactionapiautomationClawHub- Registry
- ClawHub
- Category
- Software Dev
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__NWO_RoboticsAB-Agents-Vision-MiniMax
👁️ Image analysis via MiniMax VL API. Describe images, extract text from screenshots, analyze photos. Requires MiniMax Token Plan API key (free tier availab...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__AB-Agents-Vision-MiniMaxAB_Agents_Vision
👁️ Image analysis using MiniMax VL API. Describe images, extract text from screenshots, analyze photos. Works with local files and URLs. Simple shell wrapper.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__AB_Agents_VisionClaw_Vision
Analyze local images including screenshots, receipts, and documents to extract structured text, UI elements, and provide content summaries with confidence le...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Claw_VisionImage_Vision
Analyze and interpret images by describing content, extracting text, answering questions, comparing visuals, and extracting structured data from JPG, PNG, GI...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.5
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Image_VisionLlava_Vision
Call a local llama.cpp server with the LLaVA model to analyze images.
quarantinedClawHub- Registry
- ClawHub
- Category
- Mlops· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Llava_VisionMedia_Gen_Vision_Video
Generate and analyze images, and generate videos using OpenClaw's preferred Google media workflows. Use when the user asks to create, edit, inspect, compare,...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Media_Gen_Vision_VideoMinimax_Vision_Search
Analyze images and search the web using MiniMax MCP tools
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Minimax_Vision_SearchNtriq_Vision_Product_Analyzer_Mcp
Analyze product images: identify items, extract specs, compare features, generate descriptions.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Ntriq_Vision_Product_Analyzer_McpFor_example,_add_a_2-second_wait_after_`mouse_click`_to_avoid_operation_failure_due_to_slow_system_response.
Professional Windows-only visual automation toolkit with 11 modules for screenshot, OCR, template matching, clicks, input, environment setup, and looping tasks.
quarantinedai-agentopenclawvisual-automationClawHub- Registry
- ClawHub
- Category
- AI Agents
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__OpenClaw_11-in-1_Visual_Automation_Suite_(Windows_Only)_Complete_visual_automation_toolkit_with_11_integrated_modules.__###_💰_Price_One-time_purchase:_**$2.99**_(Lifetime_access_to_all_modules_+_future_updates)__###_🚀_How_to_Purchase_1.__Pay_via_PayPal_Invoice:_____🔗_[Click_to_pay_$2.99](https:____www.paypal.com__invoice__p__#V2RC9S8LVKJ434R9)_2.__After_payment,_send_your_email_to:_**[email protected]**_3.__I_will_send_the_full_download_link_within_12_hours.__###_🖥️_Compatibility_-_Windows_10____11_only_-_Not_compatible_with_macOS____Linux_##_1._Product_Basic_Description_###_1.1_Core_Functions_Provides_professional_universal_computer_vision_automation_capabilities_covering_the_full-process_visual_automation_scenarios_such_as_environment_initialization,_full-screen_automatic_screenshot,_OCR_text_recognition,_template_matching_target_localization,_mouse_click_simulation,_keyboard_input_simulation,_and_complete_environment_initialization_&_cleanup_mechanisms._It_supports_custom_task_combination_and_cyclic_execution.__###_1.2_Version_&_Directory_Description_-_Core_Capability:_Flexible_invocation_based_on_minimum_executable_units,_supporting_parameter_customization,_result_variable_inheritance,_and_custom_skill_saving._All_functions_can_be_used_directly_with_the_`call`_command_right_after_extracting_the_package._-_Directory_Structure:____-_`claw.json`_-_Skill_package_configuration_file___-_`skills__all_skills.claw`_-_All_skill_unit_definitions___-_`templates__`_-_Directory_for_template_images_(place_your_template_images_here_for_matching)___-_Temporary_file_directory_`temp__`_(for_storing_screenshots_like_temp__screen.png)_is_automatically_created_after_executing_`init_env`;_temporary_screenshot_files_can_be_cleaned_up_via_`clean_temp`._-_Version_Info:_Current_version:_1.0.0;_Compatible_with_OpenClaw_>=_1.0.0__###_1.3_Paid_Attribute_This_automation_skill_system_(vision-auto-tool-pro)_is_a_paid_professional_toolkit._The_document_does_not_explicitly_authorize_commercial_use_of_the_toolkit._The_paid_permission_only_covers_basic_usage_(non-commercial_by_default),_and_commercial_use_requires_separate_confirmation_of_authorization_with_the_provider_(e.g.,_purchasing_a_commercial_license,_signing_a_commercial_agreement).__##_2._Complete_Skill_Invocation_Manual_###_Important_Notes_Ensure_sufficient_time_is_reserved_for_the_computer_to_respond_to_each_click_or_operation._For_example,_add_a_2-second_wait_after_`mouse_click`_to_avoid_operation_failure_due_to_slow_system_response.__###_2.1_List_of_All_Minimum_Executable_Units_|_Unit_Name_______________|_Fixed_Call_Name__________|_Function_Description_________________________________________________________________|_Individual_Call_Method__________________________|_|-------------------------|--------------------------|--------------------------------------------------------------------------------------|-------------------------------------------------|_|_Initialize_Environment__|_`init_env`_______________|_Create_directory_structure,_clear_temporary_files,_check_template_directory__________|_`call_init_env`__________________________________|_|_Full_Screen_Screenshot__|_`screenshot_full`________|_Capture_entire_screen_and_save_as_temp__screen.png____________________________________|_`call_screenshot_full`___________________________|_|_Check_Screenshot_Validity_|_`check_screenshot_valid`_|_Check_for_black_screen__freeze,_wake_up_the_interface_if_invalid______________________|_`call_check_screenshot_valid`____________________|_|_Wake_Interface__________|_`wake_window`____________|_Solve_the_problems_of_background_non-rendering_and_black_screenshot___________________|_`call_wake_window`_______________________________|_|_OCR_Recognition_________|_`ocr_recognize`__________|_Recognize_all_text_on_the_screen_and_their_corresponding_coordinates__________________|_`call_ocr_recognize`_____________________________|_|_Template_Matching_______|_`template_match`_________|_Use_template_image_to_match_and_locate_icons__buttons__________________________________|_`call_template_match_category_template_name`_____|_|_Unified_Localization____|_`locate_target`__________|_Prioritize_OCR_positioning;_use_template_matching_if_not_found,_return_coordinates___|_`call_locate_target_target_text_OR_category+template_name`_|_|_Mouse_Click_____________|_`mouse_click`____________|_Move_to_the_specified_coordinates_and_perform_click_operation________________________|_`call_mouse_click_X_Y_[click_type,_default=single_click]`_|_|_Keyboard_Input__________|_`keyboard_input`_________|_Input_text_after_locating_the_input_box______________________________________________|_`call_keyboard_input_target_coords__description_input_content`_|_|_Clean_Temporary_Files___|_`clean_temp`_____________|_Delete_temporary_screenshots_and_free_up_storage_space_______________________________|_`call_clean_temp`________________________________|_|_Loop_Restart____________|_`loop_restart`___________|_Wait_2_seconds_then_go_back_to_the_screenshot_step_and_restart_the_process___________|_`call_loop_restart`______________________________|__###_2.2_Method_for_Invoking_Individual_Units_####_Invocation_Format_```_call_[unit_call_name]_[parameter...]_```_####_Invocation_Examples_-_Initialize_environment:_`call_init_env`_-_Template_match_browser_icon_on_desktop:_`call_template_match_desktop_web`_-_Perform_double-click_at_coordinates_(100,200):_`call_mouse_click_100_200_double`__###_2.3_Combine_into_Custom_New_Tasks_By_writing_one_call_instruction_per_line_in_execution_order,_you_can_combine_them_into_a_custom_new_task,_which_supports_variable_inheritance,_looping,_and_permanent_saving.__####_Format_Example_(Open_Browser)_```_#_Task_Name:_Open_Browser_call_init_env_call_screenshot_full_call_check_screenshot_valid_call_locate_target_browser_desktop_Browser_call_mouse_click_{{resultX}}_{{resultY}}_double_call_clean_temp_```__####_Combination_Steps_1._**Write_task_name_and_description_first**_(for_easier_identification_later)_2._**In_execution_order**,_write_one_`call_unit_name_parameters`_instruction_per_line_3._Coordinates_can_use_variables_`{{resultX}}`__`{{resultY}}`_to_inherit_the_output_result_of_the_previous_unit_4._If_cyclic_execution_is_required,_add_`call_loop_restart`_at_the_end_5._**Save_custom_skill**:_Use_`save_skill_skill_name_instruction_list`_to_save_the_task_permanently,_then_call_it_directly_with_`call_skill_name`__###_2.4_Complete_Main_Flow_Invocation_Example_```_#_General_Main_Flow:_vision_auto_main_call_init_env_call_screenshot_full_call_check_screenshot_valid_call_ocr_recognize_#_If_template_matching_is_needed,_add_this_line:_call_template_match_category_name_call_locate_target_target_text_call_mouse_click_{{X}}_{{Y}}_#_If_text_input_is_needed,_replace_the_above_line_with:_call_keyboard_input_{{X}}_{{Y}}_input_content_call_clean_temp_#_Add_this_line_if_you_need_to_loop:_call_loop_restart_```___###_Important_Notes_Ensure_sufficient_time_is_reserved_for_the_computer_to_respond_to_each_click_or_operation.___>__For_example,_add_a_2-second_wait_after_`mouse_click`_to_avoid_operation_failure_due_to_slow_system_response.Openclaw_mneme_vision
Use local visual memory tools to index and search photos and videos from creator media libraries.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Openclaw_mneme_visionPeripheral_Vision
Monitors adjacent systems, upstream dependencies, and downstream consumers for changes that could affect your current work — before they break it. Like biolo...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Peripheral_VisionTiexue_Vision
Recognizes text (Chinese/English), objects, and scenes in images from chat, documents, or local files, with optional translation and auto-saving results.
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Tiexue_VisionTunnel_Vision_&_Slack
Activate when: user says 'I'm always firefighting,' 'I never have time to think,' 'everything feels urgent,' 'I can't get ahead,' or 'we keep pushing strateg...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Tunnel_Vision_&_SlackVision_Recognition_Ocr
Vehicle/animal/plant recognition plus OCR for screenshots, photos, invoices, and tables. Use when users ask 识别车型/看图识别/提取文字/OCR. Supports local path, URL, and...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Vision_Recognition_OcrVision_Simulator
Marketing strategy prediction with real-world data integration. Predicts campaign outcomes, product launches, and market reactions using multi-agent simulati...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Vision_SimulatorVision_Understanding
Turn images, video, audio, or documents into text. Use when the user says "what's in this image", "describe / caption this", "tag these photos", "read this d...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Vision_Understandinguni-vision-engine
Automated high-quality video generation (text-to-video, image-to-video) via a local jimeng-api Docker service. Features native OpenClaw image interception, a...
quarantinedClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__uni-vision-enginewith-keil-u-vision-5-c-code-explainer
Expert in interpreting embedded C code using Keil uVision 5 and Proteus
quarantinedmicrocontrollerc codeeducationLobeHub- Registry
- LobeHub
- Category
- Research
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:LobeHub__LobeHub__with-keil-u-vision-5-c-code-explaineraxiom-vision
Indexed by skills.sh from charleswiltgen/axiom
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- charleswiltgen
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__axiom-visioncomputer-vision
Indexed by skills.sh from aj-geddes/useful-ai-prompts
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- aj-geddes
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__computer-visioncomputer-vision-expert
Indexed by skills.sh from sickn33/antigravity-awesome-skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- sickn33
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__computer-vision-expertcomputer-vision-opencv
Indexed by skills.sh from mindrally/skills
quarantinedskills.sh- Registry
- skills.sh
- Category
- Autonomous Ai Agents· inferred
- Version
- 1.0.0
- Author
- mindrally
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__computer-vision-opencvdeepstream-import-vision-model
Indexed by skills.sh from nvidia/skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- nvidia
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__deepstream-import-vision-modeldefining-product-vision
Indexed by skills.sh from refoundai/lenny-skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- refoundai
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__defining-product-visionfal-vision
Indexed by skills.sh from nexu-io/open-design
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- nexu-io
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__fal-visiongame-analytics-platform-computer-vision
Indexed by skills.sh from aradotso/data-skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- aradotso
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__game-analytics-platform-computer-visionhuggingface-vision-trainer
Indexed by skills.sh from huggingface/skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- huggingface
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__huggingface-vision-trainernorth-star-vision
Indexed by skills.sh from owl-listener/designer-skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- owl-listener
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__north-star-visionpdf-vision-reader
Indexed by skills.sh from childbamboo/claude-code-marketplace-sample
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- childbamboo
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__pdf-vision-readerproduct-vision
Indexed by skills.sh from phuryn/pm-skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- phuryn
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__product-visionqianwen-vision
Indexed by skills.sh from qianwen-ai/qianwen-ai
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- qianwen-ai
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__qianwen-visionqwencloud-vision
Indexed by skills.sh from qwencloud/qwencloud-ai
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- qwencloud
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__qwencloud-visionreact-native-vision-camera
Indexed by skills.sh from margelo/react-native-skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- margelo
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__react-native-vision-camerasenior-computer-vision
Indexed by skills.sh from alirezarezvani/claude-skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.5
- Author
- alirezarezvani
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__senior-computer-visionvision-analysis
Indexed by skills.sh from minimax-ai/skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- minimax-ai
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__vision-analysisvision-framework
Indexed by skills.sh from dpearson2699/swift-ios-skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- dpearson2699
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__vision-frameworkvision-multimodal
Indexed by skills.sh from lobbi-docs/claude
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- lobbi-docs
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__vision-multimodalvision-support
Indexed by skills.sh from penfick/skills
quarantinedskills.sh- Registry
- skills.sh
- Version
- 1.0.0
- Author
- penfick
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__vision-supportLanguage_Learning_Mastery
Adaptive language learning system with placement tests, structured weekly curricula, vocabulary spaced repetition, grammar lessons, pronunciation coaching, c...
quarantinedchineseconversationeducationClawHub- Registry
- ClawHub
- Category
- Research
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Language_Learning_MasteryLanguage_Learning_Tutor
AI language tutor for learning ANY language through conversation, vocab drills, grammar lessons, flashcards, and immersive practice. Use when the user wants to: learn a new language, practice vocabulary, study grammar, do flashcard drills, translate phrases, practice conversat…
quarantinedDELEDELFHSKClawHub- Registry
- ClawHub
- Category
- Research
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Language_Learning_TutorLanguage_Practice_Skill
Guides structured help for Language Practice using clear templates, checks, and safe defaults (category: Language Learning).
quarantinedlanguagelanguage-practiceopenclawClawHub- Registry
- ClawHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Language_Practice_SkillMalayalam_Language_Skill_(മലയാളം)
Understands Malayalam and Manglish WhatsApp messages and replies politely using culturally appropriate language in matching script style.
quarantinedindialanguagemalayalamClawHubbad-language-helper
Specializing in teaching the charm of language and creative responses
quarantinedLanguage LearningDialogue ExamplesLobeHub- Registry
- LobeHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:LobeHub__LobeHub__bad-language-helperturkish-language-tutor
AI Turkish Language Mentor: Introduce, teach, and support beginners in learning Turkish.
quarantinedturkish-languagelanguage-learningteachingLobeHub- Registry
- LobeHub
- Category
- Translation
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:LobeHub__LobeHub__turkish-language-tutorenglish-language-c-1-mastery-coach
English Conversation Partner for C1 Level
quarantinedenglish-conversationlanguage-proficiencyadvanced-levelLobeHub- Registry
- LobeHub
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:LobeHub__LobeHub__english-language-c-1-mastery-coachlanguage-fixer
Checks for typos and grammatical errors
quarantinedgrammaticaltypolanguageLobeHub- Registry
- LobeHub
- Category
- Creative
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:LobeHub__LobeHub__language-fixermulti-language-2-chinese-or-reverse
Multilingual translation, Chinese to English and Japanese, foreign languages to Chinese
quarantinedTranslationMultilingualLanguage ProcessingLobeHub- Registry
- LobeHub
- Category
- Translation
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:LobeHub__LobeHub__multi-language-2-chinese-or-reversetao-port-huggingface-model
Integrate a HuggingFace Computer Vision model into the NVIDIA TAO Toolkit ecosystem (tao-core config, tao-pytorch trainer, tao-deploy TensorRT pipeline). Use when the user asks to "integrate a HuggingFace model into TAO", "add an HF model to TAO Toolkit", "wire a HuggingFace V…
quarantinedtaohuggingfaceintegrationNVIDIA- Registry
- NVIDIA
- Category
- Vision AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__tao-port-huggingface-modelBTCvision_Donation_Nudge
Donation reminder for BTC-vision.org — checks monthly funding progress and sends a contextual Lightning tip request.
quarantinedClawHub- Registry
- ClawHub
- Category
- Productivity· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__BTCvision_Donation_Nudge