Inference AI
30 shown on page 1
Decision_Dynamo
Run a weighted decision matrix to score and rank 2-4 options across 5 configurable criteria. Use when a user needs help choosing between options, comparing t...
quarantinedClawHub- Registry
- ClawHub
- Category
- Inference Ai· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Decision_DynamoMem_Skill
Self-evolving memory and knowledge accumulation system for AI agents. Acts as a persistent 'second brain' that automatically retrieves past experiences, capt...
quarantinedClawHub- Registry
- ClawHub
- Category
- Inference Ai· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__Mem_Skilldynamo-interconnect-check
Validate that a Dynamo deployment's NIXL/UCX/NCCL interconnect is ready for disaggregated serving over RDMA/NVLink. Use after recipe-runner brings a deployment up (especially disagg/multi-node) to confirm the KV transport is correct; use troubleshoot for diagnosing already-fai…
quarantinedClawHub- Registry
- ClawHub
- Category
- Inference Ai· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__dynamo-interconnect-checkdynamo-recipe-runner
Select, validate, patch, and deploy existing NVIDIA Dynamo Kubernetes recipes. Use for model/backend/GPU/deployment-mode recipe bring-up; use router-starter for router-only mode work and troubleshoot for broken deployments.
quarantinedClawHub- Registry
- ClawHub
- Category
- Inference Ai· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__dynamo-recipe-runnerdynamo-router-starter
Start or patch Dynamo router modes and run router endpoint smoke checks. Use for round-robin, KV-aware, least-loaded, or device-aware routing setup; use recipe-runner for recipe deployment and troubleshoot for failure diagnosis.
quarantinedClawHub- Registry
- ClawHub
- Category
- Inference Ai· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__dynamo-router-starterdynamo-troubleshoot
Diagnose failed or unhealthy Dynamo deployments. Use when pods, model-cache jobs, PVCs, workers, frontend/router health, endpoints, or benchmark jobs fail; use recipe-runner/router-starter before this for normal bring-up.
quarantinedClawHub- Registry
- ClawHub
- Category
- Inference Ai· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__dynamo-troubleshootmem
Search local memory index (local-first). Use for /mem queries in Telegram.
quarantinedClawHub- Registry
- ClawHub
- Category
- Inference Ai· inferred
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:ClawHub__ClawHub__memdynamo-interconnect-check
Validate that a Dynamo deployment's NIXL/UCX/NCCL interconnect is ready for disaggregated serving over RDMA/NVLink. Use after recipe-runner brings a deployment up (especially disagg/multi-node) to confirm the KV transport is correct; use troubleshoot for diagnosing already-fai…
quarantinedNVIDIA- Registry
- NVIDIA
- Category
- Inference AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__dynamo-interconnect-checkdynamo-recipe-runner
Select, validate, patch, and deploy existing NVIDIA Dynamo Kubernetes recipes. Use for model/backend/GPU/deployment-mode recipe bring-up; use router-starter for router-only mode work and troubleshoot for broken deployments.
quarantinedNVIDIA- Registry
- NVIDIA
- Category
- Inference AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__dynamo-recipe-runnerdynamo-router-starter
Start or patch Dynamo router modes and run router endpoint smoke checks. Use for round-robin, KV-aware, least-loaded, or device-aware routing setup; use recipe-runner for recipe deployment and troubleshoot for failure diagnosis.
quarantinedNVIDIA- Registry
- NVIDIA
- Category
- Inference AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__dynamo-router-starterdynamo-troubleshoot
Diagnose failed or unhealthy Dynamo deployments. Use when pods, model-cache jobs, PVCs, workers, frontend/router health, endpoints, or benchmark jobs fail; use recipe-runner/router-starter before this for normal bring-up.
quarantinedNVIDIA- Registry
- NVIDIA
- Category
- Inference AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__dynamo-troubleshootjetson-inference-mem-tune
Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson.
quarantinedNVIDIA- Registry
- NVIDIA
- Category
- Inference AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__jetson-inference-mem-tunejetson-llm-benchmark
Benchmark Jetson LLM/VLM serving performance across vLLM, llama.cpp, and Ollama with structured JSON output.
quarantinedNVIDIA- Registry
- NVIDIA
- Category
- Inference AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__jetson-llm-benchmarkjetson-llm-serve
Stand up vLLM or SGLang serving on Jetson, using upstream vLLM on Thor and Orin JetPack 7.2+, and NVIDIA-AI-IOT vLLM on older Orin.
quarantinedNVIDIA- Registry
- NVIDIA
- Category
- Inference AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__jetson-llm-servejetson-speculative-decoding
Add EAGLE-3 or draft-model speculative decoding to a Jetson vLLM server when TPOT is the bottleneck.
quarantinedNVIDIA- Registry
- NVIDIA
- Category
- Inference AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__jetson-speculative-decodingtao-run-inference-service
Start, query, and stop a network-specific TAO inference microservice ({network_arch}-inference-microservice) by delegating container execution to the appropriate platform skill. Handles container image resolution, job-payload JSON construction, and the service registry. Use wh…
quarantinedinferencemicroserviceworkflowNVIDIA- Registry
- NVIDIA
- Category
- Inference AI
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:NVIDIA__NVIDIA__tao-run-inference-servicecompare-and-build-llm-model
Search and compare OpenRouter's 350+ LLMs by cost, speed (throughput/latency/uptime), context length, modalities, and use-case category; pick the best fit; then build with it via OpenAI-compatible chat completions. API-first — no scraping, no auth for reads.
quarantinedllmopenroutermodel-selectionbrowse.sh- Registry
- browse.sh
- Category
- Ai Models
- Version
- 1.0.0
curl -s /v1/skills/community__axehub:browse.sh__browse.sh__compare-and-build-llm-modelbabysit
Indexed by skills.sh from thedotmack/claude-mem
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- thedotmack
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__babysitdflash-mlx-speculative-decoding
Indexed by skills.sh from aradotso/trending-skills
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- aradotso
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__dflash-mlx-speculative-decodingdo
Indexed by skills.sh from thedotmack/claude-mem
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- thedotmack
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__dodynamo-interconnect-check
Indexed by skills.sh from nvidia/skills
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- nvidia
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__dynamo-interconnect-checkdynamo-recipe-runner
Indexed by skills.sh from nvidia/skills
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- nvidia
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__dynamo-recipe-runnerdynamo-router-starter
Indexed by skills.sh from nvidia/skills
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- nvidia
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__dynamo-router-starterdynamo-troubleshoot
Indexed by skills.sh from nvidia/skills
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- nvidia
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__dynamo-troubleshoothf-mem
Indexed by skills.sh from huggingface/skills
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- huggingface
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__hf-memmem
Indexed by skills.sh from runablehq/memory
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- runablehq
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__memmem-search
Indexed by skills.sh from thedotmack/claude-mem
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- thedotmack
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__mem-searchpathfinder
Indexed by skills.sh from thedotmack/claude-mem
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- thedotmack
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__pathfinderspeculative-decoding
Indexed by skills.sh from davila7/claude-code-templates
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.5
- Author
- davila7
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__speculative-decodingwowerpoint
Indexed by skills.sh from thedotmack/claude-mem
quarantinedskills.sh- Registry
- skills.sh
- Category
- Inference Ai· inferred
- Version
- 1.0.0
- Author
- thedotmack
curl -s /v1/skills/community__axehub:skills.sh__skills.sh__wowerpoint