bynvidia-ai-iot· 35 skills
Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson.
$npx -y skills add nvidia-ai-iot/jetson-device-skills --skill jetson-inference-mem-tuneInstalls into the current project.
Run `npx skills use "https://github.com/nvidia-ai-iot/jetson-device-skills" --skill "nvidia-ai-iot/jetson-device-skills/jetson-inference-mem-tune"` and follow the generated skill instructions now. Read its complete output, redirecting it to a temporary file first if necessary. Resolve relative paths from the supporting-files directory it provides.
Use the skills in "https://github.com/nvidia-ai-iot/jetson-device-skills" that are relevant to the current task. Run `npx skills add "https://github.com/nvidia-ai-iot/jetson-device-skills"` and select the relevant skills, then follow their instructions.