Agent skill · nvidia

jetson-inference-mem-tune

Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson.

Reference it in any coding agent with:

@skills nvidia/jetson-inference-mem-tune

Browse the @skills marketplace