Jetson Inference Mem Tune

nvidia/skills/skills/jetson-inference-mem-tune

by nvidiacf5224d14250Apache-2.03.5K starsListed Oct 8, 2026Updated Oct 8, 2026Repository updated today

Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson.

Includes scriptsDevOps & Cloud
  1. cf5224d14250Currentcommit cf5224dPublished Oct 8, 2026

Source and attribution

Source:nvidia/skillsinskills/jetson-inference-mem-tuneat commitcf5224d

License: Apache-2.0

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal