Jetson Inference Mem Tune

nvidia/skills/skills/jetson-inference-mem-tune

by nvidiacf5224d14250Apache-2.03.5K starsListed Oct 8, 2026Updated Oct 8, 2026Repository updated today

Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson.

Includes scriptsDevOps & Cloud

Only the file list is public. File contents are available once the skill is installed in a workspace.

PathSizeType
BENCHMARK.md3.3 KBtext/markdown
evals/evals.json5.4 KBapplication/json
scripts/recommend.py10 KBtext/plain
skill-card.md3.7 KBtext/markdown
SKILL.md11.2 KBtext/markdown
skill.oms.sig4.7 KBtext/plain

Source and attribution

Source:nvidia/skillsinskills/jetson-inference-mem-tuneat commitcf5224d

License: Apache-2.0

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal