Jetson LLM Benchmark

Benchmark Jetson LLM/VLM serving performance across vLLM, llama.cpp, and Ollama with structured JSON output.

About this skill

Benchmark Jetson LLM/VLM serving performance across vLLM, llama.cpp, and Ollama with structured JSON output.

Maintained by NVIDIA. The source includes the instructions and any supporting files needed to use this skill.

Inside the instructions

  • 01Purpose
  • 02Prerequisites
  • 03Available Scripts
  • 04Instructions
  • 05Expected Workflow
  • 06When to use

Before you start

  1. Read the instructions and check tool or account requirements.
  2. Install the complete folder when the skill references scripts or other files.
  3. Provide your task context, then review the agent's output.

Source

nvidia/skills / jetson-llm-benchmark

Source reviewed October 2, 2026 · Apache-2.0 / CC-BY-4.0; see source notices