
Jetson Diagnostic
OfficialFreeGet a comprehensive health snapshot of your Jetson device.
Free · Opens the source repo
What Jetson Diagnostic does
The Jetson Diagnostic skill provides a unified, read-only view of the health status of a running Jetson device. This skill simplifies the process of obtaining vital information about the device by consolidating data from various sources such as tegrastats, jtop, and nvpmodel. Instead of needing to remember which command provides specific information, users can rely on this skill to deliver a comprehensive snapshot of device identity, memory usage, GPU performance, thermal conditions, power status, storage, and active services.
When activated, the skill captures a health snapshot that answers common queries about the Jetson device. For example, users can inquire about the device's SKU, memory availability, or the current processes consuming resources. This is particularly useful for diagnosing performance issues, such as a device running slowly or overheating, by providing real-time data that can be analyzed to identify the root cause.
The skill is designed for use in scenarios where live data is essential, such as during development, testing, or troubleshooting of Jetson-based applications. It can be invoked when users need to understand the current state of their device or when they require a baseline measurement before running other Jetson-specific skills.
However, it is important to note that this skill is read-only and does not allow for any modifications to the device's state. Users should not attempt to change power modes or stop services using this skill; it is strictly for reporting observed conditions. For actions requiring changes, users should refer to other action-oriented skills after obtaining the necessary data from this diagnostic tool.
When to use it
Use this skill when you need a quick and comprehensive overview of a Jetson device's health or performance metrics.
When not to use it
Avoid using this skill for modifying device settings or performing actions that require write access to the system.
What you can build with it
Performance Troubleshooting
When a user reports that their Jetson device is running slowly, they can use this skill to obtain real-time metrics on memory and CPU usage.
Device Identity Verification
A developer can quickly check the SKU and memory specifications of their Jetson device to ensure compatibility with their applications.
Pre-Deployment Baseline Measurement
Before running intensive applications, users can capture a health snapshot to establish a baseline for performance metrics.
How to install Jetson Diagnostic
View source1. Install with the skills CLI
npx skills add nvidia/skills/jetson-diagnostic --agent claude-code2. Or install it manually
Download the skill folder and drop it into ~/.claude/skills/ for all projects, or .claude/skills/ to scope it to one repo. Restart Claude Code so it picks up the new skill.
Anthropic's agentic coding CLI, and the reference implementation of Agent Skills. Drop a skill folder into ~/.claude/skills and Claude Code loads it automatically whenever a task matches the skill's description. Claude Code docs
Inside SKILL.md
Written by nvidiaJetson Diagnostic
A unified, agent-friendly view of a running Jetson device. Replaces the need to remember which of tegrastats, jtop, procrank, /sys/kernel/debug/nvmap, nvpmodel, free, swapon, df, and systemctl list-units produces which slice of the truth.
Purpose
Capture a read-only health snapshot from the Jetson host so agents can answer device identity, memory, GPU, thermal, power, storage, and service-state questions using live data instead of guesses.
When to use
Activate when the user asks:
- "What is this Jetson? What SKU? How much memory?"
- "What's running on this Jetson right now?"
- "Why is my Jetson slow / hot / out of memory?"
- "Give me a snapshot of GPU / CPU / power usage."
- "What does my tegrastats output mean?"
- "Which services are running that I could turn off?"
- The user has installed
jetson-memory-audit,jetson-headless-mode,jetson-inference-mem-tune,jetson-llm-benchmark,jetson-llm-serve, orjetson-packageand needs a baseline measurement before running them.
Do not use this skill to change power modes, drop caches, stop services, install packages, serve models, or tune inference flags. Report the observed state, then hand off to the action-oriented skill.
Prerequisites
- Run on the Jetson host, or in a sandbox/container with host-visible Jetson system paths and process data.
Available Scripts
| Script | Purpose | Arguments |
|---|---|---|
scripts/snapshot.sh | Emits the all-in-one JSON snapshot for identity, memory, GPU, thermal, power, disk, top processes, and candidate services. | --human, --tegra-secs N, --top-procs N. |
scripts/mem_summary.sh | Emits a compact human-readable RAM/GPU/swap summary. | --short, --watch, --interval N. |
scripts/detect_jetson.sh | Exports or prints canonical Jetson SKU/generation/product-line fields for this repo. | No arguments. |
If your agent runtime supports run_script, use it to run scripts/snapshot.sh or scripts/mem_summary.sh and summarize the returned output. Otherwise run the scripts with bash from the repository root.
Instructions
- Run
scripts/snapshot.shfor the all-in-one JSON view (preferred default). - For a quick human-readable memory line, run
scripts/mem_summary.sh. - To explain a single tegrastats line the user has pasted, see
references/tegrastats-fields.md. - To explain the NvMap clients output, see
references/nvmap-clients.md.
Reporting guidance
Run the matching helper script before summarizing device state, and report only fields returned by that script. If direct execution is blocked by the runtime, run it with bash {baseDir}/scripts/<script-name> rather than trying to chmod files.
- For "what is this Jetson" questions, quote
product_modelorsku,variant,l4t_version, andmem_total_gb. - For "slow and hot" questions, run
snapshot.shand summarize both sides of the symptom:thermal_cfor heat, plustop_processes,gpu_processes,nvmap.top_clients, orgpu_sourcefor load. End with a concrete handoff such asjetson-memory-audit,jetson-headless-mode, orjetson-inference-mem-tune. - For "which process is using memory" questions, run
snapshot.shand name the leading process aspid <number>,cmd, and itspss_kb/ MiB value. If NvMap GPU memory is the relevant signal, also quotegpu_sourceand the topnvmap.top_clientsorgpu_processesentry.
If your agent runtime does not automatically execute helper scripts relative to this skill directory, resolve script paths with the AgentSkills {baseDir} placeholder:
{baseDir}/scripts/snapshot.sh
{baseDir}/scripts/mem_summary.sh
Do not call jetson-diagnostic as a tool name unless the runtime explicitly registers skills as callable tools; Agent Skills are normally instructions plus files, not direct tool functions.
All scripts source the canonical platform detector at skills/jetson-diagnostic/scripts/detect_jetson.sh (exports JETSON_SKU, JETSON_GENERATION, JETSON_PRODUCT_LINE, JETSON_VARIANT, JETSON_MEM_GB, JETSON_L4T_VERSION, JETSON_PRODUCT_MODEL). Other skills may source this detector rather than duplicating Jetson identification logic. Exits 2 with a remediation message off-platform.
Limitations
- Seeing this skill file does not guarantee access to Jetson host hardware. If
/proc/device-tree/model,/etc/nv_tegra_release,tegrastats,nvpmodel,nvidia-smi, or/sys/kernel/debug/nvmapare missing inside a NemoClaw/OpenClaw sandbox, say the sandbox lacks Jetson host visibility and ask the user to run on the Jetson host or relaunch with a host-visible sandbox profile. - NvMap debugfs often requires root, so unprivileged runs may report
gpu_source: "none"or incompletenvmapfields. - This skill reports observed state only. Do not fabricate memory, GPU, thermal, service, or reclamation data when a tool is missing or inaccessible.
Error handling
- If a helper exits off-platform, report that the current environment is not a Jetson host or lacks host visibility; do not substitute generic Linux values.
- If
tegrastats,nvpmodel,nvidia-smi, or NvMap debugfs are unavailable, preserve the correspondingnull,false, or empty fields from the JSON and explain which signal is limited. - If
snapshot.shemits malformed JSON, report the raw failure and rerun after fixing the helper output; do not hand-edit a synthetic device snapshot.
Output contract for snapshot.sh
{
"sku": "orin-nano",
"generation": "orin",
"product_line": "orin-nano",
"variant": "orin-nano-8gb",
"mem_total_gb": 8,
"l4t_version": "36.4.0",
"product_model": "nvidia jetson orin nano developer kit",
"memory_kb": { "total": 8123456, "available": 4123456, "swap_total": 0, "swap_free": 0, "cached": 1234567 },
"tegrastats_sample": "RAM 4011/8138MB (lfb 8x4MB) ...",
"thermal_c": { "CPU": 52.3, "GPU": 49.0, "AO": 47.0 },
"power": { "nvpmodel_id": 0, "nvpmodel_name": "MAXN" },
"disk": [ { "mount": "/", "used_pct": 41 } ],
"gpu_source": "nvmap:iovmm-clients",
"gpu_devices": [],
"gpu_processes": [],
"nvmap": {
"readable": true,
"total_kb": 654321,
"stats_total_bytes": 669985280,
"top_clients": [ { "pid": 1234, "cmd": "vlm-server", "kb": 524288 } ]
},
"top_processes": [ { "pid": 4321, "cmd": "vllm", "pss_kb": 4000000 } ],
"candidate_services": { "gdm3": { "active": "inactive", "enabled": "disabled" } }
}
gpu_source names the specific datum the skill used to attribute per-process GPU memory, so the caller can tell exactly what the numbers represent:
"nvidia-smi:compute-apps"— per-processused_memoryvalues fromnvidia-smi --query-compute-apps. Used on the unifiednvidia.kostack (Thor family today). Note: on this stacknvidia-smi's device-levelmemory.usedquery returns[N/A]on some BSPs, which is why the skill sums the per-process list rather than reading a top-level total. The summed total appears ingpu_processes[*].used_mib."nvmap:iovmm-clients"— per-process sizes from/sys/kernel/debug/nvmap/iovmm/clients. Used on thenvgpustack (Orin family today), wherenvidia-smiis a stub that returns[N/A]for every compute/memory query. Per-process entries appear innvmap.top_clients; the kernel-side total is innvmap.total_kband (when readable)nvmap.stats_total_bytes."none"— no authoritative source reachable. Typical when running unprivileged on annvgpu-stack Jetson (debugfs under/sys/kernel/debug/nvmapneedssudo); rerun withsudoto populate thenvmapfields.
The agent should present the salient parts back to the user (SKU, available memory, top GPU consumer per gpu_source, hottest zone, power mode) and offer to drill into specifics (top_processes, gpu_processes / nvmap, services).
Safety
This skill is read-only. It does not change nvpmodel, does not run jetson_clocks, does not modify services. To act on findings, hand off to:
jetson-memory-audit— focused memory snapshot + drop_caches verify loopjetson-headless-mode— disable GUI + auxiliary daemons (safe, reversible)jetson-inference-mem-tune— pick runtime + memory flags (vLLM / SGLang / llama.cpp / TensorRT Edge-LLM)jetson-llm-serve— vLLM and related GHCR images with Jetson defaultsjetson-llm-benchmark— reproducible latency / throughput benchmarksjetson-package— GHCR + Jetson AI Lab PyPI indexes vs generic ARM wheels
Cross-platform behavior
| Family | Variants the skill recognises | tegrastats | nvidia-smi | nvpmodel | NvMap debugfs |
|---|---|---|---|---|---|
| Jetson Thor | thor-t5000, thor-t4000 | yes | yes (full) | yes | yes (root) |
| Jetson AGX Orin | orin-agx-64gb, orin-agx-32gb, orin-agx-industrial | yes | yes (stub, nvgpu)* | yes | yes (root) |
| Jetson Orin NX | orin-nx-16gb, orin-nx-8gb | yes | yes (stub, nvgpu)* | yes | yes (root) |
| Jetson Orin Nano | orin-nano-8gb, orin-nano-4gb | yes | yes (stub, nvgpu)* | yes | yes (root) |
* On Jetsons whose GPU is driven by the in-tree nvgpu kernel driver, the nvidia-smi binary is present but most fields (Memory-Usage, power, utilisation, compute-process table) report Not Supported / N/A. To decide which source to trust at runtime, the script does a capability probe — nvidia-smi --query-gpu=memory.used --format=csv,noheader,nounits — and only uses nvidia-smi for per-process GPU memory when that query returns a real integer. When it doesn't, the script falls back to /sys/kernel/debug/nvmap/iovmm/clients, which on nvgpu-stack Jetsons is the authoritative per-process GPU-memory source.
The script handles each tool's presence gracefully and reports null / false for tools it cannot reach (typical when the agent isn't running with the privilege needed for /sys/kernel/debug). Variant detection uses the /proc/device-tree/model string first (recognising names like T5000 / T4000) and falls back to memory-size heuristics when the model string is generic.
Frequently asked questions about Jetson Diagnostic
Similar skills
Spring Boot Testing
Master testing techniques for Spring Boot 4 applications.
GitHub Issues
Manage GitHub issues efficiently with MCP tools.
Geofeed Tuner
Optimize your IP geolocation feeds in CSV format.
Batch Files
Master Windows batch scripting for automation and task management.
Adobe Illustrator Scripting
Automate your Illustrator workflows with ExtendScript.
Plugin Structure
Create and organize Claude Code plugins effectively.
