# \#monitoring

**URL:** https://forums.developer.nvidia.com/tag/monitoring/516.md

[Latest](https://forums.developer.nvidia.com/latest.md) · [Categories](https://forums.developer.nvidia.com/categories.md) · [Tags](https://forums.developer.nvidia.com/tags.md)

---

## [Native Monitoring Terminal App](https://forums.developer.nvidia.com/t/native-monitoring-terminal-app/371303)

<div class="topic-metadata">

**Author:** [@overbid.poxes\_7y](https://forums.developer.nvidia.com/u/overbid.poxes_7y)\
**Replies:** 1\
**Last updated:** [May 25, 2026, 9:23pm UTC](https://forums.developer.nvidia.com/t/native-monitoring-terminal-app/371303 "2026-05-25T21:23:48Z")

</div>

I am about to publish a terminal app for Spark, since the current options like nvidia-smi or nvtop, all have some broken aspects since their primary focus is not the Spark.

---

## [Cluster Monitoring Tool - Docker Image](https://forums.developer.nvidia.com/t/cluster-monitoring-tool-docker-image/368717)

<div class="topic-metadata">

**Author:** [@chronosolidus](https://forums.developer.nvidia.com/u/chronosolidus)\
**Replies:** 3\
**Last updated:** [May 2, 2026, 3:49am UTC](https://forums.developer.nvidia.com/t/cluster-monitoring-tool-docker-image/368717 "2026-05-02T03:49:39Z")

</div>

I got my hands on a few DGX Spark units and have been experimenting with various use cases. The major pain point throughout was being forced to continuously open the various CLI monitoring tools for hardware observabili…

---

## [Nv-monitor now supports Jetson](https://forums.developer.nvidia.com/t/nv-monitor-now-supports-jetson/365610)

<div class="topic-metadata">

**Author:** [@wentbackward](https://forums.developer.nvidia.com/u/wentbackward)\
**Replies:** 3\
**Last updated:** [April 7, 2026, 1:44pm UTC](https://forums.developer.nvidia.com/t/nv-monitor-now-supports-jetson/365610 "2026-04-07T13:44:55Z")

</div>

nv-monitor now supports Jetson Orin (Nano, NX, AGX) We’ve added full GPU monitoring support for Jetson devices in nv-monitor (GitHub - wentbackward/nv-monitor: A lightweight terminal system monitor built for the \*\*NVIDIA…

---

## [Nv-monitor: Add RDMA/InfiniBand metrics (data export only) - testers needed](https://forums.developer.nvidia.com/t/nv-monitor-add-rdma-infiniband-metrics-data-export-only-testers-needed/365419)

<div class="topic-metadata">

**Author:** [@wentbackward](https://forums.developer.nvidia.com/u/wentbackward)\
**Replies:** 8\
**Last updated:** [April 4, 2026, 5:11pm UTC](https://forums.developer.nvidia.com/t/nv-monitor-add-rdma-infiniband-metrics-data-export-only-testers-needed/365419 "2026-04-04T17:11:52Z")

</div>

Could anyone with the appropriate hardware access test nv-monitor is correctly exporting RDMA metrics please? I don’t have access to the hardware to do so. Binaries are built by the ci, so should be quick to dump it to C…

---

## [Small monitor program for DGX Spark](https://forums.developer.nvidia.com/t/small-monitor-program-for-dgx-spark/364123)

<div class="topic-metadata">

**Author:** [@wentbackward](https://forums.developer.nvidia.com/u/wentbackward)\
**Replies:** 18\
**Last updated:** [April 1, 2026, 6:14am UTC](https://forums.developer.nvidia.com/t/small-monitor-program-for-dgx-spark/364123 "2026-04-01T06:14:31Z")

</div>

This is turning out to be very useful to me … In case it’s useful to you: GitHub - wentbackward/nv-monitor: A lightweight terminal system monitor built for the \*\*NVIDIA DGX Spark\*\* (Grace CPU + GB10 GPU). · GitHub 73…

---

## [Extremely fast flickering image graphic like a flash of lightning when different scaling](https://forums.developer.nvidia.com/t/extremely-fast-flickering-image-graphic-like-a-flash-of-lightning-when-different-scaling/351464)

<div class="topic-metadata">

**Author:** [@meinertreu74](https://forums.developer.nvidia.com/u/meinertreu74)\
**Replies:** 1\
**Last updated:** [November 16, 2025, 2:39pm UTC](https://forums.developer.nvidia.com/t/extremely-fast-flickering-image-graphic-like-a-flash-of-lightning-when-different-scaling/351464 "2025-11-16T14:39:39Z")

</div>

monitor 1: Dell U2410 —\> scale = 100% monitor 2: Asus ProArtPA279CRV —\> scale= 200% nvidia-driver-version: 580.95.05 nvidia-bug-report is here…. nvidia-bug-report.log.gz (552.0 KB)

---

## [AGX Thor shows abnormal display when using the DP interface to connect the monitor](https://forums.developer.nvidia.com/t/agx-thor-shows-abnormal-display-when-using-the-dp-interface-to-connect-the-monitor/350227)

<div class="topic-metadata">

**Author:** [@2420565083](https://forums.developer.nvidia.com/u/2420565083)\
**Replies:** 2\
**Last updated:** [November 5, 2025, 9:57am UTC](https://forums.developer.nvidia.com/t/agx-thor-shows-abnormal-display-when-using-the-dp-interface-to-connect-the-monitor/350227 "2025-11-05T09:57:50Z")

</div>

Hello to everyone browsing this post, Regarding the use of the AGX Thor device, I have encountered an abnormal phenomenon. After powering on the AGX Thor device and connecting a 4K monitor via the DP interface wi…

---

## [RTX 5090 disconnection HDMI/DISPLAY PORT during heavy AI loading](https://forums.developer.nvidia.com/t/rtx-5090-disconnection-hdmi-display-port-during-heavy-ai-loading/343897)

<div class="topic-metadata">

**Author:** [@shazello.shazellano.shaza](https://forums.developer.nvidia.com/u/shazello.shazellano.shaza)\
**Replies:** 0\
**Last updated:** [September 3, 2025, 12:06am UTC](https://forums.developer.nvidia.com/t/rtx-5090-disconnection-hdmi-display-port-during-heavy-ai-loading/343897 "2025-09-03T00:06:11Z")

</div>

First of all, greetings to those who are reading me. My hardware is: Ryzen 9 9950X3D processor, ROG STRIX X670-E GAMING WIFI motherboard, 64GB of RAM G.SKILL EXPO, my graphics card is: AORUS GeForce RTX™ 5090 MASTER 32G…

---

## [Bright Cluster Manager - Grafana Dashboard Examples](https://forums.developer.nvidia.com/t/bright-cluster-manager-grafana-dashboard-examples/324548)

<div class="topic-metadata">

**Author:** [@phil.hale](https://forums.developer.nvidia.com/u/phil.hale)\
**Replies:** 0\
**Last updated:** [February 20, 2025, 10:23pm UTC](https://forums.developer.nvidia.com/t/bright-cluster-manager-grafana-dashboard-examples/324548 "2025-02-20T22:23:16Z")

</div>

Just wanted to see if anyone has sample Grafana Dashboards they are using with the built-in prometheus data source. If so, would you be willing to share? Thanks, Phil

---

## [Jtop refresh setting for custom flask server](https://forums.developer.nvidia.com/t/jtop-refresh-setting-for-custom-flask-server/320461)

<div class="topic-metadata">

**Author:** [@charles.fonbonne](https://forums.developer.nvidia.com/u/charles.fonbonne)\
**Replies:** 3\
**Last updated:** [January 20, 2025, 6:55pm UTC](https://forums.developer.nvidia.com/t/jtop-refresh-setting-for-custom-flask-server/320461 "2025-01-20T18:55:26Z")

</div>

I’m currently trying to monitor the energy consumption of my AGX Orin jetson with the python package jtrop. I’m using prometheus and grafana to get a remote visualisation, and I’d like to know how I can increase the jtop…

---

## [Any way to monitor GPU-specific power usage on Jetson Orin Nano?](https://forums.developer.nvidia.com/t/any-way-to-monitor-gpu-specific-power-usage-on-jetson-orin-nano/318764)

<div class="topic-metadata">

**Author:** [@v39zhang](https://forums.developer.nvidia.com/u/v39zhang)\
**Replies:** 1\
**Last updated:** [January 6, 2025, 2:28am UTC](https://forums.developer.nvidia.com/t/any-way-to-monitor-gpu-specific-power-usage-on-jetson-orin-nano/318764 "2025-01-06T02:28:20Z")

</div>

I am trying to monitor the power consumption of the GPU on my Jetson Orin Nano. When I run tegrastats, it outputs the power usage for VDD\_CPU\_GPU\_CV, which seems to include the combined power consumption for the CPU, GPU…

---

## [Support hwmon for GPU monitoring](https://forums.developer.nvidia.com/t/support-hwmon-for-gpu-monitoring/305234)

<div class="topic-metadata">

**Author:** [@AthanSpod2](https://forums.developer.nvidia.com/u/AthanSpod2)\
**Replies:** 4\
**Last updated:** [August 31, 2024, 4:35pm UTC](https://forums.developer.nvidia.com/t/support-hwmon-for-gpu-monitoring/305234 "2024-08-31T16:35:08Z")

</div>

First, I am aware of nvidia-smi for retrieving monitoring data. However, my needs centre around using lm\_sensors, fancontrol and other software that expect to be able to access such data via the Linux Kernel hwmon inter…

---

## [ORIN DP out with MST hub](https://forums.developer.nvidia.com/t/orin-dp-out-with-mst-hub/293233)

<div class="topic-metadata">

**Author:** [@larson2](https://forums.developer.nvidia.com/u/larson2)\
**Replies:** 1\
**Last updated:** [May 17, 2024, 7:43am UTC](https://forums.developer.nvidia.com/t/orin-dp-out-with-mst-hub/293233 "2024-05-17T07:43:21Z")

</div>

Hi NVIDIA team some previously display issue ,ORIN with MST Is the bug fixed ? Any newer news?

---

## [Trying to use Nsight Graphics on Windows to connect to Linux box](https://forums.developer.nvidia.com/t/trying-to-use-nsight-graphics-on-windows-to-connect-to-linux-box/259143)

<div class="topic-metadata">

**Author:** [@Uncorp](https://forums.developer.nvidia.com/u/Uncorp)\
**Replies:** 1\
**Last updated:** [August 22, 2023, 2:23pm UTC](https://forums.developer.nvidia.com/t/trying-to-use-nsight-graphics-on-windows-to-connect-to-linux-box/259143 "2023-08-22T14:23:02Z")

</div>

I have a windows machine that I installed Nvidia Nsight for GPU trace and profiling, etc. And I want it to connect to a Linux box executing a process that I want to remotely monitor GPU trace, etc. According to User Gu…

---

## [Enabling RAS drivers on NX dev kit (Reliability, Accessibility and Serviceability)](https://forums.developer.nvidia.com/t/enabling-ras-drivers-on-nx-dev-kit-reliability-accessibility-and-serviceability/240912)

<div class="topic-metadata">

**Author:** [@shaantam](https://forums.developer.nvidia.com/u/shaantam)\
**Replies:** 3\
**Last updated:** [February 22, 2023, 8:36am UTC](https://forums.developer.nvidia.com/t/enabling-ras-drivers-on-nx-dev-kit-reliability-accessibility-and-serviceability/240912 "2023-02-22T08:36:20Z")

</div>

Hi Nvidia community, I’m interested in using the Xavier NX’s built-in Reliability, Accessibility and Serviceability (RAS) drivers to understand ECC performance of the Carmel CPUs in an extreme environment. 2 goals : (1…

---

## [Query power consumption directly instead of using tegrastats](https://forums.developer.nvidia.com/t/query-power-consumption-directly-instead-of-using-tegrastats/241741)

<div class="topic-metadata">

**Author:** [@terunofuji](https://forums.developer.nvidia.com/u/terunofuji)\
**Replies:** 8\
**Last updated:** [February 8, 2023, 5:48pm UTC](https://forums.developer.nvidia.com/t/query-power-consumption-directly-instead-of-using-tegrastats/241741 "2023-02-08T17:48:08Z")

</div>

On TX2, users can directly query the instant power consumption of each hardware component. Can we do the same on Jetson Orin? More specifically, can we query VDD\_GPU\_SOC, VDD\_CPU\_CV, VIN\_SYS\_5V0, and VDDQ\_VDD2\_1V8AO pow…

---

## [2 Gpu 6 Monitoring](https://forums.developer.nvidia.com/t/2-gpu-6-monitoring/241003)

<div class="topic-metadata">

**Author:** [@imwuuuu](https://forums.developer.nvidia.com/u/imwuuuu)\
**Replies:** 0\
**Last updated:** [January 30, 2023, 7:07am UTC](https://forums.developer.nvidia.com/t/2-gpu-6-monitoring/241003 "2023-01-30T07:07:50Z")

</div>

Hi. I have 2 Nvidia a6000 graphics cards and I connect them to 6 monitors with display cables. but since it has only 4 outputs, I can display 4 monitors. I can’t get an image from 2 dislays from the other card. I tried t…

---

## [Provide solution for "GPU MEM used by PID but no GPU LOAD"](https://forums.developer.nvidia.com/t/provide-solution-for-gpu-mem-used-by-pid-but-no-gpu-load/237753)

<div class="topic-metadata">

**Author:** [@ron.koss](https://forums.developer.nvidia.com/u/ron.koss)\
**Replies:** 2\
**Last updated:** [January 10, 2023, 10:52am UTC](https://forums.developer.nvidia.com/t/provide-solution-for-gpu-mem-used-by-pid-but-no-gpu-load/237753 "2023-01-10T10:52:00Z")

</div>

Hi all, as a service provider, I am managing some DGX-1 and DGX-2 machines for a customer in a multi-.user-shared DL/ML environment on Linux. We use nvidia-smi regularly. In some cases, the tool does not provide the inf…

---

## [dcgmUpdateAllFields returns "Timeout"](https://forums.developer.nvidia.com/t/dcgmupdateallfields-returns-timeout/237764)

<div class="topic-metadata">

**Author:** [@dmitrygx](https://forums.developer.nvidia.com/u/dmitrygx)\
**Replies:** 1\
**Last updated:** [December 27, 2022, 10:13pm UTC](https://forums.developer.nvidia.com/t/dcgmupdateallfields-returns-timeout/237764 "2022-12-27T22:13:38Z")

</div>

Hi NVIDIA experts, I use golang DCGM API to watch fields and the update their values through Fields API. nv-hostengine is launched as a standalone process in my setup. And DCGM client is connected through TCP-connectio…

---

## [NVIDIA Orin Devkit no display](https://forums.developer.nvidia.com/t/nvidia-orin-devkit-no-display/236015)

<div class="topic-metadata">

**Author:** [@ee18resch11006](https://forums.developer.nvidia.com/u/ee18resch11006)\
**Replies:** 0\
**Last updated:** [December 2, 2022, 10:43am UTC](https://forums.developer.nvidia.com/t/nvidia-orin-devkit-no-display/236015 "2022-12-02T10:43:30Z")

</div>

I have got a new Nvidia Orin Devkit. When I powered it for the first time, nothing came up on the monitor. It just says, no signal detected. please help…

---

## [DLI - Course - NameError: name '\<\<\<\<FIXME\>\>\>\>' is not defined - Disaster Risk Monitoring Using Satellite Imagery - Flood Detection](https://forums.developer.nvidia.com/t/dli-course-nameerror-name-fixme-is-not-defined-disaster-risk-monitoring-using-satellite-imagery-flood-detection/234838)

<div class="topic-metadata">

**Author:** [@diyanderson](https://forums.developer.nvidia.com/u/diyanderson)\
**Replies:** 0\
**Last updated:** [November 18, 2022, 6:02pm UTC](https://forums.developer.nvidia.com/t/dli-course-nameerror-name-fixme-is-not-defined-disaster-risk-monitoring-using-satellite-imagery-flood-detection/234838 "2022-11-18T18:02:06Z")

</div>

Hi, Congratulations nVIDIA is an excellent course. but I’m having a hard time understanding the execution of this step: In Notebook assessment.ipynb Instructions : 3.1 Execute the below cell to load dependencies, se…

---

## [Display](https://forums.developer.nvidia.com/t/display/220431)

<div class="topic-metadata">

**Author:** [@lahroodi](https://forums.developer.nvidia.com/u/lahroodi)\
**Replies:** 1\
**Last updated:** [July 13, 2022, 2:29pm UTC](https://forums.developer.nvidia.com/t/display/220431 "2022-07-13T14:29:57Z")

</div>

I followed the instruction. Connect all cables including Display-HDMI. My monitor ( HDMI 1 port) shows nothing. I tested it with three different monitors. None of them worked. I was able to see the Terminal by using my P…

---

## [How can I measure memory bandwidth caused by CPU/GPU/DLA separately?](https://forums.developer.nvidia.com/t/how-can-i-measure-memory-bandwidth-caused-by-cpu-gpu-dla-separately/220162)

<div class="topic-metadata">

**Author:** [@sunlingyu](https://forums.developer.nvidia.com/u/sunlingyu)\
**Replies:** 2\
**Last updated:** [July 10, 2022, 1:39am UTC](https://forums.developer.nvidia.com/t/how-can-i-measure-memory-bandwidth-caused-by-cpu-gpu-dla-separately/220162 "2022-07-10T01:39:47Z")

</div>

Hi, I understand that on Xavier, CPU/GPU/DLA shares one memory. I would like to get the memory bandwidth caused by each of them when all of the CPU/GPU/DLA are running at the same time. I know that tegrastats can give t…

---

## [Tegrastats monitoring](https://forums.developer.nvidia.com/t/tegrastats-monitoring/217088)

<div class="topic-metadata">

**Author:** [@Harry-S](https://forums.developer.nvidia.com/u/Harry-S)\
**Replies:** 2\
**Last updated:** [June 22, 2022, 2:57am UTC](https://forums.developer.nvidia.com/t/tegrastats-monitoring/217088 "2022-06-22T02:57:54Z")

</div>

Hello, I would like to get more explaination about the log in tegrastats. When I run $ sudo tegrastats, I have the following log: In this image you can see the usage of RAM, etc… BUT what I need to see, is the power…
