Hi,您好~
我们用的Orin NX的核心卡,定制的客户底板 (这个底板已经在Xavier NX上已经大量使用的),目前使用的R36.4.3版本(JP6.2),在量产阶段,概率性出现启动断言错误,目前发现两种错误:
1,第一种错误,也是概率性的,很难出现,但是5秒后可以重启
ÿTC: Reserved shared memory is disabledI/TC: Dynamic shared memory is enabledI/TC: Normal World virtualization support is disabledI/TC: Asynchronous notifications are disabledI/TC: WARNING: Test UEFI variable auth key is being used !I/TC: WARNING: UEFI variable protection is not fully enabled !
Camera-FW on t234-rce-safe startedTCU early console enabled.Camera-FW on t234-rce-safe ready SHA1=e2238c99 (crt 0.891 ms, total boot 55.003 ms)
ASSERT [NvmExpressDxe] /out/nvidia/bootloader/uefi/Jetson_RELEASE/edk2/MdeModulePkg/Bus/Pci/NvmExpressDxe/NvmExpressHci.c(772): (Private->Cap.Mpsmin + 12) <= 12
Resetting the system in 5 seconds.
ÿ怢Shutdown state requested 1
Rebooting system ...
2,第二种错误, 也是概率性的,但是这个问题最严重:
[2025/9/9 17:23:52][ 16.301555] ERROR: mounting PARTUUID=43f6f058-79bb-4539-9c3d-7d833cd0b311 as /mnt fail...
[2025/9/9 17:23:52][ 16.303047] ERROR: PARTUUID=43f6f058-79bb-4539-9c3d-7d833cd0b311 mount fail...
[2025/9/9 17:23:52][ 16.304789] ttyTCU0: Press [ENTER] to start bash in 30 seconds...
[2025/9/9 17:23:55][ 19.305303] ttyTCU0: Press [ENTER] to start bash in 27 seconds...
[2025/9/9 17:23:58][ 22.305818] ttyTCU0: Press [ENTER] to start bash in 24 seconds...
[2025/9/9 17:24:01][ 25.306394] ttyTCU0: Press [ENTER] to start bash in 21 seconds...
[2025/9/9 17:24:04][ 28.306969] ttyTCU0: Press [ENTER] to start bash in 18 seconds...
[2025/9/9 17:24:07][ 31.307544] ttyTCU0: Press [ENTER] to start bash in 15 seconds...
[2025/9/9 17:24:10][ 34.308119] ttyTCU0: Press [ENTER] to start bash in 12 seconds...
[2025/9/9 17:24:13][ 37.308695] ttyTCU0: Press [ENTER] to start bash in 9 seconds...
[2025/9/9 17:24:16][ 40.309202] ttyTCU0: Press [ENTER] to start bash in 6 seconds...
[2025/9/9 17:24:19][ 43.309776] ttyTCU0: Press [ENTER] to start bash in 3 seconds...
[2025/9/9 17:24:22][ 46.312045] No ttyAMA0 is found
[2025/9/9 17:24:22][ 46.312088] Rebooting system...
[2025/9/9 17:24:22][ 46.314021] sysrq: Resetting
[2025/9/9 17:24:22][ 46.314047] ------------[ cut here ]------------
[2025/9/9 17:24:22][ 46.314048] Voluntary context switch within RCU read-side critical section!
[2025/9/9 17:24:22][ 46.314055] WARNING: CPU: 4 PID: 481 at kernel/rcu/tree_plugin.h:316 rcu_note_context_switch+0x37c/0x460
[2025/9/9 17:24:22][ 46.314069] Modules linked in: pwm_fan pwm_tegra tegra_bpmp_thermal tegra_xudc ucsi_ccg typec_ucsi typec nvme nvme_core phy_tegra194_p2u pcie_tegra194
[2025/9/9 17:24:22][ 46.314085] CPU: 4 PID: 481 Comm: reboot Not tainted 5.15.148-rt-tegra #24
[2025/9/9 17:24:22][ 46.314087] Hardware name: NVIDIA NVIDIA Jetson Orin NX Engineering Reference Developer Kit/Jetson, BIOS 36.4.3-gcid-38968081 01/08/2025
[2025/9/9 17:24:22][ 46.314089] pstate: 604000c9 (nZCv daIF +PAN -UAO -TCO -DIT -SSBS BTYPE=--)
[2025/9/9 17:24:22][ 46.314092] pc : rcu_note_context_switch+0x37c/0x460
[2025/9/9 17:24:22][ 46.314094] lr : rcu_note_context_switch+0x37c/0x460
[2025/9/9 17:24:22][ 46.314095] sp : ffff80000b36ba60
[2025/9/9 17:24:22][ 46.314096] x29: ffff80000b36ba60 x28: 0000000000000001 x27: 0000000000000001
[2025/9/9 17:24:22][ 46.314098] x26: ffff000087148f80 x25: 0ce6b7365f8d0db0 x24: ffffb736603bb008
[2025/9/9 17:24:22][ 46.314100] x23: 0000000000000000 x22: ffff000087148f80 x21: ffffb73660d1aca8
[2025/9/9 17:24:22][ 46.314102] x20: 0000000000000000 x19: ffff0003f02829c0 x18: 0000000000000000
[2025/9/9 17:24:22][ 46.314104] x17: 0000000000000000 x16: 0000000000000000 x15: 0000000000000000
[2025/9/9 17:24:22][ 46.314106] x14: 0000000000000000 x13: 216e6f6974636573 x12: 206c616369746972
[2025/9/9 17:24:22][ 46.314108] x11: 6320656469732d64 x10: 6165722055435220 x9 : 206e696874697720
[2025/9/9 17:24:22][ 46.314110] x8 : 6863746977732074 x7 : 7865746e6f632079 x6 : 7261746e756c6f56
[2025/9/9 17:24:22][ 46.314112] x5 : ffff0003f0271bc8 x4 : 00000000fffff23d x3 : ffffb736609d2a28
[2025/9/9 17:24:22][ 46.314114] x2 : 0000000000000000 x1 : 0000000000000000 x0 : ffff000087148f80
[2025/9/9 17:24:22][ 46.314116] Call trace:
[2025/9/9 17:24:22][ 46.314118] rcu_note_context_switch+0x37c/0x460
[2025/9/9 17:24:22][ 46.314119] __schedule+0xc8/0x810
[2025/9/9 17:24:22][ 46.314125] schedule+0x90/0x120
[2025/9/9 17:24:22][ 46.314128] schedule_timeout+0xa4/0x1e0
[2025/9/9 17:24:22][ 46.314130] msleep+0x40/0x60
[2025/9/9 17:24:22][ 46.314134] pr_flush+0x1a0/0x1e0
[2025/9/9 17:24:22][ 46.314137] kmsg_dump+0xdc/0x190
[2025/9/9 17:24:22][ 46.314140] emergency_restart+0x20/0x40
[2025/9/9 17:24:22][ 46.314144] sysrq_handle_reboot+0x24/0x30
[2025/9/9 17:24:22][ 46.314150] __handle_sysrq+0x98/0x1b0
[2025/9/9 17:24:22][ 46.314151] write_sysrq_trigger+0x12c/0x190
[2025/9/9 17:24:22][ 46.314153] proc_reg_write+0xd0/0x140
[2025/9/9 17:24:22][ 46.314157] vfs_write+0xf8/0x2e0
3,这里有一个参考,错误虽然一样,但是R36.4.3没有办法关闭PCI,我们需要使用PCIE功能的
另外这个问题也不一样,看介绍这个问题你们在R36.4.3里面已经解决了