Measured sm_110 / sm_110a facts for NVIDIA Jetson AGX Thor: tensor-core capability matrix (2:4 sparsity works on plain sm_110; tcgen05 + NVFP4 block-scale is sm_110a-only), CUDA 11/12->13 migration breaks, CPU ISA, and the bandwidth roofline. With probes.
gpu cuda jetpack nvidia cutlass aarch64 jetson ptx structured-sparsity tensor-cores blackwell llm-inference nvfp4 jetson-thor cuda-13 agx-thor tcgen05 sm110 ptxas sm110a
-
Updated
Aug 16, 2026 - Shell