Qwen2.5 inference runtime for Jetson Orin with BF16 Tensor Core and W8A16 CUDA kernel optimization.
-
Updated
Aug 15, 2026 - C++
8000
Qwen2.5 inference runtime for Jetson Orin with BF16 Tensor Core and W8A16 CUDA kernel optimization.
Add a description, image, and links to the bf16 topic page so that developers can more easily learn about it.
To associate your repository with the bf16 topic, visit your repo's landing page and select "manage topics."