24 minutes 36 seconds on one Thor
NVIDIA’s developer blog says TensorRT Edge-LLM completed the MLPerf Inference v6.1 Edge Agentic benchmark 6.4 times faster on Jetson AGX Thor. AI Daily Post, edited by Brian Petersen and dated 16 September, puts the TensorRT run at 24 minutes 36 seconds on a single Jetson AGX Thor Developer Kit, against 2 hours 37 minutes for the llama.cpp reference on the same hardware and the same Qwen3.6-27B model. It calls that a 6.4 times gap. The 24-minute and 2-hour-37-minute prints are in the AI Daily Post copy; NVIDIA’s pool title names the 6.4 times Thor result without repeating those clock times.[1], [2]
