Date: 2026-05-26 18:18Backend: llama.cpp CUDA (-ngl 99) Concurrency: 1Sweep: prompt in {128,512,1024,2048} tok × gen in {64,128,256} tok
Artifacts:
https://huggingface.co/datasets/YuvrajSingh9886/jetson-non-reasoning-benchmark-maxn
Cells marked — = OOM (server crashed or skipped). Power = VDD_CPU_GPU_CV average over… See the full description on the dataset page:
https://huggingface.co/datasets/YuvrajSingh9886/jetson-non-reasoning-benchmark-maxn.