Profile Job Results

Jobs
j7gj3rnxp
Results Ready
Name
whisper_tiny_en_WhisperDecoder
Target Device
  • Snapdragon X Elite CRD
  • Windows 11
  • Snapdragon® X Elite | SC8380XP
Creator
ai-hub-support@qti.qualcomm.com
Target Model
Input Specs
x: int32[1, 1]
index: int32[1, 1]
k_cache_cross: float32[4, 6, 64, 1500]
v_cache_cross: float32[4, 6, 1500, 64]
k_cache_self: float32[4, 6, 64, 224]
v_cache_self: float32[4, 6, 224, 64]
Completion Time
8/11/2024, 5:38:20 AM
Versions
  • QNN: v2.24.0.240626131148_96320
  • QNN Backend API: 5.24.0
  • QNN Core API: 2.17.0
  • Windows: Windows 11 (26100)
  • AI Hub: aihub-2024.08.01.0
Estimated Inference Time
2.06 ms
Estimated Peak Memory Usage
10 MB
Compute Units
NPU
447
StageTimeMemory
First App Load
1.16 s10 MB
Subsequent App Load
1.11 s10 MB
Inference
2.06 ms10 MB
QNNValue
context_options.htp_options.performance_modeBURST
default_graph_options.htp_options.precisionFLOAT16

Sign up to run this model on a hosted Qualcomm® device!

Run on device