Skip to content

[Bug] recent change of param "block_size_x/y, unroll" in dlight/gpu/matmul.py significantly decrease q4f16_1 prefill speed on android 8gen3 device #39

[Bug] recent change of param "block_size_x/y, unroll" in dlight/gpu/matmul.py significantly decrease q4f16_1 prefill speed on android 8gen3 device

[Bug] recent change of param "block_size_x/y, unroll" in dlight/gpu/matmul.py significantly decrease q4f16_1 prefill speed on android 8gen3 device #39

Triggered via issue October 30, 2024 06:31
Status Skipped
Total duration 1h 39m 39s
Artifacts

github-command-test.yml

on: issue_comment
run_command
0s
run_command
Fit to window
Zoom out
Zoom in