Skip to content

fix(vulkan): refresh ggml for AMD submission handling - #3482

Merged
LauraGPT merged 1 commit into
mainfrom
codex/fix-amd-vulkan-3479
Aug 11, 2026
Merged

fix(vulkan): refresh ggml for AMD submission handling#3482
LauraGPT merged 1 commit into
mainfrom
codex/fix-amd-vulkan-3479

Conversation

@LauraGPT

Copy link
Copy Markdown
Collaborator

Summary

  • refresh the pinned llama.cpp revision to 803b7fca, including AMD submission-size heuristics, the batching-threshold fix, and DeviceLost diagnostics
  • adapt Fun-ASR-Nano to the updated penalties sampler API
  • document Windows AMD Vulkan diagnostics and the reliable CPU fallback

Refs #3479.

Verification

  • clean CPU configure and build: llama-funasr-cli, llama-funasr-paraformer, llama-funasr-sensevoice
  • clean Linux Vulkan configure and build: llama-funasr-sensevoice
  • confirmed linkage to libvulkan.so.1
  • verified the pin contains llama.cpp commits 1051ecd2, 51eae8cf, 0bbc87b1, and 803b7fca
  • git diff --check

The Windows AMD runtime result still needs reporter hardware validation; this PR intentionally does not claim the driver crash is fixed.

Signed-off-by: LauraGPT <18321252+LauraGPT@users.noreply.github.com>
@LauraGPT
LauraGPT merged commit 500956b into main Aug 11, 2026
10 checks passed
@LauraGPT
LauraGPT deleted the codex/fix-amd-vulkan-3479 branch August 11, 2026 04:19
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant