Files
ggml-org-llama-cpp-mirror/conversion
Hrishith Thadicherla a11f57ba93 model : fix DFlash output head sharing (#30111)
* llama : fix DFlash output head sharing

Assisted-by: Codex

* dflash : read tied output weights from GGUF metadata

Assisted-by: Codex

* llama : share tied word embedding metadata

Assisted-by: Codex

* llama : remove DFlash embedding head fallback

Assisted-by: Codex
2026-10-08 17:42:28 +02:00
..
2026-09-07 21:10:06 +02:00
2026-10-01 14:13:27 +03:00
2026-08-17 10:15:11 +02:00