master
ggml: backend-agnostic tensor parallelism (experimental) (#19378)
CUDA: manage NCCL communicators in context (#21891)
vulkan: Support F16 OP_FILL (#22177)
vulkan : cmake integration (#8119)
ggml : bump version to 0.10.0 (ggml/1463)