opencl: fold the gpt-oss MoE per-expert bias adds into the epilogue (op/kernel fusion) (#26431)
* opencl: fold the gpt-oss MoE bias adds into swiglu_oai Default on, opt out with GGML_OPENCL_FUSE_MOE_BIAS_GLU=0. * opencl: fold the MoE down-projection bias into the combine Default on, opt out with GGML_OPENCL_FUSE_MOE_BIAS_COMBINE=0.
H
Hongqiang Wang committed
3af988fabcf79fd81f8720505e684d2aa5bfc786
Parent: 9a286ac
Committed by GitHub <noreply@github.com>
on 8/21/2026, 9:24:33 PM