BeefAndPoultry@lemmus.org to LocalLLaMA@sh.itjust.worksEnglish · 2 days agoIntroducing Qwen3.8-27B Dynamic v3 Unsloth GGUFshuggingface.coexternal-linkmessage-square10fedilinkarrow-up161file-textcross-posted to: hackernews@lemmy.bestiver.se
arrow-up161external-linkIntroducing Qwen3.8-27B Dynamic v3 Unsloth GGUFshuggingface.coBeefAndPoultry@lemmus.org to LocalLLaMA@sh.itjust.worksEnglish · 2 days agomessage-square10fedilinkfile-textcross-posted to: hackernews@lemmy.bestiver.se
minus-squareAvid Amoeba@lemmy.calinkfedilinkEnglisharrow-up2·2 days agoThere’s no point in this if one runs Q8 right? I guess could save VRAM by dropping down from Q8 to say UD Q6.
minus-squarehumanspiral@lemmy.calinkfedilinkEnglisharrow-up2·2 days agothey have full range of quants.
minus-squareblob42@lemmy.mllinkfedilinkEnglisharrow-up1·2 days agoBut is it worth considering UD Q8_K_L if one is running Q8 K XL ?
minus-squareBeefAndPoultry@lemmus.orgOPlinkfedilinkEnglisharrow-up2·2 days agoThey have a graph, differences are tiny at that high end
There’s no point in this if one runs Q8 right? I guess could save VRAM by dropping down from Q8 to say UD Q6.
they have full range of quants.
But is it worth considering UD Q8_K_L if one is running Q8 K XL ?
They have a graph, differences are tiny at that high end