BeefAndPoultry@lemmus.org to LocalLLaMA@sh.itjust.worksEnglish · 1 month agoIntroducing Qwen3.8-27B Dynamic v3 Unsloth GGUFshuggingface.coexternal-linkmessage-square9linkfedilinkarrow-up163arrow-down13file-textcross-posted to: hackernews@lemmy.bestiver.se
arrow-up160arrow-down1external-linkIntroducing Qwen3.8-27B Dynamic v3 Unsloth GGUFshuggingface.coBeefAndPoultry@lemmus.org to LocalLLaMA@sh.itjust.worksEnglish · 1 month agomessage-square9linkfedilinkfile-textcross-posted to: hackernews@lemmy.bestiver.se
minus-squareAvid Amoeba@lemmy.calinkfedilinkEnglisharrow-up3·1 month agoThere’s no point in this if one runs Q8 right? I guess could save VRAM by dropping down from Q8 to say UD Q6.
minus-squarehumanspiral@lemmy.calinkfedilinkEnglisharrow-up3·1 month agothey have full range of quants.
minus-squareblob42@lemmy.mllinkfedilinkEnglisharrow-up2·1 month agoBut is it worth considering UD Q8_K_L if one is running Q8 K XL ?
minus-squareBeefAndPoultry@lemmus.orgOPlinkfedilinkEnglisharrow-up4·1 month agoThey have a graph, differences are tiny at that high end
There’s no point in this if one runs Q8 right? I guess could save VRAM by dropping down from Q8 to say UD Q6.
they have full range of quants.
But is it worth considering UD Q8_K_L if one is running Q8 K XL ?
They have a graph, differences are tiny at that high end