Seems like smarter and more efficient quants than normal
Is a 35b-a3b planned for 3.8?
They don’t tell us things lol. All they said was

Which sounds like there’s something “better” coming, but doesn’t deny the possibility of 35b-a3b. Which is weird because “better” is subjective and depends on your hardware. It could be smaller and smarter than 3.6 35b, but then people are gonna ask for a 3.8 35b because it should be even smarter.
Yeah I was gonna say, 64b or whatever would be completely useless for me. There’s a reason for that size.
But I guess this at least means they are still looking at that range.
There’s no point in this if one runs Q8 right? I guess could save VRAM by dropping down from Q8 to say UD Q6.
they have full range of quants.
But is it worth considering UD Q8_K_L if one is running Q8 K XL ?
They have a graph, differences are tiny at that high end
Your mom’s a more efficient quant.





