zlacker

[return to "Qwen3-Coder-Next"]
1. zokier+lM[view] [source] 2026-02-03 19:09:59
>>daniel+(OP)
For someone who is very out of the loop with these AI models, can someone explain what I can actually run on my 3080ti (12G)? Is this something like that or is this still too big; is there anything remotely useful runnable with my GPU? I have 64G RAM if that helps (?).
◧◩
2. cirrus+UX[view] [source] 2026-02-03 20:00:22
>>zokier+lM
This model is exactly what you’d want for your resources. GPU for prompt processing, ram for model weights and context length, and it being MoE makes it fairly zippy. Q4 is decent; Q5-6 is even better, assuming you can spare the resources. Going past q6 goes into heavily diminishing resources.
[go to top]