Hacker News
new
past
comments
ask
show
jobs
points
by
psychoslave
1 days ago
|
comments
by
nsbk
1 days ago
|
[-]
I run a 2x 3090 rig, but a single 3090 already provides a great experience at a reasonable quant and context size. On a single card I used to run Qwen_Qwen3.6-27B-Q4_K_M or similarly quantized 35B MoE at 65536 context size
reply
by
xiconfjs
22 hours ago
|
parent
|
[-]
you should look into
https://github.com/noonghunna/club-3090
reply