Hacker News
new
past
comments
ask
show
jobs
points
by
xienze
2 hours ago
|
comments
by
Auracle
2 hours ago
|
[-]
Sure, but shouldn’t the programs to run the LLMs go “the user has this much vram and the model is this size, so I’ll start with sensible defaults based on that”?
You could override, obviously.
reply
by
zargon
2 hours ago
|
parent
|
[-]
Yes, llama.cpp does that.
reply