Hacker News
new
past
comments
ask
show
jobs
points
by
pplonski86
5 hours ago
|
comments
by
chii
5 hours ago
|
[-]
> I had older NVIDIA card (RTX 3070) and llama.cpp instalation was smooth
what model was it that you were able to run with the rtx 3070?
reply
by
pplonski86
3 hours ago
|
parent
|
[-]
I was able to fit only small models Qwen3.5-4B in RTX3070 which is not very useful for Python and SQL generation thought. When I wan to test larger open LLM models I often just use cloud resources.
reply