upvote
I feel like for localAI t/s is less of an issue. Just make a PRD and run a ralph loop. For big slogging projects like reverse engineering, or converting a codebase to a new language it actually doesn't matter if it takes a day or seven days.
reply
Yeah this is my experience. My 24GB 3090 + 64GB RAM takes a couple hours to crank out some code with largest Gemma 4 and Qwen3.8 models it can run

But in the meantime I get dishes done, vacuum, flip laundry... etc etc

Frontier models also seem in such a rush to emit anything they produce a mess that needs steering all day anyway

While I have not tested it, it feels like my local setup going slower is better at producing code that works the first time as its not trying to look fast for marketing sake

reply
Sounds like it's worth waiting for M7 anyways, no point investing too much right now

https://news.ycombinator.com/item?id=48676795

reply
At least wait until someone gets one in hand and post a review, and if it does work decently well there probably is going to be a long backlog.
reply