upvote
Oh yeah. I discovered this back in February. Faster and cheaper is better, because faster means you can focus.

Frenetic multitasking is for suckers.

reply
I’ve also found that fast cheap models are more sustainable because they make it easier to keep technical debt under control. Partially because I’m keeping sustained attention in one place. And partially because small fast models with low reasoning don’t seem to love technical debt either and will (unintentionally) start to give you negative feedback signals when you’re letting things get messy.
reply
Do you use luna to plan as well, or is it just code? I am interesting in understanding the process you follow.

My workflow is the most common one - use something like Opus to help create a spec document (I decide the specs and then get the LLM to ask me questions to harden it), then generate a plan, and then implement. I use Opus the whole time. Sonnet sometimes if it is a simple task.

reply
I've had bad experiences with gpt 6 luna, but 5.6 luna worked real nicely with e.g React, you can do a back n forth and it'll work like a teammate vs just a "done!" where you then have to clean everything up afterwards
reply
Yeah, I've come to this realization as well. It's like the inverse of that classic interruption comic. The waiting makes me lose the context.
reply
That's a great idea, I will try that too. The LLMs are so capable, but the slowness annoys me. I always end up working on a few things at once and always doing rounds of reading its output, giving new orders, and going back to Reddit/HN/YouTube while it's doing its thing for the next 5 minutes.
reply