upvote
What quantization are you using? Which infra provider?

Baseten.co's version got into a loop rather rapidly... I've since added loop detection and adjusted some other settings on the pi coding agent and have yet to notice it again. I also switched to DeepInfra ... who serves an fp4 version admittedly, but I've had no issues with it as of yet and it's the top provider on openrouter.ai volume wise.

reply
I think this one requires a bit of strong prompting. I am also normally a Pi user, but my experience in OpenCode with this model has been drastically better than in Pi, where it overthinks a lot and gets distracted by random things.

It might be even better in Codex or Oh My Pi according to this bench I saw earlier: https://nitter.net/composio/status/2085330847951970801

reply
using the platform.deepseek API version I haven't had this happen once in my usage (which has been exclusive since its release). Which provider are you using? I also use a pretty bare bones Pi.
reply
I've been using for work, from opencode $10/mo subscription plan, on high effort (which is better than max imo), and haven't had any issue.

When it was first available in opencode, it was kinda slow for me, I guess because everyone wanted to try the new shiny. But now it's back to being screamingly fast and Opus 4.8 level of smart, for penies.

reply
Yeah, I saw the same thing - quite annoying. It can be mitigated through the prompt.
reply