upvote
[flagged]
reply
Are you saying scripts from agents are deterministic? :)
reply
Why don't you try to dispove me. Yes, they are _more_ deterministic than tool calls and consume less tokens.
reply
There is nothing to disprove as you don't understand what does a word mean. Deterministic is not a spectrum, they can either be deterministic or not. In both cases, they are not.
reply
Oh, I'm so sorry I touched your paper feelings.

How dared I to imply that some LLM output is more deterministic than the other, your LLM majesty. Shame on me and my entire family! For generations to come!

So sorry I implied that the code that doesn't work and has to be fixed later is deterministic in its execution and can be reused later instead of being re-generated from scratch!

Will I ever wash it off my name, your grace?

reply
They're talking about writing a file with a harness-native Edit tool. They're saying the agents aren't doing that, but are using ad-hoc methods of writing the files. (My agents seem to prefer see these days.)
reply
Why do you think your agents prefer to create scripts instead of doing tool calls these days?

I wonder, is it easier to modify a script that agent wrote before to satisfy your prompt, or is it easier to write a new one from scratch each time a retry happens?

Are input tokens more expensive than output tokens?

reply
My god. They are not reusing the scripts. They are adhoc, inline Python scripts just used to make a single edit. You seem to fundamentally not understand what everyone else is talking about
reply
deleted
reply