upvote
Pretty crazy that the model doesn't know that it needs to stop before it hits 128k output tokens. I guess it has no sense of how many tokens in it is? Wouldn't this be possible to work into the architecture?
reply
I think this is a bug. I've not seen this problem from any of the other frontier models.
reply
I would also consider this a bug. I think ajy reasonable consumer would.
reply
Do other models put a hard cap on the output tokens it can generate?
reply