I was saying the opposite. The core feature of an LLM is that it is subjective. The implications of a written expression (prompt) are not well-defined objective truth, but instead a vague probability. We can compute the probability, but that doesn't ever intersect with logical reduction or arithmetic; so the implications we get are just vague guesses on the progression of the story. With enough examples and training, the LLM can guess correct arithmetic answers, but critically, it does not actually perform the arithmetic that verifies them.