Agentic models brush up against reality, this gives a way around the hallucination problem.
Here is a recent talk showing that hallucination and discovery are actually positively coupled. https://www.youtube.com/live/ZNlZsI9kBm4?si=nhn4ancXu7s6qtom...
That border is determined not as much by ground truth, as by people's sentiment. As such, many famous artists and creatives never succeed navigating - their contemporaries would call them crazy, and then decades or centuries later, the border shifts, and suddenly they're remembered as greatest creators of their era.
Until an AI can have desires or experience life it cannot tell a story.
Even if it could it is questionable if such alien experience qualifies as story.
It can retell stories. It can swirl around extant media like a child playing with its dinner. It cannot blend in its own lived and new perspectives which is what gives new art value.
It's interesting to compare your opinion with Andrew Wiles quote:
> Perhaps I could best describe my experience of doing mathematics in terms of entering a dark mansion. You go into the first room and it's dark, completely dark. You stumble around, bumping into the furniture. Gradually, you learn where each piece of furniture is. And finally, after six months or so, you find the light switch and turn it on. Suddenly, it's all illuminated and you can see exactly where you were. Then you enter the next dark room...
There; now the statement is far less vague and handwavy, because I'm making a specific claim that is, AFAIK, well-understood.
Also, you seem to have misunderstood me a little. I didn't mean "prove me wrong", I meant "If I'm missing something, please tell me about it." More of a conversational request than staking a claim in an argument. Many people at HN seem to like to take argumentative, debate-competition stances — but I usually prefer more "Hey, let's discuss this interesting idea, point out mistakes each other is making, and learn together" kind of interactions. That's what I was asking for.
Well, I guess the answer isn't too different. I believe model collapse is a limitation of the current AI tech but maybe not the future ones. You can see humanity as a huge model that trains itself. What is novel about AI is that we built a machine with some intelligence traits that is free of biological constraints. If we can emulate the aggregate intelligence of a civilization inside a machine, it could improve itself forever but at a much faster pace.
If your training run dies at 1 am and you’re sleeping, you won’t find out about it until the next day. You can lose up to 18 hours of work depending on when it happens. Based on the error it might be as simple as tweaking a single hyperparameter and rebooting, which is something LLMs are usually capable of.
Even just that task means I can kick off multiple runs over the weekend and have confidence they’ll finish. It’s a game changer.
But I'd classify this as LLM being used to automate a sysadmin task, rather than calling that self-training.
In a way, “recursive self improvement” just means tools helping us to create better tools. At least that’s what the words mean.
Now humans are just as bad, but they're not moving at the speed of compute so the posion and fabricates can dissolve over time, or just, as you've noticed turning on your news, get stuck in very stupid positions. So humans are clearly capable but clearly don't tend to do this either.
So then we dont have a real road map. The error rates, although small, acrete at exponential levels and will wash out improvements.
So I also had the idea that "if we just give it enough context, surely it'll be more powerful". But the error rates hit that squarely. The larger the context grows, the more likely it hasn't properly organized its knowledge to avoid overlapping facts.
In programming, it's worse, because a lot of the code is purposefully "DRY" and reuseable. Everye C program has a main(); is it remembering the correct main? or any of the number of same variables?
You can see an LLM is powerful but it's not ominipotent. It'll suffer very much when it starts hallucinations and context poisoning.
So, sure you can try a super ralph wiggum loop with memory, fallback safeties, etc, but you basically then need another turtle that does the same thing, and at that point, you're positing a infinite jest of ralph wiggum loops tracking each other, recursively, forever.
https://www.rameznaam.com/p/471bbae4-1163-4048-944b-18f8b0bf...